adrianco/retort

Platform Evolution Engine. Distill the best from the combinatorial mess.

199Stars on the repository
11Mods indexed here, across every type
2d agoLast push, which is what freshness is scored on
Apache-2.0Licence, which decides whether bodies are shown

beads

01

adrianco/retort

Skill Claude CodeCodex

Use when working in a repository that uses bd or Beads for durable project task tracking, issue dependencies, blocker management, multi-session handoff, or shared work memory. Trigger when the user asks to find ready work, claim or close tasks, create follow-up work, inspect blockers, recover project context, or…

199 2d ago A 74 tokens copy · 100% Apache-2.0

compare-runs

02

adrianco/retort

Skill Claude CodeCodex

Compare evaluated runs in a retort experiment along factor dimensions. Surfaces effects of each factor, aggregates across replicates, and highlights cells that diverge qualitatively — complementing (not replacing) retort's ANOVA analysis.

199 2d ago A 50 tokens original Apache-2.0

diagnose-failed-run

03

adrianco/retort

Skill Claude CodeCodex

Determine the TRUE cause of a failed retort run before attributing it. Ground-truth every failure (run its tests, read its agent logs, inspect its workspace) and classify it as an infrastructure false-fail, a genuine model miss, or an environment issue — never trust the gate verdict or a log signature alone. Use…

199 2d ago A 109 tokens original Apache-2.0

evaluate-run

04

adrianco/retort

Skill Claude CodeCodex

Evaluate a single retort experiment run. Score the generated code against the task's TASK.md requirements, run its build and tests, compute metrics, and emit a structured evaluation report plus a machine-readable findings file.

199 2d ago A 45 tokens original Apache-2.0

file-run-issues

05

adrianco/retort

Skill Claude CodeCodex

Aggregate a retort run's findings.jsonl into a machine-readable assessment.json summary with severity counts, penalty score, requirement coverage, and top findings.

199 2d ago A 35 tokens original Apache-2.0

run-summary

06

adrianco/retort

Skill Claude CodeCodex

Summarize the architecture of code generated by a single retort run. Produces module-level structure, interfaces, and control flow in a form suitable for cross-run comparison — not a full codebase-summary.

199 2d ago A 45 tokens original Apache-2.0

update-optimal-blog

07

adrianco/retort

Skill Claude CodeCodex

Refresh the data tables in optimal-blog.md from master.db. Checks the data for integrity problems FIRST, then runs the generator that picks per-language winners and splices every GEN-marked table, then reconciles the surrounding prose. Use after new experiment results land, or when the optimal-blog numbers are stale.

199 2d ago A 66 tokens original Apache-2.0