Agent Claude Code
Independent read-only judge for slices that change structure — new modules, cross-boundary dependencies, new interfaces, or added packages — returning a one-line verdict.
A public toolkit of reusable AI agents, skills, and other artifacts
Agent Claude Code
Independent read-only judge for slices that change structure — new modules, cross-boundary dependencies, new interfaces, or added packages — returning a one-line verdict.
Agent Claude Code
Implements one slice against its numbered spec clauses, and on a fix round reads the verify log or reviewer findings it was handed and clears what failed.
Agent Claude Code
Runs an approved spec slice by slice — delegates implementation, runs the verify command, dispatches independent review, bounds the fix rounds, and reports once.
Agent Claude Code
Independent read-only judge for one slice — renders a verdict per spec clause with receipts, writes the findings file, and returns a one-line verdict.
Agent Claude Code
Independent read-only judge for slices that touch the trust boundary — untrusted input, authentication, authorization, secrets, or dependencies — returning a one-line verdict.
Agent Claude Code
Independent read-only judge for slices touching {{SURFACE}} ({{GLOBS}}) — reviews for {{FOCUS}} and returns a one-line verdict.
Agent Claude Code
Analyze blind comparison results to understand WHY the winner won and generate improvement suggestions.
Agent Claude Code
Compare two outputs WITHOUT knowing which skill produced them.
Agent Claude Code
Evaluate expectations against an execution transcript and outputs.