Design-integrity guardrails for Codex and Claude Code: AGENTS.md, review skill, and deterministic hooks against speculative catches, fallbacks, and API detours.
Audit whether tests, evals, benchmarks, autoresearch tasks, generated artifacts, public APIs, or user-visible workflows provide independent behavioral evidence. Use at a completion checkpoint, not for ordinary implementation or documentation-only changes.
Route one frozen candidate change through a focused structural-integrity review, or to behavioral-acceptance-review when an explicit evaluation surface changed. Use at a completion checkpoint, not during ordinary edits; do not use for documentation-only, formatting-only, or routine test maintenance.
Establish a short, evidence-backed decision gate before architecture or broad implementation when missing facts, requirements, external behavior, or acceptance criteria could change the direction. Use only for consequential or materially uncertain work.
Preserve failure evidence and prove that tests detect the behavior they claim to cover. Use for bug fixes, regression coverage, test changes, flaky-test investigation, and behavior changes with a concrete acceptance contract; do not use for documentation-only or formatting-only work.
★not rated 0 4d agoA56 tokens
originalMIT
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: