debug

A routing guide for handling a specific, repeatable software failure, such as a broken test, bug, or failed deployment. It selects debugging methods and puts them in order.

In plain words
What is it for?
Use it to investigate reproducible failures, check hypotheses before changing code, and apply extra reasoning when timing or asynchronous behavior may be involved.
Why use it?
It prevents developers from jumping straight to a fix without first investigating the cause and testing their explanation.

Skill for Claude CodeCodex

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/langerrr/zforge/debug
Any agent
npx skills add Langerrr/zforge --skill debug
Clone the repo
git clone --depth 1 https://github.com/Langerrr/zforge

Made for: Claude Code, Codex.

Per session 20 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 536 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00020 $0.00536
Opus 5 $0.00010 $0.00268
Sonnet 5 $0.00004 $0.00107
Haiku 4.5 $0.00002 $0.00054

Measured yesterday against content hash 86544ea34bda, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

debug scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured yesterday.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

codex/skills/debug/SKILL.md · 42 lines

How it starts

The opening of the file, as written. The whole thing — 42 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Debugging Router

A concrete failure — a broken test, a reproducible bug, a failing deploy. This skill owns no debugging opinion of its own; it names which methodology applies and in what order, because the failure mode is applying one of them and skipping the rest.

Not for open-ended exploration ("this feels wrong"). That is $compound-engineering:ce-ideate territory.

Sequence

1. $superpowers:systematic-debugging — start here when installed. Investigation discipline comes before any fix, including when the cause seems obvious. When Superpowers is unavailable, use the host's equivalent systematic-debugging workflow; do not skip the investigation step.

2. $compound-engineering:ce-debug — add when installed. Its hypothesis-and-prediction gate goes before any code change: state what you believe is wrong and what you expect to see if you are right, then check. Without Compound Engineering, apply that same hypothesis-and-prediction gate directly. A fix applied without a failed prediction is a guess that happened to be typed confidently.

3. $zforge:async-reasoning — add when the failure involves timing. Async state, race conditions, cache invalidation, init order, stale reads, write-then-read gaps within one component.

4. $distributed-architect:dist-debug — add when that compatible plugin is installed and the failure crosses a boundary. Cross-service, cross-process, distributed races. Backward tracing from the symptom to the boundary where the cause lives. Check the project's topology file first if it has one.

While debugging a planned feature

If the failure is in a feature with a zforge directory, the ledger is evidence:

  • decision_review.md may already contain the decision that caused this.
  • An open standing flag may already name this defect class — a bug in the gap a flag describes is a flag being discharged the expensive way.
  • The phase's ## Evidence Required shows what was never verified, which is usually where to look first.

Read the full file on GitHub · 42 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. yesterday First seen · 42 lines · 20 tokens per session scan A 86544ea34bda

Subscribe to this mod's changes

debug is a skill published in the GitHub repository Langerrr/zforge (10 stars, last pushed yesterday), licensed MIT. It adds 20 tokens to every session and 536 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other skills, from other repositories

espalier-migrate

Migrate an existing harness/espalier install to the current Espalier version — auto-detects which of v0.1→v0.2, v0.3→v0.4, v0.4→v0.5, the v0.5.3 coder-agent patch, v0.5→v0.6 (Stage 1 grill), v0.6→v0.7 (read-only /espalier-ask lane), v0.7→v0.8 (requirements approval gate), the v0.8.1 impact-analysis agent patch, the…

Junhanliu-dev/espalier-engineering · 1,095 tokens

espalier-init

Analyze any existing codebase, discover its patterns and best practices, and generate an Espalier structure (rules, skills, wiki, pipeline) so AI coders produce code matching that project's language and conventions. Greenfield repos take the two-pass Decide-Then-Bind path (/espalier-map charts the conventions, init…

Junhanliu-dev/espalier-engineering · 118 tokens

math-unicode

Use whenever you need to express mathematical notation — equations, filters, set-builder, statistics, calculus, logic, ratios, drops, counts. Emit Unicode math glyphs INLINE; never wrap in $…$, \(...\), or $$...$$. Terminal coding agents (Claude Code, Codex CLI, and others) do not render LaTeX, so raw delimiters…

JustMichael-80/MasterVault · 104 tokens

sdd-serve

Serve the SDD Builder's AI request queue: claim requests with sddnextrequest, draft the proposal, answer with sddrespondrequest. Never writes spec files — the user accepts each proposal in the builder. Use when the user asks to attend, serve or listen to the SDD board queue. / Atiende la cola de peticiones del SDD…

juanklagos/spec-driven-development-template · 80 tokens

eo-brainstorming

对不成形的想法做发散、对抗、拆解和方向决策。触发:帮我想想 / 头脑风暴 / brainstorming / /eo-brainstorming。 NOT FOR: 普通技术讨论或具体实现问题(只在用户明确"想想方向"时触发)。.

SimpleEve/eo-skills · 75 tokens

sdd-workflow

Guide a project with Spec-Driven Development (SDD) discipline - idea, approved spec, consistent plan, tasks, a gate that verifies approval and consent, implementation, validation, and logbook. Bilingual EN/ES. Use when the user wants to start, spec, plan, implement, or validate work with SDD, or mentions specs, plans…

juanklagos/spec-driven-development-template · 84 tokens