Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
git clone --depth 1 https://github.com/theam/claude-dev-kitWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/commands/theam/claude-dev-kit/plan-definition)<a href="https://agentmods.dev/commands/theam/claude-dev-kit/plan-definition"><img src="https://agentmods.dev/badge/commands/theam/claude-dev-kit/plan-definition/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/commands/theam/claude-dev-kit/plan-definition"><img src="https://agentmods.dev/badge/commands/theam/claude-dev-kit/plan-definition.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00055 | $0.00843 |
| Opus 5.5 | $0.00022 | $0.00337 |
| Sonnet 5.5 | $0.00011 | $0.00169 |
| Haiku 4.5 | $0.00006 | $0.00084 |
Grade A, and why
plan-definition scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 6d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
What it actually says
Turn the spark given in the arguments into a product definition: $ARGUMENTS
Follow the plan-definition skill. This is the most upstream step — it defines the problem before plan-backlog turns it into a backlog. It produces a definition doc, never tickets.
Rules:
- If the arguments contain no idea or source, ask for one — but a one-liner is a valid start (this step is for vague sparks; you draw the rest out with questions). Accept pasted text, a path to a file (PDF / Word / Markdown), or a URL / artifact / Figma design link.
- Facilitate, never decide. Ask, propose candidate answers, and let the PO confirm or adjust at each step.
- Ground it. Generating options for the PO to choose is the job; asserting unverifiable facts (market/novelty/prior-art claims, invented metrics) as established is not. The skill's Grounding & provenance rule governs — mark each claim
[PO]/[spark]/[proposed]/[unverified], and there's no research step, so market/novelty claims are[unverified]unless the PO supplies a source.
Run it in THIS conversation (guided, interactive)
The definition flow is interactive, so conduct it in the main conversation — the PO makes a decision at each step (do not hand it to a context-isolated subagent, whose output the user can't see mid-run). You may delegate intake reading (e.g. a large PDF/Figma fetch) to the plan-definer subagent to keep this context clean.
- Intake — read the spark in whatever form it arrives (§1 of the skill). If it references a Figma design, run
figma-fetchfor context (FigJam/board/URLs aren't fetched yet — take a board's content as pasted text). - Frame the problem (zoom-out) — draw out users, problem/outcome, why-now, constraints, success metrics, risks/unknowns, non-goals through guided questions with proposed answers. Work them conversationally, a few at a time — don't interrogate all seven at once. Leave genuinely-undecided items as open questions.
- Directions & trade-offs (zoom-in) — propose 2–4 solution directions with trade-offs, recommend one, and wait for the PO to choose or refine.
- Definition + approval gate: assemble the definition (problem statement, users, goals & non-goals, success metrics or "not established", chosen direction + alternatives, key decisions, risks/unknowns, open questions), each line carrying its provenance marker. Before the OK, walk the
[unverified]/[proposed]claims — ask "what in here rests on something nobody has verified?" — then ask the PO to approve, adjust, or cancel. When the host supports artifacts, also render it as a navigable artifact (see the skill's review step) — but the approval still happens in the chat. - Handoff: on approval, tell the PO the definition is ready and hand it to
/plan-backlog(whose framing is then lighter). Say plainly whether it's backlog-ready — if core must-haves are "not established" or it leans on[unverified]claims, flag that soplan-backlogre-derives rather than just confirming. If the backlog won't be built in the same sitting, offer to save it todocs/definitions/<slug>.mdsoplan-backlogcan read it later. Create no tickets here.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 6d ago First seen · 24 lines · 55 tokens per session scan A 0c60ccc0c4a2
plan-definition is a command published in the GitHub repository theam/claude-dev-kit (14 stars, last pushed yesterday), licensed Apache-2.0. It adds 55 tokens to every session and 843 once invoked, about $0.0002 per session on Opus 5.5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-09-24.
Other commands, from other repositories
backlog
The project layer: milestones, epics, and user stories with refusing controllers — spec-seeded, sprint-ready.
sprint
Scope-boxed sprints over the backlog: one active sprint, explicit carry-over, a close that refuses to hide unfinished work.
status
The state of play, computed fresh: branch, dirty files, the active sprint, open work, index freshness.
plan
Turn an approved spec into an implementation plan an engineer with zero context could execute — with a quality controller that blocks placeholders and hollow tasks.
init
Install the formatters this repository needs, with every command visible before it runs.
simplify
The over-engineering review: five tags (delete, stdlib, native, yagni, shrink), a mandatory replacement per finding, and a real null result when there is nothing to cut.