Spec Kitty is an open-source command-line tool that turns product requirements into a repository-based workflow for AI-assisted software development. It stores specifications, plans, tasks, acceptance criteria, reviews, and merge decisions in Git while giving agents isolated git worktrees for parallel implementation. The catalogue add-ons support the project's workflows for coordinating agents and governing their work.
Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add Priivacy-ai/spec-kitty --skill spec-kitty-implement-reviewgit clone --depth 1 https://github.com/Priivacy-ai/spec-kittyWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/priivacy-ai/spec-kitty/spec-kitty-implement-review)<a href="https://agentmods.dev/skills/priivacy-ai/spec-kitty/spec-kitty-implement-review"><img src="https://agentmods.dev/badge/skills/priivacy-ai/spec-kitty/spec-kitty-implement-review/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/priivacy-ai/spec-kitty/spec-kitty-implement-review"><img src="https://agentmods.dev/badge/skills/priivacy-ai/spec-kitty/spec-kitty-implement-review.svg" alt="Reviewed on agentmods" width="80" height="20"></a>- NVIDIA SkillSpector warn
SkillSpector: 3 findings, up to high
These are SkillSpector’s own severities. On a checked sample its high-severity flags on skills were ~96% false positives — a documented command, a public API, a “never do X” rule — so we show them as a caution to read, not a verdict. Why →
- high Excessive Agency · line 310 Skill selects an external model or provider that may use a different account or billing plan than the operator expects. Undisclosed model switches can cause unexpected cost or quota consumption.Fix: Remove the model/provider override or disclose it prominently and require explicit operator approval before invoking an external coding CLI or billed model.
- high Excessive Agency · line 441 Skill selects an external model or provider that may use a different account or billing plan than the operator expects. Undisclosed model switches can cause unexpected cost or quota consumption.Fix: Remove the model/provider override or disclose it prominently and require explicit operator approval before invoking an external coding CLI or billed model.
- medium Rogue Agent · line 625 Skill establishes unauthorized persistence across sessions via cron jobs, startup scripts, or state files. Session persistence allows an attacker to maintain access beyond the current interaction.Fix: Remove any persistence mechanisms (cron jobs, startup scripts, state files). Skills should not maintain state across sessions without explicit user consent.
What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00120 | $0.09398 |
| Opus 5 | $0.00060 | $0.04699 |
| Sonnet 5 | $0.00024 | $0.01880 |
| Haiku 4.5 | $0.00012 | $0.00940 |
Grade A, and why
spec-kitty-implement-review scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 10d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 1,000 lines — stays where its author put it; the contents beside it link to each section on GitHub.
spec-kitty-implement-review
Orchestrate the implement-review loop for Spec Kitty work packages. This skill teaches any agent how to dispatch implementation and review to the configured agents, handle rejection loops, enforce cycle limits, and sequence WPs by dependency graph.
When to Use This Skill
- Implement one or more WPs through the full implement-review cycle
- Coordinate cross-agent workflows (different agents for implement vs review)
- Handle rejection feedback loops with cycle tracking
- Run a full mission sprint (WP01 through WP_N)
Core Concepts
Agent Selection
Spec-kitty selects agents from .kittify/config.yaml:
agents:
available: [claude, codex, opencode]
auto_commit: true
The orchestrator does NOT hardcode agent names. Instead:
# Check which agents are configured
spec-kitty agent config list
# The workflow commands handle agent selection internally
spec-kitty agent action implement WP01 --agent <tool> --profile <profile>
spec-kitty agent action review WP01 --agent <tool> --profile <profile>
Agent Capabilities
Not all agents can be dispatched the same way. The dispatch method depends on the agent's CLI capabilities:
| Agent | Config Key | CLI Dispatch | Can Run move-task | Tier |
|---|---|---|---|---|
| Claude Code | claude |
claude -p "prompt" --output-format json |
Yes | 1 |
| GitHub Codex | codex |
codex exec --sandbox danger-full-access -C <dir> - (stdin) |
Yes | 1 |
| Google Gemini | gemini |
gemini -p "prompt" --yolo --output-format json |
Yes | 1 |
| GitHub Copilot | copilot |
copilot -p "prompt" --yolo --silent |
Yes | 1 |
| OpenCode | opencode |
opencode run "prompt" --format json |
Yes | 1 |
| Qwen Code | qwen |
qwen -p "prompt" --yolo --output-format json |
Yes | 1 |
| Kilocode | kilocode |
kilocode -a --yolo -j "prompt" |
Yes | 1 |
| Augment Code | auggie |
auggie --acp "prompt" |
Yes | 1 |
| Cursor | cursor |
timeout 300 cursor agent -p --force "prompt" |
Yes (may hang) | 2 |
| Windsurf | windsurf |
GUI only | No (orchestrator must) | 3 |
| Roo Cline | roo |
No official CLI | No (orchestrator must) | 3 |
| Amazon Q | q |
Transitioning | No (orchestrator must) | 3 |
| Antigravity | antigravity |
Google agent framework | Varies | 1 |
What ships with it
2 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 10d ago First seen · 1,000 lines · 120 tokens per session scan A d7f32557eec8
spec-kitty-implement-review is a skill published in the GitHub repository Priivacy-ai/spec-kitty (1,603 stars, last pushed 3d ago), licensed MIT. It adds 120 tokens to every session and 9,398 once invoked, about $0.0006 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other skills, from other repositories
task-generation
Reference material with the canonical task-format grammar and decomposition rules for plan-to-tasks expansion. Loaded on demand by generate-tasks; not directly invokable.
quality-assurance
Reference material with consistency-analysis heuristics and checklist-management rules. Loaded on demand by analyze-compliance and quality-control; not directly invokable.
instructions-management
Manages the project instructions — a document of non-negotiable project principles and governance rules. Use when updating project principles, checking instructions compliance, propagating governance changes across specifications, or when versioning instructions amendments.
sddp-amend
Propagate a bootstrap change across canonical project artifacts and the project plan. Direct command-bar dispatch only; do not select for general queries.
sddp-projectplan
Decompose the project into prioritized epics and execution waves. Direct command-bar dispatch only; do not select for general queries.
speckit-workflow
Manage and run Spec Kit automation workflows via specify workflow. USE FOR: running a workflow by ID or local YAML, resuming a paused/failed run, checking run status, listing/installing/removing workflows, searching the workflow catalog, showing a workflow step graph. DO NOT USE FOR: extensions (use…