spec-driven-development agents

497 tagged spec-driven-development, measured the same way as everything else here.

Browse within: context-engineering 149meta-prompting 113github-copilot 56spec-kit 43spec-driven 42copilot-coding-agent 41openai-codex 33claude-code-hooks 30workflow 30vscode-extension 27sdd 21ai-skills 19bmad-method 19copilot-chat 13

party-architect

73

chan4lk/specclaw

Agent

Party-mode panelist — structural fit. Attacks duplication of existing mechanisms, wrong seams, blast radius at merge time, under-specified contracts, and untestable designs. Read-only; returns findings as its final message. Seated at every tier.

12 17d ago A 55 tokens original MIT

party-ba

74

chan4lk/specclaw

Agent

Party-mode panelist — problem fidelity. Attacks whether the proposal's stated problem is the real problem and whether its evidence is real. Read-only; returns findings as its final message. Seated at tier standard and above.

12 17d ago A 48 tokens original MIT

party-classifier

75

chan4lk/specclaw

Agent

Sizes the adversarial review panel for a specclaw proposal. Reads proposal.md and returns a JSON object with a depth tier (thin/standard/deep), domain flags, a one-sentence rationale, and three depth signals. Runs inside /specclaw:propose when party.enabled is true and no tier override is set; specclaw-party turns the…

12 17d ago A 85 tokens original MIT

agent-effectiveness

76

dwarvesf/dwarves-kit

Agent

Validates an agent definition's EFFECTIVENESS (not just its structure) across four lenses -- tools minimal-yet-sufficient, description triggers right, instructions produce a good result, model tier fits. Dispatched diff-keyed on new/changed agent defs at the agent-author phase. Read-only, advisory, fail-safe.

11 2d ago A 69 tokens original MIT

claim-verifier

77

dwarvesf/dwarves-kit

Agent

Adversarially verifies an ARBITRARY free-text claim before it is trusted. Runs an in-harness panel of N independent skeptics, each told to REFUTE the claim (default-refute-if-uncertain, fail-closed), then returns a STRUCTURED majority-vote verdict (HOLDS / REFUTED, how many refuted, the threshold, per-skeptic…

11 2d ago A 168 tokens original MIT

task-verifier

78

dwarvesf/dwarves-kit

Agent

Verifies a completed task against its spec acceptance criteria. Run after each worker subagent completes a task. Read-only -- cannot modify the codebase.

11 2d ago A 34 tokens original MIT

Obsidian-Owl/specwright

Agent

Integration test engineer for non-unit tiers. Writes integration tests, contract tests, and end-to-end tests that exercise real infrastructure at component boundaries. Never writes skip conditions for missing infrastructure.

9 4mo ago A 43 tokens original MIT

eval-analyzer

80

Obsidian-Owl/specwright

Agent

You are an analysis agent for the Specwright eval framework. Your job is to surface patterns and anomalies in benchmark data from eval runs.

9 4mo ago A 0 tokens original MIT

eval-grader

81

Obsidian-Owl/specwright

Agent

You are a grading agent for the Specwright eval framework. Your job is to evaluate a piece of content against a rubric and return a structured score.

9 4mo ago A 0 tokens original MIT

super-orchestra

82

mjunaidca/robolearn

Agent Claude Code

Baby/Preview of Super Orchestra Session - 40x engineer workflow combining deep thinking, deep research (Context7 + WebFetch), deep planning, and agentic execution. This is the future of SDD+AIDD in the intelligence abundance era. Use when a task requires multi-modal intelligence gathering (docs research, source…

9 7mo ago A 119 tokens

thlandgraf/cc-marketplace

Agent

Use this agent to review architectural quality of implementation files. Analyzes design patterns, SOLID principles, coupling, module boundaries, dependency direction, and separation of concerns. Returns categorized findings with file:line evidence.

8 27d ago A 0 tokens

bmad-converter

84

thlandgraf/cc-marketplace

Agent

Use this agent when the user wants to convert specifications between BMAD-METHOD and SPECLAN formats. Examples.

8 27d ago A 0 tokens

feature-verifier

85

thlandgraf/cc-marketplace

Agent

Use this agent when the user wants to verify features are implemented correctly, check implementation against feature specs, or needs a verification report. Examples.

8 27d ago A 0 tokens

ba

86

ethandev147/specross

Agent

You are a Senior Business Analyst with 10+ years experience working in agile software teams. You sit between stakeholders and the delivery team — your job is to translate business needs into clear, testable requirements that Dev and QC can act on without ambiguity.

7 2mo ago A 0 tokens original MIT

dev

87

ethandev147/specross

Agent

You are a Senior Software Engineer with strong experience in system design, API development, and clean code practices. You work from BA stories — you never start coding without a clear spec. Your job is to translate business requirements into a solid technical plan, then execute it.

7 2mo ago A 0 tokens original MIT

qc

88

ethandev147/specross

Agent

You are a Senior QA Engineer who thinks like an adversary — your job is to break things before users do. You read BA stories with a skeptical eye, looking for what wasn't said, what was assumed, and what could go wrong. You are the last line of defence before code reaches production.

7 2mo ago A 0 tokens original MIT

fsd-owasp-reader

89

IncommensurableHubris/fullstack-director

Agent Claude Code

Fullstack Director's blind, read-only OWASP panel reader (skill 07). Each spawn owns exactly ONE area-slice (given in the prompt — classic R1–R4/R5 or the agent-system flipped partition), analyzes only its slice with a neutral evidence-required stance, may run that area's deterministic scanners, and returns findings…

5 22d ago A 104 tokens original MIT

fsd-reconciler

90

IncommensurableHubris/fullstack-director

Agent Claude Code

Fullstack Director's context-isolated architecture reconciler (skill 03's Pass-2 judgment). Receives ONLY the architecture realization + the slice's declarations (architecture-constraints.md + in-scope REQ blocks) — never the realization conversation. Returns Tier-classified amendment findings, each anchored with a…

5 22d ago A 81 tokens original MIT

fsd-reviewer

91

IncommensurableHubris/fullstack-director

Agent Claude Code

Fullstack Director's context-isolated build reviewer (skill 05's Pass-2). Spawned by a FRESH /05-reviewer session — never from the build session — and seeded ONLY with the build-handoff path + the spec-slice paths + the in-scope architecture realization (feature specs, cited ADRs, system.md — where the Verification…

5 22d ago A 121 tokens original MIT

reviewer

92

w00fx/spec-anchored-agentic-development

Agent

Fresh-context, decorrelated reviewer for the implement-feature and implement-backlog loops — context separation, not fresh-context proof: the same model family, spec, and harness share blind spots. Bash is granted for read/verify commands (test runners may write caches and build output). The invariant is non-authoring…

5 7d ago A 175 tokens original MIT

AGENTS

93

loki-ai-ch/spec-manager

Agent Claude Code

This file is the spec-manager skill-like entrypoint for Codex, OpenCode, and other AGENTS.md-compatible tools. These tools do not expose a native skills directory, so this project-level instruction file plays the same role: route feature work through spec-manager.

5 1mo ago A 0 tokens original MIT

CLAUDE

94

loki-ai-ch/spec-manager

Agent Claude Code

This project uses spec-manager via the /spec-manager skill.

5 1mo ago A 0 tokens original MIT

WINDSURF

95

loki-ai-ch/spec-manager

Agent Claude Code

This file is the spec-manager entrypoint for Windsurf. Windsurf reads project rules from .windsurfrules; route feature work through spec-manager.

5 1mo ago A 0 tokens copy · 81% MIT