test-driven-development

test-driven-development is a skill for Claude Code from jcwleo/oh-no-harness. It costs 42 tokens per session (2,721 once invoked), scanned A, original, MIT.

A test-first software development method, often called TDD: write a test, see it fail, make the smallest change that passes, then refactor. Refactoring means improving code structure without changing its behavior.

In plain words
What is it for?
Use it for behavior changes and bug fixes inside an implementation workflow, not as a replacement for planning or final verification.
Why use it?
It makes the intended behavior explicit before production code changes and helps catch incorrect implementations early.

Skill for Claude Code

Written for Claude Code: argument-hint in frontmatter. Also seen: mentions subagents; mentions Claude Code.

Part of the oh-no-harness plugin — 12 skills, 12 commands, 9 agents, 1 hook shipped together

Good fit Use it for behavior changes and bug fixes inside an implementation workflow…

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/jcwleo/oh-no-harness/test-driven-development
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add jcwleo/oh-no-harness --skill test-driven-development
Clone the repo
git clone --depth 1 https://github.com/jcwleo/oh-no-harness

Made for: Claude Code.

Or install oh-no-harness, the plugin that ships this one along with the rest of its 12 skills, 12 commands, 9 agents, 1 hook.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for test-driven-development

README.md
[![agentmods](https://agentmods.dev/badge/skills/jcwleo/oh-no-harness/test-driven-development.svg)](https://agentmods.dev/skills/jcwleo/oh-no-harness/test-driven-development)
Your own site
<a href="https://agentmods.dev/skills/jcwleo/oh-no-harness/test-driven-development"><img src="https://agentmods.dev/badge/skills/jcwleo/oh-no-harness/test-driven-development.svg" alt="Measured on agentmods" height="20"></a>
Per session 42 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 2,721 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00042 $0.02721
Opus 5 $0.00021 $0.01360
Sonnet 5 $0.00008 $0.00544
Haiku 4.5 $0.00004 $0.00272

Measured 6d ago against content hash 04b8f91a3d27, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-06, from the pricing page.

Security

Grade A, and why

test-driven-development scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 6d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

plugins/oh-no-harness/skills-claude/test-driven-development/SKILL.md · 289 lines

How it starts

The opening of the file, as written. The whole thing — 289 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Test Driven Development for Claude Code

This generated file is the Claude Code-facing runtime skill document. Claude Code slash commands should read this file directly; maintainers edit the source documents listed below instead.

Generated Runtime Composition

Source order:

  • ../../docs/skill-core/test-driven-development.md
  • ../../docs/platforms/claude-code-runtime.md

The sections below are already composed for this platform. Do not ask the runtime model to load another platform's runtime document or invocation syntax.

Source: docs/skill-core/test-driven-development.md

Test Driven Development

Write the test first. Watch it fail. Write the smallest production change that makes it pass. Refactor only after green.

Software Development Stage

Test Driven Development is the internal mid-loop discipline for behavior-change work inside implementation and bug-fix execution.

Use it inside ralph, systematic-debugging, or ultrawork before changing production behavior. It is not a requirements, planning, cleanup, or final-verification substitute.

Top-Level Routing Boundary

Do not use this skill as the top-level route for ordinary implementation requests such as "add this feature", "fix this bug", "refactor this module", or "implement this behavior". This is not a top-level implementation skill. Use ralph for those concrete implementation requests so execution mode, worktree isolation, review, cleanup, verification, and final reporting stay attached to the work.

Use this skill as a top-level entry only when:

  • the user explicitly asked for TDD, test-first, RED/GREEN/REFACTOR, or "write the failing test first"
  • an already-selected workflow (ralph, systematic-debugging, or ultrawork) reaches its internal TDD gate

Direct-edit-eligible mutations are inert and never reach this skill; other small concrete edits route through ralph, which may apply its STANDARD small-task carve-out; there is no separate direct edit path that reaches this skill outside a workflow or an explicit user TDD request.

Read the full file on GitHub · 289 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 6d ago First seen · 289 lines · 42 tokens per session scan A 04b8f91a3d27

Subscribe to this mod's changes

test-driven-development is a skill published in the GitHub repository jcwleo/oh-no-harness (11 stars, last pushed 29d ago), licensed MIT. It adds 42 tokens to every session and 2,721 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.