tdd

tdd is a skill for Claude Code, Codex from EthanSei/skills. It costs 93 tokens per session (1,196 once invoked), scanned A, original, MIT.

A test-driven development workflow, meaning tests are written before the code they check. It repeats a red-green-refactor cycle: write a failing test, make it pass, then clean up the code while keeping the test passing.

In plain words
What is it for?
Use it when adding or changing behavior and you want the tests to guide the implementation. It covers writing focused tests, implementing the minimum needed, refactoring, and running the full checks.
Why use it?
It turns requirements into specific checks before implementation begins. Running tests after each step helps reveal mistakes close to where they were introduced.

Skill for Claude CodeCodex

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/ethansei/skills/tdd
Any agent
npx skills add EthanSei/skills --skill tdd
Clone the repo
git clone --depth 1 https://github.com/EthanSei/skills

Made for: Claude Code, Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for tdd

README.md
[![agentmods](https://agentmods.dev/badge/skills/ethansei/skills/tdd.svg)](https://agentmods.dev/skills/ethansei/skills/tdd)
Your own site
<a href="https://agentmods.dev/skills/ethansei/skills/tdd"><img src="https://agentmods.dev/badge/skills/ethansei/skills/tdd.svg" alt="Measured on agentmods" height="20"></a>
Per session 93 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 1,196 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00093 $0.01196
Opus 5 $0.00046 $0.00598
Sonnet 5 $0.00019 $0.00239
Haiku 4.5 $0.00009 $0.00120

Measured 5d ago against content hash 5cff26499e18, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

tdd scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 5d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

skills/tdd/SKILL.md · 144 lines

How it starts

The opening of the file, as written. The whole thing — 144 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Test-Driven Development

Write the test first. Make it pass. Clean up. Run the tests. Every time.

The Cycle

Every code change follows red-green-refactor:

  1. Red — Write a test for the next behavior. Run it. Watch it fail.
  2. Green — Write the minimum code to make the test pass. Nothing more.
  3. Refactor — Clean up duplication and improve clarity. Tests stay green.
  4. Repeat — Pick the next behavior. Start another cycle.

Run tests after every step. Not just at the end — after every step.

1. Start With a Failing Test

Before writing any implementation, write a test that:

  • Describes one specific behavior or requirement
  • Fails for the right reason (missing function, wrong return value — not a syntax error)
  • Uses clear, descriptive names that read as specifications

Do not write multiple tests at once. One test, one behavior, one cycle.

2. Write the Minimum to Pass

Make the failing test pass with the simplest code that works. Do not:

  • Handle edge cases you haven't tested yet
  • Build abstractions before you have three examples
  • Add error handling for scenarios not covered by a test
  • Optimize before the behavior is correct

If you're writing code that no test exercises, stop. Write the test first.

3. Refactor Under Green Tests

After the test passes, improve the code while keeping tests green:

  • Extract common patterns only when you see real duplication
  • Rename for clarity
  • Simplify complex conditionals
  • Remove dead code

Run tests after each refactoring move. If any test turns red, undo and try a smaller change.

4. Run Tests After Every Change

This is not optional. Run the relevant test suite:

  • After writing a new test (confirm it fails)
  • After writing implementation (confirm it passes)
  • After every refactoring move
  • Before considering any task done

If you cannot run tests (no test runner configured, unfamiliar project), tell the user and ask how to run them. Do not skip testing and move on silently.

5. Discover the Test Runner

Read the full file on GitHub · 144 lines

Files

What ships with it

2 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 5d ago First seen · 144 lines · 93 tokens per session scan A 5cff26499e18

Subscribe to this mod's changes

tdd is a skill published in the GitHub repository EthanSei/skills (2 stars, last pushed 6mo ago), licensed MIT. It adds 93 tokens to every session and 1,196 once invoked, about $0.0005 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.