test

A testing workflow for deciding what to test, writing and running tests, and reporting coverage and gaps. TDD, or test-driven development, means writing a failing test before the code and then making it pass.

In plain words
What is it for?
Adding unit, integration, and interaction tests; checking critical paths such as authentication, payments, forms, navigation, and API calls.
Why use it?
It reduces the chance that changes break important behavior, edge cases, user interactions, or connections to other services.

Skill for Claude CodeCodex

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/darkroomengineering/cc-settings/test
Any agent
npx skills add darkroomengineering/cc-settings --skill test
Clone the repo
git clone --depth 1 https://github.com/darkroomengineering/cc-settings

Made for: Claude Code, Codex.

Per session 73 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 1,417 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00073 $0.01417
Opus 5 $0.00036 $0.00709
Sonnet 5 $0.00015 $0.00283
Haiku 4.5 $0.00007 $0.00142

Measured 2d ago against content hash 42e35b05d460, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

test scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 2d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

skills/test/SKILL.md · 151 lines

How it starts

The opening of the file, as written. The whole thing — 151 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Testing Workflow

Delegates to the Tester agent for test coverage and verification.

Workflow

  1. Identify - Determine what needs testing (component, hook, utility, integration)
  2. Write - Create test files colocated with source (e.g., button.test.tsx)
  3. Run - Execute tests and verify results
  4. Report - Summarize coverage and gaps

Testing Priorities

  1. Critical paths - Auth, payments, core features
  2. Edge cases - Error states, empty states, boundaries
  3. User interactions - Forms, buttons, navigation
  4. Integration points - API calls, external services

Rationalization Counters

If you catch yourself thinking any of the following, STOP — you are skipping testing:

Rationalization Why It's Wrong
"This is too simple to test" Simple functions with edge cases are exactly what tests catch
"I'll add tests later" Later never comes; untested code ships and breaks
"The types guarantee correctness" Types check structure, not logic — add(a, b) returning a - b passes TypeScript
"It's just UI, tests don't help" Interaction tests catch regressions that visual review misses
"Manual testing is enough" Manual testing doesn't run in CI and doesn't prevent regressions
"Tests are passing immediately" Tests that pass on first run without failing first may not be testing what you think — verify the test actually exercises the code path

Red Flags

  • Tests pass immediately on first write: Suspicious. Verify the test would fail if the implementation were wrong.
  • No assertions: A test without assertions is not a test.
  • Mocking everything: If you mock the thing you're testing, you're testing the mock.
  • A batch of tests written before any has run: a batch against imagined behavior pins what you guessed — those tests fail on harmless changes and pass while the real path is broken. Write one, watch it go red, earn green, then write the next; each cycle tells you what the next test should actually assert.
  • Asserting internals: assert observable behavior through the outermost practical entry point (return values, exit codes, persisted rows, rendered output) — never which internal functions ran. A public-surface test survives a rewrite of everything underneath; an internals test breaks on every refactor and pins implementation, not behavior.

Read the full file on GitHub · 151 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 2d ago First seen · 151 lines · 73 tokens per session scan A 42e35b05d460

Subscribe to this mod's changes

test is a skill published in the GitHub repository darkroomengineering/cc-settings (42 stars, last pushed 4d ago), licensed MIT. It adds 73 tokens to every session and 1,417 once invoked, about $0.0004 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other skills, from other repositories

dotfiles-bootstrap

Bootstrap a workstation with the dotfiles framework. Takes a GitHub user / owner+repo / explicit clone URL and runs dot init (which shells out to chezmoi) with the right safety prompts. Honors the active agent profile (ask / plan / apply / audit) so it defaults to dry-run in safer modes and full apply in apply.

sebastienrousseau/dotfiles · 88 tokens

astro-dso-doc

Generates a complete, polished HTML documentation page, a processing checklist, an AstroBin post JSON, a PixInsight process icon set (XPSM), AND a ready-to-paste PixInsight project Description field for a deep-sky object (DSO) astrophotography project. Use this skill whenever the user mentions astrophotography, a DSO…

jjmartres/ai-coding-agents · 244 tokens

document-code

Apply Google Style documentation standards to Python, Go, TypeScript, and Terraform code. Use when writing or reviewing code that needs docstrings/comments/JSDoc, when asked to "document this code", "add docstrings", "follow Google Style", or when improving code documentation quality. Supports Python docstrings, Go…

jjmartres/ai-coding-agents · 88 tokens

work-on-ticket

Fetches Jira ticket details, creates an appropriately named branch, and initiates the task planning workflow. Use when the user says "work on [TICKETID]" or similar phrases.

jjmartres/ai-coding-agents · 41 tokens

datadog

Use this skill when you need to search Datadog logs, query metrics, tail logs in real-time, trace distributed requests, investigate errors, compare time periods, find log patterns, check service health, or export observability data.

jjmartres/ai-coding-agents · 51 tokens

document-project

Generate comprehensive, professional project documentation structures including README, ARCHITECTURE, USERGUIDE, DEVELOPERGUIDE, and CONTRIBUTING files. Use when the user requests project documentation creation, asks to "document a project", needs standard documentation files, or wants to set up docs for a new…

jjmartres/ai-coding-agents · 79 tokens