test-master

test-master is a skill for Claude Code from Jeffallan/claude-skills. It costs 104 tokens per session (1,148 once invoked), scanned A, original, MIT.

A software-testing guide for planning and writing tests across unit, integration, end-to-end, performance, and security checks. It also covers mocking, coverage, test failures, flaky tests, and defect reports.

In plain words
What is it for?
It helps generate test files, choose test approaches, create mocks, run tests, classify failures, stabilize flaky tests, measure coverage, and report defects with severity and suggested fixes.
Why use it?
It helps replace ad hoc testing with a defined scope and strategy, while making failures and coverage gaps explicit. It also guides investigation of flaky tests, which do not fail consistently.

Skill for Claude Code

Written for Claude Code: shipped in a Claude Code plugin.

Part of the fullstack-dev-skills plugin — 67 skills shipped together

Good fit It helps generate test files, choose test approaches, create mocks, run tests, classify failures, stabilize flaky tests, measure coverage, and report defects with severity and suggested fixes.

Compare 6 skills from other repositories ↓
Install with agentmods
npx agentmods add skills/jeffallan/claude-skills/test-master
About the project

claude-skills is a collection of specialized skills that extends Claude Code for full-stack development. Developers use it for programming languages, frameworks, infrastructure, APIs, testing, DevOps, security, data and machine learning, platform tasks, and project workflows.

Jeffallan/claude-skills · 11,766 stars · on GitHub

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Any agent
npx skills add Jeffallan/claude-skills --skill test-master
Clone the repo
git clone --depth 1 https://github.com/Jeffallan/claude-skills

Made for: Claude Code.

Or install fullstack-dev-skills, the plugin that ships this one along with the rest of its 67 skills.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for test-master

README.md
[![agentmods](https://agentmods.dev/badge/skills/jeffallan/claude-skills/test-master/github.svg)](https://agentmods.dev/skills/jeffallan/claude-skills/test-master)
Your own site
<a href="https://agentmods.dev/skills/jeffallan/claude-skills/test-master"><img src="https://agentmods.dev/badge/skills/jeffallan/claude-skills/test-master/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for test-master

Your own site · 80×15
<a href="https://agentmods.dev/skills/jeffallan/claude-skills/test-master"><img src="https://agentmods.dev/badge/skills/jeffallan/claude-skills/test-master.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 104 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 1,148 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe. Third-party audits
  • Socket pass 14 Sept 2026
  • Snyk pass 14 Sept 2026
How audits are shown
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00104 $0.01148
Opus 5.5 $0.00042 $0.00459
Sonnet 5.5 $0.00021 $0.00230
Haiku 4.5 $0.00010 $0.00115

Measured 4d ago against content hash 93c3710547af, method: parsed. Prices are Anthropic first-party input rates as of 2026-10-07, from the pricing page.

Security

Grade A, and why

test-master scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 4d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

Origin

Copies of this mod

2 near-identical copies found in the catalogue:

skills/test-master/SKILL.md · 100 lines

How it starts

The opening of the file, as written. The whole thing — 100 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Test Master

Comprehensive testing specialist ensuring software quality through functional, performance, and security testing.

Core Workflow

  1. Define scope — Identify what to test and which testing types apply
  2. Create strategy — Plan the test approach across functional, performance, and security perspectives
  3. Write tests — Implement tests with proper assertions (see example below)
  4. Execute — Run tests and collect results
    • If tests fail: classify the failure (assertion error vs. environment/flakiness), fix root cause, re-run
    • If tests are flaky: isolate ordering dependencies, check async handling, add retry or stabilization logic
  5. Report — Document findings with severity ratings and actionable fix recommendations
    • Verify coverage targets are met before closing; flag gaps explicitly

Quick-Start Example

A minimal Jest unit test illustrating the key patterns this skill enforces:

// ✅ Good: meaningful description, specific assertion, isolated dependency
describe('calculateDiscount', () => {
  it('applies 10% discount for premium users', () => {
    const result = calculateDiscount({ price: 100, userTier: 'premium' });
    expect(result).toBe(90); // specific outcome, not just truthy
  });

  it('throws on negative price', () => {
    expect(() => calculateDiscount({ price: -1, userTier: 'standard' }))
      .toThrow('Price must be non-negative');
  });
});

Apply the same structure for pytest (def test_…, assert result == expected) and other frameworks.

Reference Guide

Load detailed guidance based on context:

Topic Reference Load When
Unit Testing references/unit-testing.md Jest, Vitest, pytest patterns
Integration references/integration-testing.md API testing, Supertest
E2E references/e2e-testing.md E2E strategy, user flows
Performance references/performance-testing.md k6, load testing
Security references/security-testing.md Security test checklist
Reports references/test-reports.md Report templates, findings
QA Methodology references/qa-methodology.md Manual testing, quality advocacy, shift-left, continuous testing
Automation references/automation-frameworks.md Framework patterns, scaling, maintenance, team enablement
TDD Iron Laws references/tdd-iron-laws.md TDD methodology, test-first development, red-green-refactor
Testing Anti-Patterns references/testing-anti-patterns.md Test review, mock issues, test quality problems

Read the full file on GitHub · 100 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 4d ago First seen · 100 lines · 104 tokens per session scan A 93c3710547af

Subscribe to this mod's changes

test-master is a skill published in the GitHub repository Jeffallan/claude-skills (11,766 stars, last pushed 4d ago), licensed MIT. It adds 104 tokens to every session and 1,148 once invoked, about $0.0004 per session on Opus 5.5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-10-04.

Related

Other skills, from other repositories

req-to-test

Generates comprehensive test scenarios from requirements including BDD/Gherkin scenarios, unit tests, integration tests, and end-to-end test cases. Use when converting requirements, user stories, or specifications into testable scenarios with full coverage including happy paths, error cases, edge cases, and boundary…

ArabelaTso/Skills-4-SE · 71 tokens

ios-testing

Invoke any time a user is writing iOS/Swift tests or asking why tests behave a certain way — including XCTest versus Swift Testing (@Test/#expect) choices, async ViewModel tests with @Observable or @Published, snapshot testing across device sizes, mocking protocols for dependency injection, setUp/tearDown lifecycle…

rusel95/ios-agent-skills · 134 tokens

test-strategy-document

Create a production-ready Testing Strategy and QA Execution Plan. Covers testing levels (unit, integration, E2E, performance), mocking boundaries, test environment matrix, code coverage thresholds, and automated CI pipeline runsheets. Use when establishing a QA framework for a new system or feature set.

fattain-naime/engineering-docs · 62 tokens

test-strategy

Use when deciding what to test and at which level. Covers the test pyramid, what belongs in unit versus integration versus end-to-end tests, coverage as a signal rather than a target, and eliminating flakiness.

nimadorostkar/Claude-Skills-collection · 47 tokens

writing-tests

Write unit, integration, and E2E tests that follow the testing pyramid, the Arrange-Act-Assert pattern, the shouldXwhenY naming convention, and meet a 75% (target 80%) branch coverage gate per supported OS. Use when the user asks to write tests, add coverage, build a test suite, fix flaky tests, or verify a feature…

Cristhianzl/claude-skills-czl · 103 tokens

test-levels

This skill explains the 3 test levels (Unit, Integration, E2E) using the "Building a Car" analogy and provides guidance on when to use each type. Includes project-specific Playwright examples.

georgekhananaev/claude-skills-vault · 46 tokens