Testing

18,394 mods in this category, of every kind an agent can take. Each one carries what it costs per session, what the scan found, and whether it is the original.

代码调试

913

jianchen08/Agent-os-open

Skill Claude CodeCodex

A debugging workflow that reproduces a problem with a failing test, collects evidence, finds the underlying cause, applies a fix, and runs regression checks. Regression checks verify that the fix does not break behavior that already worked.

not rated 5 changed 5d ago A SkillSpector: pass 82 tokens original Apache-2.0

testing

914

itsnex1s/awesome-claude-skills

Skill Claude Code

Testing with Vitest, Jest, and Playwright. Use for unit tests, integration tests, E2E tests, and coverage reports.

not rated 5 6mo ago A 30 tokens original MIT

sub-zero

915

henchmarketing-rgb/sub-zero-skill

Plugin Claude Code

Bundles 1 command, 2 agents · 137 tokens together

Audit a project, plan the path to a win condition, then execute step by step with real-world verification. A fresh-context verifier sub-agent checks the world before any success is declared.

not rated 5 3mo ago A tokens not measured original MIT

agent-evaluation

916

nek1987/auto-agent-harness

Skill Claude Code

Evaluate and improve Claude Code commands, skills, and agents. Use when testing prompt effectiveness, validating context engineering choices, or measuring improvement quality.

not rated 5 5mo ago A 32 tokens

add-vertical

917

LamantinAI/mayak

Skill Claude CodeCodex

Use this skill when adding a new feature vertical to this template — a domain aggregate with its port, repository, application service, DTOs, endpoint, and tests. Trigger it for requests like "add an endpoint", "add a new entity", "build the X feature", or "wire up a new service".

not rated 5 changed yesterday A SkillSpector: warn 67 tokens original MIT

claude-code-test

918

LoNebula/lluminai

Skill Claude CodeCodex

A development rule for a FastAPI and SQLite project that requires tests to run with `pytest -v`, then requires failures to be fixed and tests rerun until coverage reaches 100%. FastAPI is a Python web framework, and pytest is a Python testing tool.

not rated 5 21d ago A 0 tokens original MIT

dev-workflow

919

seabbs/skills

Plugin Claude Code

Bundles 12 skills, 1 agent · 391 tokens together

Development workflow skills for commits, linting, testing, code review, PRs, documentation, coverage, dependencies, project scaffolding, and task automation.

not rated 5 16d ago A tokens not measured original MIT

setup-voicecheck

920

sujitnoronha/voicecheck

Skill Claude Code

Set up VoiceCheck end to end — install the package with the right extras, run a zero-key smoke test to prove the pipeline works, configure a transport (LiveKit, Daily, VAPI, or Retell) and audio providers, then scaffold, validate, and run a first scenario. Use when someone wants to install, onboard onto, or get…

not rated 5 2mo ago A 96 tokens original MIT

verify-e2e

921

pablomarin/claude-codex-forge

Agent Claude Code

E2E verification — executes user-journey use cases through user-facing interfaces (API, UI via Playwright MCP, CLI) and produces a markdown report. Read-only: cannot modify code or write files.

not rated 5 changed 7d ago A 0 tokens original MIT

persona-qa

922

akashkendre1298/Skills

Skill Claude CodeCodex

Invoke this skill for any testing or quality assurance task — writing unit tests, integration tests, E2E tests, auditing test coverage, finding edge cases, or reviewing existing test suites. Triggers on: 'write tests for this', 'add unit tests', 'find edge cases', 'audit my test coverage', 'my tests are flaky', 'write…

not rated 5 3mo ago A 184 tokens original MIT

verify

923

ChiruMori/my-claude-code

Skill Claude CodeCodex

Verify a code change does what it should by running the app.

not rated 5 5mo ago A 15 tokens

35-testing-tdd

925

feral-file/ff-cli

Cursor rule Cursor needs its repo

TDD workflow and verification gates for FF1-CLI.

not rated 5 yesterday A 188 tokens original MIT

e2e-runner

926

mh2-lee/everything-claude-code

Agent Claude Code

End-to-end testing specialist using Playwright. Use PROACTIVELY for generating, maintaining, and running E2E tests. Manages test journeys, quarantines flaky tests, uploads artifacts (screenshots, videos, traces), and ensures critical user flows work.

not rated 5 7mo ago A 59 tokens

e2e-setup

927

Code-Shock/claude-skills

Skill Claude Code

Set up Playwright E2E testing framework for any project — auto-detects stack, auth, routes, and selectors.

not rated 5 2mo ago A 31 tokens

e2e

928

colliercoder/forge-methodology

Command Claude Code

Run automated E2E browser tests for a FORGE feature, capturing screenshots and generating a comprehensive test report.

not rated 5 7mo ago A 0 tokens

unit-test-gen

929

FanaticsKang/code-wiki

Skill Claude Code

A repository-wide tool that creates and runs unit tests for Python code with pytest or C++ code with Google Test.

not rated 5 4mo ago A 128 tokens

allure-ru

930

yugoru/allure-styleguide-ru

Skill Claude CodeCodex

A Russian-language style guide and editor for Allure test annotations. Allure is a tool that turns test metadata, titles, descriptions, and steps into readable test reports.

not rated 5 1mo ago A 267 tokens

qa-expert

931

laravel-agent-kits/claude-livewire-starter-kit

Agent Claude Code

Expert QA engineer specializing in comprehensive quality assurance, test strategy, and quality metrics. Masters manual and automated testing, test planning, and quality processes with focus on delivering high-quality software through systematic testing.

not rated 5 7mo ago A 43 tokens

audit-tests

932

ashrust/lets-start-skill

Skill Claude CodeCodex

Audits a repo's test suite against a short rubric and scaffolds a comprehensive one if it's thin or missing. Detects language and test framework, measures coverage where possible, identifies gaps (no unit tests, no integration tests, no CI hook, slow or flaky suite, uncovered critical paths), and presents a scored…

not rated 5 3mo ago A 196 tokens

skill-auditor

933

tlzmw001/naiyue-skills

Skill Claude CodeCodex

A testing skill that checks whether a third-party coding-agent skill does what its documentation claims. It breaks claims into individual items and tests them in a temporary workspace.

not rated 5 1mo ago A 187 tokens original MIT

tckit

934

georgeturneruk/tckit

Plugin Claude Code

Bundles 1 MCP server

MCP server for TwinCAT 3 PLC projects: read structure, write ST code, trigger builds, deploy to targets, run TcUnit tests.

not rated 5 17d ago A tokens not measured original MIT

agentloop

935

ArcBlock/agent-skills

Plugin Claude Code

Bundles 20 skills · 2,536 tokens together

Repo-agnostic engineering loop engine: config-driven verification gate, sticky PR-comment delivery, and the review/sweep skill contracts. Repo specifics live in each repo's .claude/verify/config.ts, not here.

not rated 5 changed yesterday A tokens not measured original MIT

my-skills

936

mizzy/my-skills

Plugin Claude Code

Bundles 10 skills · 332 tokens together

Personal development workflow skills: brainstorming, planning, TDD, debugging, verification, review, and Codex integration.

not rated 5 24d ago A tokens not measured

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: