Testing

18,481 mods in this category, of every kind an agent can take. Each one carries what it costs per session, what the scan found, and whether it is the original.

wincreator

1105

winterbim/wincreator

Skill Codex

Proves that agent work is actually done: each claim becomes a ledger row with a gate that was executed, captured raw evidence, and a status a builder is not allowed to write for itself. Use when a technical result must be auditable later — shipping or migrating something users depend on, a change whose failure is…

not rated 3 yesterday A 156 tokens original MIT

specmint-tdd

1106

ngvoicu/specmint-tdd

Skill Claude CodeCodex

TDD-first spec management for AI coding workflows. Use this skill when the user explicitly mentions specs, forging, or structured planning: says "forge", "forge a spec", "write a spec for X", "create a spec", "plan X as a spec", "resume", "what was I working on", "spec list/status/pause/switch/activate", "implement…

not rated 3 2mo ago A 146 tokens copy · 97% MIT

engineering-skills

1107

int2t05/engineering-skills

Plugin Claude Code

Bundles 47 skills, 1 hook · 3,422 tokens together

Unified engineering skills pack for Claude Code — 47 skills organized by the software development lifecycle (product → research → design → develop → tune → test → verify → ship → operate), with shared engineering principles injected at session start, plus product and design principle references linked by domain skills.

not rated 3 changed 8d ago A tokens not measured original MIT

guardsman

1108

hedimanai-pro/guardsman

Plugin Claude Code

Bundles 1 skill · 221 tokens together

Right-sized code, verified before it ships. Risk-tiered engineering judgment for AI coding agents: convention detection, blast-radius tiers, mandatory verification, and a scannable logbook for deliberate shortcuts.

not rated 3 2mo ago A tokens not measured original MIT

test-generator

1109

u9401066/template-is-all-you-need

Skill Claude Code

A test-suite generation workflow for software projects. It includes static checks, unit tests, integration tests, end-to-end tests, coverage reports, type checking, linting, security scanning, and dead-code detection.

not rated 3 6mo ago A 98 tokens original Apache-2.0

android-qa

1110

willbytee-sudo/android-qa-kit

Skill Claude Code

Test an Android app on a real phone or an emulator using adb. Use when the user wants to install an APK they just compiled, walk through the app's screens, check the on-screen text, reproduce a bug, capture screenshots, read crash logs, or set up an Android testing environment from scratch. Covers both a physical…

not rated 3 1mo ago A 80 tokens original MIT

qa-go

1111

endorphin-ai/claude-code-teams

Skill Claude Code

Go testing skill with table-driven tests, httptest, testify, and integration testing patterns. Use when writing or running Go backend tests.

not rated 3 6mo ago A 30 tokens

run-tests

1112

JSchOBL/agentic-ai-learning-journey

Skill Claude Code

Run the pytest suite, report pass/fail counts and coverage, and identify untested code. Use when the user asks to run tests, check test coverage, or verify that changes didn't break anything.

not rated 3 1mo ago A 43 tokens original MIT

no-vibes

1113

Lum1104/no-vibes

Skill Claude CodeCodex

Use when completion depends on an end-to-end outcome across components, environments, or external systems.

not rated 3 1mo ago A 23 tokens original MIT

testing-arsenal

1114

FutureJJ/claude-skills

Skill Claude CodeCodex

Testing strategies including unit tests, integration tests, E2E tests, mocking, coverage analysis, and TDD workflow. Trigger when users need help writing tests, choosing testing frameworks, implementing mocking strategies, or setting up test infrastructure.

not rated 3 6mo ago A 51 tokens original MIT

vl-stage

1115

giltotherescue/velocity-agent-skills

Skill Claude CodeCodex

Use when a branch, worktree, local app, or cloud-agent task needs a staged test environment without disturbing the main checkout, especially for browser-facing web app work. Helps agents choose and set up the right strategy for browser/app testing across JavaScript/TypeScript, Python, PHP, Docker Compose, local…

not rated 3 3mo ago A 108 tokens original MIT

ironbee-ai/ironbee-devtools-skills

Skill Claude Code

CLI for driving an Android emulator (AVD) over adb. Use when the user needs to connect to an emulator (attach to a running one or boot/manage an AVD), tap/swipe/type on the UI, read the UI/accessibility tree, take screenshots or screen recordings, capture logcat, capture HTTP(S) requests in-process (Frida OkHttp hook…

not rated 3 1mo ago A 203 tokens original MIT

kill-test-first

1117

redamancy231-create/claude-skills

Skill Claude Code

A test-first review process for new research ideas, trading strategies, paper topics, models, architectures, or data pipelines. TDD means defining tests or checks before building the idea.

not rated 3 29d ago A 85 tokens CC-BY-4.0

code-verification

1118

bhaumikmaan/claude-code-master-skills

Skill Claude Code

Adversarial verification of code changes. Tries to break implementations rather than confirm they work. Produces structured PASS/FAIL/PARTIAL verdicts with evidence. Use when verifying code changes, after non-trivial implementations, before reporting task completion, or when asked to check if something works.

not rated 3 5mo ago A 63 tokens original MIT

testkit

1119

knowhowlab/agent-testkit

Plugin Claude Code

Bundles 3 skills · 288 tokens together

Universal E2E and manual acceptance testing for any repo: test-init generates project test plans (TESTSE2E.md, TESTSMANUAL.md), test-e2e executes them autonomously, test-manual conducts them with you at the controls. Every run writes a timestamped protocol file.

not rated 3 2mo ago A tokens not measured original MIT

a

1120

chafoo/anchored

Plugin Claude Code

Bundles 2 skills, 1 agent, 1 hook · 395 tokens together

Verification gate for AI work — one run file, one independent validator, evidence-gated close. CLI-only transport via the anchored binary over Bash, no MCP. Slash commands: /a:run /a:setup.

not rated 3 2mo ago A tokens not measured original MIT

harden

1121

Calvin-LLC/agentic-hardening-skill

Skill Claude CodeCodex

Evidence-driven codebase hardening. Audits security (OWASP Top 10:2025), supply chain (inventory vs SBOM vs SLSA v1.2), reliability (OTel + operational limits), tests (sandboxed tiers), and accessibility (WCAG 2.2 AA). Every finding is quoted, matrix-scored, and cited. Does not fix. Reports with file:line and a…

not rated 3 15d ago A 117 tokens original MIT

nen-conjurer

1122

rlaope/nen

Cursor rule Cursor

Failure-mode-driven reliability engineering — enumerate how it breaks, give every mode a verdict, prove every handler with a reproducing test. Use when the user says "flaky", "harden this before launch", "what if this fails", "error handling", "timeout", "retry", "idempotency", or "race condition"; when an incident or…

not rated 3 1mo ago A 126 tokens original MIT

maestro-testing

1123

eagleisbatman/maestro-skill

Plugin Claude Code

Bundles 1 skill · 176 tokens together

Generate and manage Maestro test flows for mobile (Android, iOS) and web apps. Covers React Native, Expo, Flutter, Swift, Kotlin, Jetpack Compose, SwiftUI, UIKit, Capacitor, Ionic, and web desktop testing.

not rated 3 6mo ago A tokens not measured original MIT

journey-simulation

1124

RockyHong/super-bootstrap

Skill Claude Code

Use when caller wants to observe how a stranger encounters a flow, artifact, or sandbox — triggers like "simulate a user journey", "test our onboarding / checkout / signup", "will my ICP convert", "how does a cold reader experience this README", "first-time user test", "cognitive walkthrough", or any request to…

not rated 3 today A 79 tokens original MIT

parallel-lifecycle

1125

DevOtts/parallel-lifecycle

Skill Claude Code

Run and CDP-test a feature inside its own git worktree without colliding with other parallel coding sessions. Use whenever you are testing a feature locally in a worktree, running Playwright/CDP against your own app, or running two or more agent sessions in parallel that each need a dev server and a browser. Trigger…

not rated 3 1mo ago A 157 tokens original MIT

ade-bench

1127

typedef-ai/ade-bench-plugin

Plugin Claude Code

Bundles 3 skills, 3 commands, 1 agent · 316 tokens together

Generate ADE-Bench benchmark tasks from your own dbt project. Scans your models, proposes realistic bug-injection scenarios, and writes the task scaffolding (config, patches, scripts, custom assertion tests) ready to run against AI agents.

not rated 3 4mo ago A tokens not measured original MIT

crucible

1128

gpanakkal/crucible

Skill Claude CodeCodex

Write solid unit tests using property-based testing and mutation testing. Use whenever unit tests are being written, fixed, audited, or reviewed in a TypeScript project; whether the user asks directly or test-writing occurs as a step inside another workflow (TDD, feature implementation, bug fixing, code review). Also…

not rated 3 1mo ago A 98 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: