Testing

18,394 mods in this category, of every kind an agent can take. Each one carries what it costs per session, what the scan found, and whether it is the original.

coding-quality-loop

889

zaingz/coding-quality-loop

Skill Claude CodeCodex

Use when a coding agent must turn a software goal, bug, issue, or refactor into a small, verified, independently reviewed code change.

not rated 5 1mo ago A 33 tokens original MIT

squirrel

891

flyingsquirrel0419/squirrel-skill

Skill Claude CodeCodex

Full-cycle software development agent: plans, builds, tests, lints, fixes bugs, and writes production-grade README docs. ALWAYS use this skill when the user wants to: build or scaffold a new project, add features to existing code, fix bugs, improve code quality (lint, format, refactor), write or improve a README, add…

not rated 5 4mo ago A 204 tokens original Apache-2.0

factory-render-verify

892

squidbay/factory

Skill Claude CodeCodex needs its repo

Render-and-measure receipts for any HTML page your factory builds — the render half of the design quality gate. Engineer runs it to screenshot every screen size and MEASURE what a source read or a single screenshot only guesses at: horizontal overflow, computed type sizes, tap-target sizes, safe-area presence, mono…

not rated 5 20d ago A 133 tokens original MIT

reviewing-unit-tests

893

burugo/behavior-driven-development

Skill Claude CodeCodex

Use in a dedicated subagent for post-TDD blind unit-test review against requirements and public contracts to find missing cases, wrong assertions, and brittle tests.

not rated 5 4mo ago A 36 tokens

nestjs-hexagonal

894

Softtor/nestjs-hexagonal

Plugin Claude Code

Bundles 10 skills, 8 agents · 1,211 tokens together

Skills for building NestJS bounded contexts with Hexagonal Architecture, DDD, and CQRS patterns. Covers domain modeling, application layer, infrastructure wiring, presentation, full TDD workflow, and architecture review.

not rated 5 1mo ago A tokens not measured original MIT

production-readiness

895

Meghshyams/Production-Readiness

Plugin Claude Code

Bundles 1 skill · 60 tokens together

Comprehensive production readiness audit — 75+ checks across 9 pillars: security & supply chain, visual QA, code quality, testing, error handling & observability, build, performance, accessibility (WCAG 2.2), and AI/LLM safety. Like having a senior engineer + QA tester do a final review before deploy.

not rated 5 2mo ago A tokens not measured original MIT

fable-discipline

896

assafkip/fable-discipline

Plugin Claude Code

Bundles 1 skill, 1 hook · 137 tokens together

Fable's engineering discipline for any Claude session, with one habit enforced by a hook. Two layers: running the task (stage, verify each stage with a failable check, written done-criteria) and writing the code (recon before edit, verify against a copy with a negative self-test, single-writer chokepoints…

not rated 5 2mo ago A tokens not measured original MIT

bench-creator

898

EvanLuo42/bench-creator

Plugin Claude Code

Bundles 1 skill · 116 tokens together

Capture private, reproducible AI coding benchmark cases from real repository work.

not rated 5 1mo ago A tokens not measured original MIT

write-tests

899

Anyesh/skillprobe

Skill Claude CodeCodex needs its repo

Write skillprobe YAML tests for LLM skills. Use when asked to write tests, create test scenarios, test a skill, generate skillprobe tests, or check whether a skill activates correctly.

not rated 5 5mo ago A 41 tokens original MIT

sap-fiori-testing

900

efeumutaslan/SAP-SKILLS

Skill Claude CodeCodex

SAP Fiori/UI5 testing, accessibility, and modern UI skill. Use when writing wdi5 E2E tests, OPA5 integration tests, implementing WCAG 2.1 accessibility, using UI5 Web Components, or migrating UI5 to TypeScript. If the user mentions wdi5, OPA5, Fiori test, UI5 accessibility, WCAG, or UI5 TypeScript migration, use this…

not rated 5 5mo ago A 93 tokens original MIT

meto-tester

901

iLomer/Metho_agentic

Agent Claude Code

Validate work in tasks-in-testing.md. Full acceptance criteria are in the task block. One item at a time, always sequential. Never fixes bugs, only flags and sends back.

not rated 5 3mo ago A 41 tokens original MIT

archetypeai/agent-skills

Skill Claude CodeCodex

Run Archetype AI's managed Task Verification (TVA) agent over the Agents API — upload a recording AND a reference procedure (an SOP), create a bundle from the tva blueprint, run it, poll, download a per-step PASSED / FAILED / MISSING verdict per step. Use when the user has a recording of work that should have followed…

not rated 5 21d ago A 216 tokens original Apache-2.0

skill-suite-tests

904

andreferraro/skill-suite-tests

Skill Codex

Analisa riscos e cria, adapta, executa e valida testes automatizados em projetos existentes. Use quando o usuário pedir testes para uma tela, fluxo, regra, serviço, API, banco, evento, integração, bug ou atributo de qualidade e esperar código integrado à stack, execução e evidências reais.

not rated 5 28d ago A 65 tokens original MIT

hookdeck/agent-skills

Cursor rule Cursor

Use hookdeck listen for local webhook testing—prefer it over ngrok or public endpoints; never test in production.

not rated 5 22d ago A 23 tokens original MIT

dsh-chaos-test

906

cyanseek/dsh-tool-chaos

Skill Codex

Design, install, run, and report deterministic DeepSeek Harness tool-failure experiments. Use when a user wants to prove retry or fallback behavior, timeout or cooperative cancellation, policy-denial handling, blocked-result recovery, Code Mode nested-call resilience, or CI evidence for a DSH agent/plugin. Complete…

not rated 5 14d ago B 91 tokens original MIT

cc-autopilot

907

gbotev1/cc-autopilot

Plugin Claude Code

Bundles 1 skill, 23 agents · 1,049 tokens together

Autonomous judge agents fix your app's debt, build what's missing, and verify every change through a browser or gate suite before it lands, then find the next thing and keep going until you say stop. Each fixer owns a file-disjoint slice of the tree, so no two can ever collide, and nothing unverified reaches git. A…

not rated 5 1mo ago A tokens not measured original Apache-2.0

trapstreet/trapstreet-skills

Skill Claude CodeCodex

Design and scaffold a new trapstreet.run task to evaluate a given agent/skill/tool -- the reverse of trapstreet-solution-scaffold (a solution for an existing task). Generates the mechanical parts (traptask.yaml, judge.py/grader.py on the TRAPTASKMANIFEST contract, buildcases.py's validate-then-render pipeline) and…

not rated 5 15d ago A 231 tokens original MIT

playwright-cli

909

IvanCampos/agents

Skill Codex

Translate natural-language browser automation requests into exact playwright-cli commands for interactive web testing and debugging. Use when requests involve opening/navigating pages, interacting with elements, capturing snapshots/screenshots/PDFs, using tabs, inspecting console/network, mocking routes, managing…

not rated 5 7mo ago A 74 tokens original MIT

zmr

912

johnmikel/zeno-mobile-runner

Plugin Claude Code

Bundles 1 skill · 50 tokens together

Agent-native mobile UI automation and verification for Expo, React Native, Flutter, and native Android/iOS apps. Gives Claude Code eyes and hands on emulators and simulators: semantic snapshots, typed actions, waits, assertions, and replayable traces.

not rated 5 27d ago A tokens not measured original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: