Testing

18,217 mods in this category, of every kind an agent can take. Each one carries what it costs per session, what the scan found, and whether it is the original.

code-quality-agent

289

smithclay/claudetainer

Agent Claude Code

Comprehensive code quality specialist ensuring production-ready code through formatting, linting, and testing with zero violations.

not rated 107 5mo ago A 24 tokens original MIT

soba-labs/langchain-agent-skills

Skill Claude CodeCodex

Use this skill when you need to test or evaluate LangGraph/LangChain agents: writing unit or integration tests, generating test scaffolds, mocking LLM/tool behavior, running trajectory evaluation (match or LLM-as-judge), running LangSmith dataset evaluations, and comparing two agent versions with A/B-style offline…

not rated 106 20d ago A 99 tokens original MIT

tdd

291

smartfrog/opencode-froggy

Skill Claude CodeCodex

Apply Test-Driven Development workflow for new features and bugfixes.

not rated 105 3mo ago A 17 tokens original MIT

verify

292

leanchy/BingoCode

Skill Claude CodeCodex

Verify a code change does what it should by running the app.

not rated 104 2mo ago A 15 tokens

entrix

293

phodal/entrix

Plugin Claude Code

Bundles 1 skill · 70 tokens together

Harness Engineering plugin exposing executable fitness functions for architecture quality, change-aware validation, and MCP-powered review context.

not rated 104 3mo ago A tokens not measured original MIT

swarmauri/swarmauri-sdk

Skill Codex

Add a second-class standalone Swarmauri package under pkgs/community. Use when Codex needs community package scaffolding, workspace membership, pyproject metadata, README branding, entry points, second-class citizenship rows, exports, tests, and validation.

not rated 103 24d ago A 59 tokens original Apache-2.0

omk-planning

295

KaimingWan/oh-my-kiro

Skill Claude CodeCodex

Full plan lifecycle: deep understanding → write plan with TDD checklist → parallel review → Ralph Loop execution. Trigger when user says 'plan', 'design', 'implement', 'build', 'architect', '@plan', '@execute', or describes a multi-step task that needs structured breakdown. Also trigger for feature requests, system…

not rated 103 5mo ago A 76 tokens original MIT

triage-issue

296

Teaonly/SKILL.mk

Skill Claude CodeCodex

Triage a bug or issue by exploring the codebase to find root cause, then create a GitHub issue with a TDD-based fix plan. Use when user reports a bug, wants to file an issue, mentions "triage", or wants to investigate and plan a fix for a problem.

not rated 103 4mo ago A 65 tokens original MIT

route-tester

297

blencorp/claude-code-kit

Skill Claude CodeCodex

Framework-agnostic HTTP API route testing patterns, authentication strategies, and integration testing best practices. Supports REST APIs with JWT cookie authentication and other common auth patterns.

not rated 101 9mo ago A 36 tokens original MIT

skillkit

298

rfxlamia/skillkit

Plugin Claude Code

Bundles 3 skills, 3 commands · 236 tokens together

Professional skill creation with TDD workflow. Features dual-mode (fast/full), behavioral validation, and automated quality gates for 9.0/10+ scores.

not rated 101 5mo ago A tokens not measured original Apache-2.0

grafana/agento11y

Skill Claude CodeCodex needs its repo ✓ vendor

Use early in an AI-agent project — before ship, before real traffic — to decide which evaluations to set up and to scaffold a starter experiment. Reads the agent's own code (system prompt, tools, task), recommends specific evaluators with reasons that cite real lines, and writes a labeled draft test suite as an Agent…

not rated 99 today A 0 tokens original Apache-2.0

launchworthy

302

Wunderlandmedia/launchworthy

Plugin Claude Code

Bundles 1 skill · 161 tokens together

Production readiness audit for apps built with AI coding tools. Detects your stack, audits 5 domains, and produces a scored punch list with copy-paste fixes.

not rated 95 7d ago A tokens not measured original MIT

checkup

303

agentvitals/checkup

Skill Claude CodeCodex

Give your AI agent a professional health checkup (AgentVitals). Use when the user asks the agent to run a checkup / test itself / benchmark itself ("run a checkup", "check your vitals", "test yourself", "how stable are you", "/checkup"), or an advanced personality checkup (backbone, proactivity, creativity). 给 AI…

not rated 95 19d ago C 129 tokens AGPL-3.0

pict-test-designer

304

omkamal/pypict-claude-skill

Plugin Claude Code

Bundles 1 skill · 70 tokens together

Design comprehensive test cases using PICT (Pairwise Independent Combinatorial Testing) for any piece of requirements or code. Analyzes inputs, generates PICT models with parameters, values, and constraints for valid scenarios using pairwise testing. Outputs the PICT model, markdown table of test cases, and expected…

not rated 94 +1 5mo ago A tokens not measured

e2e-runner

305

krishnakanthb13/everything-antigravity

Agent Claude Code

End-to-end testing specialist using Vercel Agent Browser (preferred) with Playwright fallback. Use PROACTIVELY for generating, maintaining, and running E2E tests. Manages test journeys, quarantines flaky tests, uploads artifacts (screenshots, videos, traces), and ensures critical user flows work.

not rated 91 +1 6mo ago A 69 tokens

agent-workspace-linux

307

agent-sh/agent-workspace-linux

Skill Codex

Use when a task needs an isolated hidden Linux desktop or workspace-owned browser: GUI app QA, web/browser/shopping automation, sandboxed app observation, or stale workspace cleanup. Routes agent-workspace-linux MCP tools on demand. Does NOT apply to host desktop/Chrome control, generic MCP setup, or pure code/file…

not rated 91 +3 today B 70 tokens copy · 86% MIT

dom-extraction-tester

308

WebMCP-org/npm-packages

Agent Claude Code

Use this agent when you need to test progressive DOM reading implementations by navigating to websites and extracting specific information. This agent works as a driver that receives instructions from a navigator AI about what website to visit and what data to extract, then attempts the extraction and reports back on…

not rated 91 today A 0 tokens original MIT

config

309

ruvnet/sublinear-time-solver

Command Claude Code

Complete configuration guide for pair programming sessions.

not rated 89 1mo ago A 0 tokens original MIT

behat-steps

310

ivangrynenko/cursorrules

Cursor rule Cursor

Cursor rule "behat-steps" from ivangrynenko/cursorrules, covering behat steps - claude memory, available steps, index of generic steps, index of drupal steps and cookietrait.

not rated 88 10mo ago A 15,618 tokens original MIT

add-cucumber-tests

311

Decathlon/tzatziki

Skill Claude CodeCodex

Generates Tzatziki-based Cucumber BDD tests (.feature files) from a functional specification. Use this skill whenever a user wants to write Cucumber tests, add BDD scenarios, create feature files, generate tests, or test application behaviors with Gherkin — especially in Java/Spring projects using Tzatziki step…

not rated 88 yesterday A 130 tokens original Apache-2.0

python-testing

312

Jamkris/everything-gemini-code

Skill Claude CodeCodex

Python testing strategies using pytest, TDD methodology, fixtures, mocking, parametrization, and coverage requirements.

not rated 87 3mo ago A 24 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: