Testing

18,443 mods in this category, of every kind an agent can take. Each one carries what it costs per session, what the scan found, and whether it is the original.

10XTeams/trio

Agent Claude Code

Use this subagent to execute E2E test cases via Playwright MCP. It performs login (if required), runs each test case's steps in a real browser, captures snapshots/screenshots, judges pass/fail against expected results, and returns a structured result per test case (including bug details for failures). Invoke it from…

not rated 5 4mo ago A 113 tokens

postmcp

989

Bencibr/postmcp

MCP server Claude CodeCodexCursor +2

MCP server for AI-driven API testing — HTTP, GraphQL, WebSocket, OAuth2, assertions, chained test suites, and SQLite-persisted history with JSON diff comparison. Runs locally from the @bencibro/postmcp npm package.

not rated 5 1mo ago A tokens not measured original MIT

semantix-ai

990

labrat-akhona/semantix-ai

MCP server Claude CodeCodexCursor +2

A Semantic Type System for AI outputs — validate intent, not just shape. Runs locally from the semantix-ai Python package.

not rated 5 1mo ago A tokens not measured original MIT

sentinel

992

Nelsonochoam/nelsonochoam-claude-plugins

Plugin Claude Code

Bundles 1 skill · 0 tokens together

QA artifact generator — point it at a PR or give freeform instructions. Navigates the UI, captures screenshots, and stitches a GIF. Fully configurable auth and artifact storage via /.sentinel/config.json.

not rated 5 3mo ago A tokens not measured

blacksmith

994

grahamnotgrant/blacksmith-mcp

MCP server Claude CodeCodexCursor +2

MCP server for Blacksmith CI - query runs, analyze test failures, detect flaky tests. Runs locally from the blacksmith-mcp npm package. Needs 2 environment variables to run.

not rated 5 4mo ago A tokens not measured original MIT

zendesk-local

995

fruggr/zendesk-mcp-server

MCP server Claude CodeCodexCursor +2

MCP server "zendesk-local" as configured in fruggr/zendesk-mcp-server. Launched with pnpm exec tsx src/index.ts --mode all --dev. Needs 1 environment variable to run.

not rated 5 +1 today A tokens not measured original MIT

skill-tester

996

topprismdata/skill-tester

Skill Claude CodeCodex

Tests and evaluates any Claude Code skill for structural validity, quality, and trigger accuracy. Implements the cc-plugin-eval 4-stage pipeline (Analysis → Generation → Execution → Evaluation) and the 4D scoring rubric (Documentation/Code/Completeness/Usability 25% each). Use before packaging or deploying any skill.

not rated 5 +1 14d ago A 70 tokens

aiagentminder

997

lwalden/AIAgentMinder

Plugin Claude Code

Bundles 15 skills, 16 agents, 5 hooks · 624 tokens together

An opinionated governance layer for Claude Code, built for solo developers: autonomous sprint execution in isolated git worktrees, mandatory TDD, quality gates, pre-PR code review, and structured planning — workflow enforced by hooks, not just rules.

not rated 5 +1 1mo ago A tokens not measured original MIT

case-design

998

xiaozhi86/qamaster

Cursor rule Cursor

A test-case design workflow that turns requirements or prototypes into test cases in Markdown or Excel, with coverage analysis.

not rated 5 4d ago A 34 tokens original MIT

test-in-serum

999

Celian-mrc/serum-mcp

Skill Claude Code needs its repo

Run the pre-real-Serum-test checklist on one or more .SerumPreset files -- automated CBOR wire-type scan plus a human-readable summary -- then hand off to the user for the real load-it-in-Serum test. Use this after any generatepreset/editpreset call on serum-mcp, or whenever a preset file needs to be verified before…

not rated 5 +2 10d ago A 108 tokens original MIT

harnessay

1000

nks0614/harnessay

Plugin Claude Code

Bundles 1 skill · 58 tokens together

Profile your Claude Code transcripts: context budget report, skill-promotion candidates, and a skill regression test harness. Local-only, stdlib-only.

not rated 5 17d ago A tokens not measured original MIT

test-automation

1002

drvoss/harness-100-copilot

Skill Claude CodeCodex

Use when you need to generate a complete test suite for a project — dispatches test-strategist, unit-test-writer, integration-test-writer, e2e-test-writer, and test-reviewer in sequence to produce strategy, implementation, and review artifacts. Covers test pyramid design, unit/integration/E2E implementation, coverage…

not rated 5 3mo ago A 128 tokens

mcp-harness

1003

sabbour/agentweaver

Skill Claude CodeCodex needs its repo

Use this harness to validate Agentweaver's MCP protocol surface, capture the complete tools/call request/response evidence, and emit a normalized agentweaver.persona-judge-verdict/v1 JSON verdict. It is for MCP end-to-end validation, MCP tool-contract regression checks, and investigation of an MCP-reported issue; use…

not rated 5 changed 7d ago A 0 tokens original MIT

zebrunner-test-impact

1004

maksimsarychau/mcp-zebrunner

Skill Claude CodeCodex

Analyze PR or local code changes and find Zebrunner test cases to run for regressions and new coverage gaps. Use when the user mentions test impact, PR test planning, which tests to run, regression coverage, sprint PR rollups, or Zebrunner + pull request / code changes.

not rated 5 today A 63 tokens AGPL-3.0

anti-confabulation

1005

0ryant/engineering-doctrine

Skill Claude CodeCodex

Primes an artefact-producing agent to separate what it intended, what it materialised, what it re-checked, and what remains unverified before it reports completion. Load for build-class tasks (code generation, evidence-pack emission, tool-wrapper generation, implementation work whose claims can be re-tested against…

not rated 5 5d ago A 68 tokens original Apache-2.0

yellow-sheet-corpus

1006

shawnclybor/clybor-claude-tooling

Skill Claude CodeCodex

Run the Yellow Sheet chain end to end against ANY matter, from the original PDFs, and read the result honestly. Use whenever someone asks what has been tested end to end, whether a matter produces a usable sheet, how to onboard a NEW matter, or asks to re-run after a code change. Carries the intake mechanism that…

not rated 5 7d ago A 173 tokens

flowproof-config

1007

automators-com/flowproof

Skill Claude CodeCodex

Configure flowproof's SAP GUI, Fiori, and AI authoring credentials by walking the user through flowproof config sap / flowproof config fiori / flowproof config ai. Use when the user wants to set up, change, or check their SAP/Fiori login or model authoring key, or when a flow run fails because SAPUSER, FIORIPASSWORD…

not rated 5 yesterday A 106 tokens original Apache-2.0

prestashop-pr-qa

1008

PrestaShop/skills

Skill Claude CodeCodex

QAs a PrestaShop pull request against an environment that is already running, in a real browser, on the command line or over HTTP, and writes an HTML report stating whether it is approved, with the recording as proof. Use when the user says "QA this PR", "test this pull request", "check that this fix works"…

not rated 5 4d ago A 90 tokens AFL-3.0

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: