Testing

18,481 mods in this category, of every kind an agent can take. Each one carries what it costs per session, what the scan found, and whether it is the original.

opengate-mcp

1441

nickjlamb/opengate

MCP server Claude CodeCodexCursor +2

Check whether an AI answer is grounded in its context — deterministic, no LLM judge. Runs locally from the @pharmatools/opengate-mcp npm package.

not rated 1 8d ago A tokens not measured original MIT

prodpoke-mcp

1442

prodpoke/prodpoke-mcp

MCP server Claude CodeCodexCursor +2

AI QA tester — real browsers scan sites for bugs, SEO, perf, and accessibility issues via chat. Remote server at prodpoke.com.

not rated 1 4mo ago A tokens not measured original MIT

rule-drift

1443

ri7in/rule-drift

MCP server Claude CodeCodexCursor +2

Test whether AI agents retain critical instructions as conversations grow. Runs locally from the rule-drift npm package.

not rated 1 yesterday A tokens not measured original MIT

veris

1445

vighriday/Veris

Skill Claude CodeCodex

Behavioral verification intelligence for autonomous coding agents. Use this skill when the user asks "what could break in this PR?", "which workflows changed?", "what should I test before merging?", "is this change risky?", "what's the blast radius of removing this function?", or any question about understanding the…

not rated 1 5d ago A 116 tokens original MIT

mimiq-mcp

1446

victorgulchenko/mimiq-mcp

MCP server Claude CodeCodexCursor +2

Your agent tests pages, copy, and flows on simulated users while you build. Remote server at mcp.mimiqai.com.

not rated 1 6mo ago A tokens not measured original MIT

vola-trebla/playwright-network-chaos-mcp

MCP server Claude CodeCodexCursor +2

MCP server that gives AI agents dynamic network chaos control over Playwright browser sessions. Runs locally from the playwright-network-chaos-mcp npm package.

not rated 1 3mo ago A tokens not measured original MIT

44-pixels/handover-mcp

Skill Claude CodeCodex

Run a complete Handover MCP continuity test across separate authenticated people or service agents. Use when asked to test an agent handoff end to end, qualify a Handover integration, verify multi-file round-trip, exercise revision-anchored review and correction, or test denied, read-only, stale-revision, and…

not rated 1 1mo ago A 76 tokens original MIT

mcp

1449

gaffer-sh/mcp

MCP server Claude CodeCodexCursor +2

Test analytics for AI agents: test history, flaky tests, failure clusters, coverage. Runs locally from the @gaffer-sh/mcp npm package. Needs 2 environment variables to run.

not rated 1 1mo ago A tokens not measured original MIT

n8n-workflow-tester

1450

Souzix76/n8n-workflow-tester-safe

Skill Claude CodeCodex

Test, score, and inspect n8n workflows via MCP. Safe by design — no credential access, no destructive auto-repair. Use when testing workflows, debugging executions, searching nodes, or building workflows incrementally.

not rated 1 5mo ago A 52 tokens original MIT

xcelium-sim

1451

hslee-cmyk/xcelium-mcp

Skill Claude Code

A guide for using the Xcelium simulator, a tool for testing digital hardware designs, including its commands and routing rules.

not rated 1 1mo ago A 175 tokens

mockhunter

1453

CodeShuX/mockhunter

Plugin Claude Code

Bundles 1 skill · 90 tokens together

Audit live web pages for fake/mock data — 5-phase Playwright + Claude Code skill that classifies every visible value as REAL, MOCK, LLM, HARDCODED, BROKEN, or UNKNOWN.

not rated 1 4mo ago A tokens not measured original MIT

qtest-mcp-server

1454

Usman-Ghani123/qtest-mcp-server

MCP server Claude CodeCodexCursor +2

Unofficial MCP server for qTest Manager — browse modules, fetch test cases, build test execution folders. Runs locally from the qtest-mcp-server npm package.

not rated 1 3mo ago A tokens not measured original MIT

web-test-case-gen

1455

automata-network/agent-skills

Skill Claude Code

Generate persistent test cases from project analysis, or add individual test cases interactively. Supports full project analysis or adding single test cases via prompt description with browser exploration.

not rated 1 8mo ago A 37 tokens original MIT

webapp-testing

1456

rockexe0000/my-awesome-copilot

Skill Claude CodeCodex

Toolkit for interacting with and testing local web applications using Playwright. Supports verifying frontend functionality, debugging UI behavior, capturing browser screenshots, and viewing browser logs.

not rated 1 8mo ago A 35 tokens copy · 100% MIT

test-execution

1457

zavora-ai/skill-test-execution

Skill Claude Code

Orchestrate test execution — run unit, integration, and E2E tests, collect coverage reports, and analyze failures. Use when running tests, checking coverage, debugging test failures, or validating code changes before merge.

not rated 1 3mo ago A 48 tokens

DAA-Master

1459

gigayaya/DAA-Master

Plugin Claude Code

Bundles 5 skills · 149 tokens together

Declarative Action Architecture (DAA) — a strict three-layer separation pattern for E2E automation testing. Provides skills for generating, reviewing, and architecting DAA-compliant test code.

not rated 1 4mo ago A tokens not measured original Apache-2.0

tkolleh/skills

Skill Claude CodeCodex

Formal model-checking layer over Allium specs using Alloy 6 + Electrod + nuXmv. Trigger on: "rigorous structural analysis", "model-check this spec", "verify this Allium spec formally", "check for structural divergence with Alloy", "run the Alloy loop", "does the code actually satisfy this spec", temporal/CTL/LTL…

not rated 1 3d ago A 237 tokens original MIT

ai-workflow

1462

vindm/dotclaude

Skill Claude CodeCodex

Set up LLM workflow discipline for projects that use AI / LLM calls in production or eval suites. Authors an eval-cost-watcher agent that projects token cost BEFORE regression evals run, plus an AI-workflow-discipline rule covering mock-mode placement, fixture freshness, and multi-stage cost accumulation. Optionally a…

not rated 1 4d ago A 88 tokens original MIT

blind-ab-verify

1463

aaronartistzhang-afk/DailyWork

Skill Claude CodeCodex

A workflow for comparing two versions of a prompt, skill, or piece of text without revealing which one is the new version. Separate agents create the versions, and a person scores them without seeing their labels.

not rated 1 14d ago A 305 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: