Testing

18,324 mods in this category, of every kind an agent can take. Each one carries what it costs per session, what the scan found, and whether it is the original.

preflight

649

Nancy-Chauhan/preflight

Plugin Claude Code

Bundles 1 skill · 98 tokens together

Know exactly where your system breaks before you ship. Scans your codebase, traces every service call, and simulates what happens at scale - finding bottlenecks, cost cliffs, rate limits, and breaking points.

not rated 13 6mo ago A tokens not measured original MIT

playwright-cli

650

kissgyorgy/coding-agents

Skill Claude Code

Automates browser interactions for web testing, form filling, screenshots, and data extraction. Use when the user needs to navigate websites, interact with web pages, fill forms, take screenshots, test web applications, or extract information from web pages.

not rated 14 today A 52 tokens

Cheryl-station/analyze-change-test-scope

Skill Codex

Prepare safe GitHub PR workspaces, scan all supported source files in a complete repository, build a lightweight caller-to-callee graph, analyze frontend and backend Git changes with symbol and API references, infer direct and transitive business impact, and produce a prioritized, evidence-backed test scope. Use when…

not rated 13 +1 1mo ago A 147 tokens original Apache-2.0

testing

652

akngs/s4

Cursor rule Cursor

When writing tests, always follow these best practices.

not rated 13 8mo ago A 0 tokens original MIT

issue-workflow

653

vchelaru/XnaFiddle

Cursor rule Cursor

GitHub issue workflow and mandatory manual-test handoff for UI/behavior changes.

not rated 13 3d ago A 399 tokens original MIT

NoesisVision/nasde-toolkit

Skill Claude CodeCodex

Run coding agent benchmarks and verify results with nasde. Use this skill when the user wants to: Run a benchmark (all tasks, single task, specific variant) Re-run assessment evaluation on existing trial results Check or verify results in Opik (traces, feedback scores, experiments) Troubleshoot a failed benchmark run…

not rated 12 today A SkillSpector: warn 139 tokens original MIT

antrieb

656

jade-pico/antrieb-mcp-server

MCP server Claude CodeCodexCursor +2

Validates AI infra code on real VMs. Self-corrects until it works. No containers, no sandboxes. Remote server at antrieb.sh.

not rated 12 4mo ago A tokens not measured original Apache-2.0

verify-cli

657

NimbleBrainInc/mpak

Skill Claude Code

Run smoke tests against the mpak CLI to verify all commands work correctly before publishing a new release.

not rated 12 1mo ago A 23 tokens

plan-validator

659

sonomirco/agents-and-commands

Agent Claude Code

Use this agent when you have created a plan (e.g., implementation plan, architecture design, refactoring strategy, feature specification) and need to validate and iteratively improve it before execution. This agent should be invoked:\n\n- After drafting any significant technical plan that will guide implementation…

not rated 12 6mo ago A 0 tokens original Apache-2.0

quality-check-command

660

Kazuki-tam/next-stage

Command Claude CodeCursor

This command performs comprehensive code quality checks. Use it before commits or when implementation is complete.

not rated 12 1mo ago A 0 tokens original MIT

ai_tdd_workflow

661

LuthienResearch/luthien_control

Cursor rule Cursor

This rule outlines the strict Test-Driven Development process to be followed by the AI assistant when implementing new features or fixing bugs. This complements the broader developmentworkflow by providing specific TDD execution steps for the AI.

not rated 12 9mo ago A 0 tokens

captain-obvious

662

shmulc8/captain-obvious

Plugin Claude Code

Bundles 1 skill, 1 hook · 203 tokens together

Deterministic scanner that finds and deletes tests that can never fail — assertions the type checker already guarantees, tautologies, mock-echo tests, dead/swallowed assertions, and duplicates. TypeScript (Jest/Vitest/bun:test) and Python (pytest + mypy).

not rated 12 23d ago A tokens not measured original MIT

harmonyos-dev-mcp

663

Deslord319/harmonyos-dev-mcp

MCP server Claude CodeCodexCursor +2

HarmonyOS MCP service for device automation, app deployment, UI interaction, E2E support, and log validation. Runs locally from the harmonyos-dev-mcp Python package.

not rated 12 1mo ago A tokens not measured

bagisto-theme-testing

664

bagisto/agent-skills

Skill Codex

Audit and prove Bagisto storefront themes with source inspection, ownership mapping, admin-to-storefront mutation tests, and Playwright commerce journeys. Use when checking that visible content is dynamic and merchant-controlled; validating theme customizations, channels, CMS, categories, products, search, filters…

not rated 12 today A 97 tokens

cypress-mcp

665

yashpreetbathla/cypress-mcp

MCP server Claude CodeCodexCursor +2

MCP server for AI-driven Cypress test execution. Run, debug, and iterate on E2E tests directly from your AI agent. Runs locally from the cypress-mcp npm package.

not rated 12 6mo ago A tokens not measured

debug-e2e-workflow

666

djscheuf/agentic-dev-ecosystem-template

Skill Claude Code needs its repo

Complete E2E test debugging workflow (composite orchestrator). Starts by reviewing the provided test failure evidence, then forms hypotheses, applies fixes, and verifies results for a presumed E2E playwright test suite.

not rated 12 today A SkillSpector: pass 49 tokens original MIT

swagger-mcp

667

amrsa1/swagger-mcp

MCP server Claude CodeCodexCursor +2

MCP Server for Swagger/OpenAPI documentation and API testing. Runs locally from the swagger-mcp npm package.

not rated 12 1y ago A tokens not measured copy · 100% MIT

add-test

668

open-metadata/ai-sdk

Skill Claude Code

Use when adding unit or integration tests. Provides test patterns, naming conventions, and fixtures for Python (pytest), TypeScript (vitest), Java (JUnit/Mockito), and Rust.

not rated 12 +1 yesterday A 40 tokens

e2e-test-conventions

669

agentmantis/test-skills

Skill Claude Code

Core conventions and rules for Playwright E2E testing with TypeScript. Covers project structure, naming conventions, selector strategy, authentication, navigation, environment configuration, test independence, and parallelism. Automatically loaded when writing or modifying E2E tests. Use when: generating E2E tests…

not rated 12 5mo ago A 84 tokens original MIT

gedd-chat

670

aws-samples/sample-GEDD

Command Claude Code ✓ vendor

You are a GEDD coaching assistant. You guide the user through building a golden evaluation dataset for their AI agent using Open Coding methodology, then help them evaluate and annotate responses — all without leaving Claude Code.

not rated 12 1mo ago A 0 tokens MIT-0

full-development

671

Deepank308/hermes-swe-agent

Skill Claude CodeCodex

End-to-end workflow for feature requests, enhancements, refactors, and tasks. Covers planning, TDD implementation, verification, integration testing, and PR creation.

not rated 12 +1 5mo ago A 35 tokens original MIT

relentless-tester

672

readysettech/rdst

Skill Claude Code needs its repo

Autonomous QA tester that systematically tests every rdst command via tmux harness, applies a quality rubric, and files bugs in beads.

not rated 12 +1 changed 3d ago A SkillSpector: warn 32 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: