Testing

18,225 mods in this category, of every kind an agent can take. Each one carries what it costs per session, what the scan found, and whether it is the original.

register-ldm-task

457

yzailab/Large-Discovery-Models

Skill Codex needs its repo

Scaffold, implement, register, scientifically qualify, and production-check an LDM domain task in this repository. Use when adding or repairing a task adapter, task manifest, experiment.json benchmark contract, proposal-provider capabilities, metric roles, qualification evidence, official evaluation budget, campaign…

not rated 30 +2 12d ago A 86 tokens original MIT

archive

459

stn1slv/spec-kit-archive

Command Claude Code needs its repo

Archive a feature specification into main project memory after merge, resolving gaps and conflicts.

not rated 29 +1 13d ago A 15 tokens original MIT

postmancer

461

hijaz/postmancer

MCP server Claude CodeCodexCursor +2

Standalone MCP server for REST API testing and management. Runs locally from the postmancer npm package. Needs 2 environment variables to run.

not rated 28 1y ago A tokens not measured original MIT

frozenlib/mcp-attr

Cursor rule Cursor

A set of Cursor coding rules for testing generated code from a Rust Model Context Protocol server attribute system. Code generation means producing source code automatically from declarations or annotations.

not rated 28 11mo ago A 745 tokens

ue-benchmark

463

blackplume233/UnrealMCPHub

Skill Claude Code

An evaluation framework for testing coding agents against defined benchmarks, or repeatable tests used to compare results.

not rated 28 5mo ago A 56 tokens

qa-use

464

desplega-ai/qa-use

Plugin Claude Code

Bundles 1 skill, 6 commands, 5 agents · 350 tokens together

Simplified CLI-integrated plugin for E2E testing with essential slash commands plus dynamic API CLI workflows (qa-use api). Provides AI-first feature verification, browser automation, and test management. All CLI operations documented in SKILL.md and accessible via qa-use docs for harness compatibility.

not rated 27 3mo ago A tokens not measured original MIT

wp-dev-skills

465

mralaminahamed/wp-dev-skills

Plugin Claude Code

Bundles 19 skills, 1 command · 5,573 tokens together

WordPress plugin development skills for Claude Code — GitHub flow (scoped commits, PR, hotfix, revert), WooCommerce extensions, plugin testing (PHPUnit/Brain\Monkey/redirect harness), coding standards (PHPCS/WPCS), CI/QA triage, PHPStan stubs scaffolding, build tools (@wordpress/scripts), background processing (Action.

not rated 27 1mo ago A tokens not measured original MIT

tdd

466

Vooster-AI/cursorrules-to-claudemd

Cursor rule Cursor

This document defines the REQUIRED process for all code changes. No exceptions without explicit team approval.

not rated 27 1y ago A 644 tokens

Blazemeter/bzm-mcp

Skill Claude CodeCodex

Comprehensive guide for BlazeMeter Performance Testing, including load configuration, reporting, JMeter configuration, Taurus, scenarios, and advanced features. Use when working with Performance tests for (1) Configuring load settings and distribution, (2) Creating and running tests (JMeter, Browser, URL/API…

not rated 26 9d ago A 138 tokens original Apache-2.0

localstack-mcp-server

470

localstack/localstack-mcp-server

MCP server Claude CodeCodexCursor +2

A LocalStack MCP Server providing essential tools for local cloud development & testing. Runs locally from the @localstack/localstack-mcp-server npm package. Needs 11 environment variables to run.

not rated 26 4d ago A tokens not measured original Apache-2.0

e2e-testing

471

hairyf/skills

Skill Claude Code

End-to-end testing patterns with Playwright for full-stack Python/React applications. Use when writing E2E tests for complete user workflows (login, CRUD, navigation), critical path regression tests, or cross-browser validation. Covers test structure, page object model, selector strategy (data-testid > role > label)…

not rated 25 27d ago A 101 tokens fork MIT

flow-dev

473

aitanjp/flow-x

Skill Claude CodeCodex

A task-focused development assistant that completes one task from TASK.md using TDD, a method of writing a failing test, making it pass, and then improving the code.

not rated 25 2mo ago A 131 tokens

vasu-playwright-utils

474

vasu31dev/playwright-ts-lib

Skill Claude Code

Use the vasu-playwright-utils library for Playwright browser automation with simplified action, assertion, locator, element, page, and API utilities.

not rated 25 5mo ago A 35 tokens original MIT

qa-tester

475

Ekioo/KittyClaw

Skill Claude CodeCodex

Verifies programmer deliveries when a ticket reaches Review. Actually runs the application/tests/endpoints to confirm the change works, sets up missing test tooling when needed, and blocks the ticket if execution is impossible. Posts a PASS/FAIL/BLOCKED report; on FAIL, returns the ticket to Todo.

not rated 25 +1 yesterday A 63 tokens AGPL-3.0

kwb

476

punt-labs/prfaq

Agent Claude Code

You are inspired by Kent Beck — creator of Extreme Programming and Test-Driven Development, co-author of JUnit, and author of Smalltalk Best Practice Patterns (1997), Test-Driven Development: By Example (2002), and Implementation Patterns (2007).

not rated 25 6d ago A 62 tokens original MIT

specflow

477

Hulupeep/Specflow

Skill Claude CodeCodex

Spec-driven development with executable contracts.

not rated 25 +1 1mo ago A 9 tokens original MIT

lead-eng

479

DUBSOpenHub/dark-factory

Agent Claude Code

Senior engineer that implements features per the architecture and writes open tests.

not rated 25 +2 1mo ago A 16 tokens original MIT

run-aeon-benchmark

480

AEON-7/Aeon-Bench-Pod

Skill Claude CodeCodex

Use when asked to run, benchmark, evaluate, or score an LLM with AEON Bench. You run the AEON Bench Pod on the user's machine, point it at a model, run the benchmark, and submit the signed result to the public leaderboard at aeon-bench.com. All work happens on the pod. The mothership only shows the board and accepts…

not rated 25 21d ago A 83 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: