Testing plugins

1,296 tagged Testing, measured the same way as everything else here.

Browse within: claude-code-plugin 205claude-code-skills 51commands 38plugins 35ai-coding 31claude-plugin 30github-copilot 30agentic-coding 27claude-code-marketplace 25claude-code-plugins 24claude-ai 22claude-code-skill 22code-quality 21ai-skills 20

dev-workflow

481

ziduzichong/dev-workflow

Plugin Claude Code

Bundles 1 skill · 92 tokens together

7-phase full-stack development workflow with TDD, systematic debugging, code review, and documentation generation. Supports C/C++/Python/TypeScript/Go/Rust/Java and embedded STM32 development.

not rated 2 4mo ago A tokens not measured original MIT

dev-helpers

482

refo/claude-dev-helpers

Plugin Claude Code

Bundles 13 commands, 1 hook · 269 tokens together

Stack-agnostic agentic dev workflow helpers: TDD, commits, PRDs, issue triage, pre-commit setup, design grilling. Reads per-project config from .claude/helpers.json.

not rated 2 2mo ago A tokens not measured original MIT

aws-test-plugin

483

whitewhiteqq/aws-test-plugin

Plugin Claude Code

Bundles 6 skills, 1 agent · 671 tokens together

AI-powered test generation skills for AWS Python projects — Lambda, API Gateway, Step Functions, and Batch. Generates E2E, integration, contract, performance, and load tests by reading actual handler code.

not rated 2 2mo ago A tokens not measured original Apache-2.0

codelore

485

qa-vault/codelore

Plugin Claude Code

Bundles 4 skills · 777 tokens together

Skills for critically exploring and documenting code. exploratory-qa surfaces non-obvious decisions in a feature; document-feature captures how and why a feature works for future maintainers; consulting-project-docs routes existing docs into AI sessions; migrate-project-docs prepares pre-existing docs for that router.

not rated 2 changed today A tokens not measured original MIT

kahea

487

copyleftdev/kahea

Plugin Claude Code

Bundles 1 skill, 1 MCP server · 102 tokens together

Deterministic HTTP and finite WebSocket planning, testing, and policy-gated execution for AI agents.

not rated 2 7d ago A tokens not measured original Apache-2.0

regressguard

490

Bharath-code/regressGuard

Plugin Claude Code

Bundles 1 MCP server

Agent-native API regression verification. Snapshot a known-good baseline, then let the agent check its own edits for broken contracts (removed fields, status changes, failing tests) before reporting done. Deterministic — no LLM judging, no cloud.

not rated 1 1mo ago A tokens not measured original MIT

evals-mcp-server

491

cyanheads/evals-mcp-server

Plugin Claude Code

Bundles 32 skills, 1 MCP server · 2,220 tokens together

Author verifiable eval records through a draft → review → revise → submit loop with server-enforced graders; compile to JSONL/CSV/Inspect/lm-eval via MCP. STDIO or Streamable HTTP.

not rated 1 12d ago A tokens not measured original Apache-2.0

roslyn-mcp

492

darylmcd/Roslyn-Backed-MCP

Plugin Claude Code

Bundles 44 skills, 5 agents, 2 hooks, 2 MCP servers · 4,090 tokens together

Semantic C# analysis and refactoring for AI agents, powered by Roslyn. Load .sln/.csproj files, inspect the live MCP catalog at runtime, and use 32 bundled agent skills for analysis, refactoring, testing, architecture review, and release workflows.

not rated 1 changed today A tokens not measured original MIT

awt marketplace

493

ksgisang/awt-skill

Plugin Claude Code

Self-healing DevQA loop: scan websites, generate YAML test scenarios with AI, execute with Playwright, auto-fix failures. Supports 4 AI providers, vision matching, pattern learning.

not rated 1 3mo ago A tokens not measured AGPL-3.0

mockhunter

494

CodeShuX/mockhunter

Plugin Claude Code

Bundles 1 skill · 90 tokens together

Audit live web pages for fake/mock data — 5-phase Playwright + Claude Code skill that classifies every visible value as REAL, MOCK, LLM, HARDCODED, BROKEN, or UNKNOWN.

not rated 1 3mo ago A tokens not measured original MIT

DAA-Master

496

gigayaya/DAA-Master

Plugin Claude Code

Bundles 5 skills · 149 tokens together

Declarative Action Architecture (DAA) — a strict three-layer separation pattern for E2E automation testing. Provides skills for generating, reviewing, and architecting DAA-compliant test code.

not rated 1 3mo ago A tokens not measured original Apache-2.0

dotclaude

498

vindm/dotclaude

Plugin Claude Code

Bundles 7 skills, 5 agents, 5 hooks · 1,038 tokens together

Code review, pre-flight, worktree isolation, guard hooks, and risk-weighted test coverage for Claude Code. Consumed as-is; /dotclaude:coding and /dotclaude:testing elicit your project's bar.

not rated 1 changed today A tokens not measured original MIT

rust-intel

499

PHPCraftdream/rust-intel

Plugin Claude Code

Bundles 2 skills · 246 tokens together

Defense against LLM Rust failure modes: 59 categories of bugs that survive rustc, clippy, and cargo test — async cancellation, unsafe/FFI, concurrency, crypto, supply-chain, tests-that-pass-by-luck, semantic conformance, and systemic performance cost. Ships the rust-intel skill plus /rust-intel:audit, /rust-intel:fix.

not rated 1 changed today A tokens not measured original Apache-2.0

lsa

500

NVZver/claude-marketplace

Plugin Claude Code

Bundles 7 skills, 1 agent, 1 hook · 333 tokens together

Living Spec Architecture — a technology-agnostic spec layer that authors a grounded spec and verifies it before and after an external implementer builds it. LSA is NOT the implementer: any coding agent (Claude Code, Cursor) or human writes the code. Skills: discover (extract intent + gather codebase facts; the…

not rated 1 9d ago A tokens not measured original MIT

sll

501

MohamedEmbarak/SLL

Plugin Claude Code

Bundles 2 hooks

Three hooks, no agents. A stated test result must reproduce, a fabricated import never reaches disk, and every command's real output is recorded. Active in every project the moment it is installed.

not rated 1 18d ago A tokens not measured original MIT

ab-harness

502

wan-huiyan/claude-ecosystem-hygiene

Plugin Claude Code

Bundles 1 skill · 199 tokens together

Counterfactual A/B + layered-ablation harness for Claude Code setup. Measures whether your /.claude/ stack (skills, lessons, axioms, memory) plus in-repo discipline (docs/runbooks, decisions, findings) actually helps on real tasks. Binary A/B (setup-ON vs setup-OFF) and layered ablation (strip one layer at a time, ran.

not rated 1 18d ago A tokens not measured original MIT

wan-huiyan/claude-ecosystem-hygiene

Plugin Claude Code

Bundles 1 skill · 162 tokens together

A check reported clean without examining the case it exists for — "green" meant "nothing ran". Use when about to trust a gate, test or CI step you have never watched fail; when a guard is green but the bug it guards shipped anyway; when a check passes on a shallow clone, a disabled object, a malformed input or an…

not rated 1 18d ago A tokens not measured original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: