Testing plugins

784 tagged Testing, measured the same way as everything else here.

Browse within: code-quality 20claude-plugin 17claude-code-marketplace 15claude-code-skills 15browser-automation 14claude-ai 14claude-code-plugins 14claude-code-skill 14antigravity 12accessibility 11agentic-coding 10playwright 10plugins 10ai-testing 9

pythonica

409

Jartan-LLC/grimoire

Plugin Claude Code

Bundles 17 skills · 379 tokens together

Comprehensive Python development -- patterns, testing, async, error handling, packaging, configuration, type safety, resilience, observability, Pydantic, and more.

2 15d ago A tokens not measured original MIT

agent-reliability

410

ByteStack-Labs/claude-plugins

Plugin Claude Code

Bundles 4 skills · 933 tokens together

Claude skills for AI agent and ML reliability: reproduce the eval-to-production gap, catch confidently-wrong outputs, and prove root cause with verified numbers. Start with production-autopsy.

2 2mo ago A tokens not measured original MIT

cantrips

411

toverux/cantrips

Plugin Claude Code

Bundles 25 skills, 1 agent · 847 tokens together

The core engineering loop for coding agents (Claude Code, Codex CLI): grill, spec, tickets, implement with TDD at agreed seams, review, commit, plus a user-gated compound step that turns session learnings into durable project memory. Basic spells a caster always has prepared.

2 3d ago A tokens not measured original MIT

pentest-toolkit

412

nibzard/skills

Plugin Claude Code

Bundles 1 skill · 40 tokens together

AI-Powered Security Testing Toolkit - Comprehensive penetration testing tools for authorized security assessments. Use when conducting professional security testing, vulnerability assessments, or penetration testing on systems you own or have explicit authorization to test.

2 23d ago A tokens not measured original MIT

audit

413

edloidas/skills

Plugin Claude Code

Bundles 9 skills · 512 tokens together

CI, script, security, skill, workspace, Three.js, React, and test-suite auditing skills.

2 3d ago A tokens not measured original MIT

proofrag

414

unshDee/proofrag

Plugin Claude Code

Bundles 1 skill, 1 command · 117 tokens together

Evaluate a RAG/LLM app: generate a golden set from your docs, run LLM-as-judge + retrieval metrics, and produce a shareable HTML scorecard with a CI gate.

2 23d ago A tokens not measured original MIT

boardless-pcb

415

raisoninme/boardless-pcb

Plugin Claude Code

Bundles 10 skills, 1 agent · 520 tokens together

A seven-step, user-invoked workflow for zero-hardware, simulation-only PCB + firmware development: 7 step skills (one step per session) + status + methodology skill + a setup skill for Codex/Cursor, an amnesia-verifier subagent, and an anti-downstream-patching hook. One skill set runs on Claude Code, OpenAI Codex and…

2 20d ago A tokens not measured original MIT

skillme

416

shennawardana23/skillme

Plugin Claude Code

Bundles 55 skills, 2 commands · 5,010 tokens together

A testable knowledge base of real engineering practice — every skill backed by its own deterministic eval suite, checked automatically with smeval, this repository's own eval runner.

2 5d ago A tokens not measured original Apache-2.0

pre-flight-check

417

mirekondro/The-Pre-Flight-Check

Plugin Claude Code

Bundles 1 skill · 45 tokens together

Fail-fast quality gate (Typecheck → Lint → Test → Security Audit) that runs before any task is declared done or code is committed. Auto-detects Node.js and Python projects.

2 2mo ago A tokens not measured original MIT

flow-tester

418

kushagraagent47/vibe-tester

Plugin Claude Code

Bundles 1 skill · 114 tokens together

Discovers a web app's user flows, drives a real Chromium browser through them with a live local dashboard, flags functional/content/visual/console bugs, and runs a read-only security audit on local source. Identify-only, never fixes.

2 2mo ago A tokens not measured

ai-stack

419

nazmulnahid-git/Ai-Stack

Plugin Claude Code

Bundles 4 skills · 361 tokens together

Useful AI things an engineer uses daily. Ships prodcheck (senior-engineer production review), prodtest (senior-QA test pass with Playwright), mergemain (context-aware merge of main into the current branch), and shipit (clean commits and PRs, with or without AI attribution).

2 26d ago A tokens not measured original MIT

auto-pilot

420

IronRookieCoder/auto-pilot

Plugin Claude Code

Bundles 6 skills, 2 hooks · 445 tokens together

Long-cycle coding workflow plugin - Structured state machine with TDD-driven milestone execution and deterministic gates for sustained AI coding sessions.

2 4mo ago A tokens not measured

gut-skill

424

rockerBOO/gut-skill

Plugin Claude Code

Bundles 1 skill · 50 tokens together

Claude Code skill for the GUT (Godot Unit Testing) framework — test structure, assertions, doubles, async helpers, memory management, and CLI usage for GDScript.

2 4mo ago A tokens not measured GPL-3.0

qa-explore

425

victoraguilarsantamariadev/qa-explore

Plugin Claude Code

Bundles 6 skills · 841 tokens together

A team of AI agents that test any web app like human QA: explore the live app, screenshot and visually judge rendering + data, capture trace/HAR/console/video evidence, adversarially verify findings, learn from rejected ones, codify them into a self-growing Playwright/Cypress suite — and close the loop by filing…

2 1mo ago A tokens not measured original MIT

slopstop marketplace

426

iansmith/slopstop

Plugin Claude Code

Lists 1 plugin

Marketplace hosting the slopstop plugin: ticket-anchored AI development for Linear, JIRA, and GitHub Issues — TDD-first planning, scope control, adversarial review, and multi-agent orchestration.

2 3d ago A tokens not measured

claim-check

428

bhumik154/claim-check

Plugin Claude Code

Bundles 1 skill, 3 hooks · 85 tokens together

Verifies test-count claims against what your suite actually did, before a commit lands. Supports pytest, vitest and jest. If the message makes no claim, it does nothing.

2 6d ago A tokens not measured original MIT

swe-workbench

429

lugassawan/swe-workbench

Plugin Claude Code

Bundles 1 skill, 24 commands, 32 agents, 4 hooks · 2,193 tokens together

Senior-engineer toolkit: principled design (Clean Arch, DDD, SOLID, TDD, patterns, observability, concurrency), security review, language expertise (Bash, C#, Dart, Go, Java, Kotlin, Python, Ruby, Rust, SQL, Swift, TypeScript), and pragmatic workflows.

2 3d ago A tokens not measured original MIT

user-test

431

cosmos-makers/user-test

Plugin Claude Code

Bundles 4 skills · 148 tokens together

Multi-agent User Testing — spawn persona agents to test your spec, service, or API and compile actionable feedback.

2 5mo ago A tokens not measured

yoke

432

HECer/yoke

Plugin Claude Code

Bundles 34 skills, 1 hook · 1,585 tokens together

Cross-agent coding harness: one curated skill canon (TDD, brainstorming, plans, reviews, shipping, design verification) plus mechanical safety gates and an autonomous loop via the yoke CLI.

2 12d ago A tokens not measured original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: