Testing agents

6,768 tagged Testing, measured the same way as everything else here.

Browse within: code-quality 57agent-orchestration 47harness 40spec-driven-development 40agentic-workflow 39Multi-Agent 38playwright 36agentic-coding 32github-copilot 31rtl 31verification 31agentic 29copilot 29context-engineering 27

implementation-agent

145

ThibautBaissac/rails_ai_agents

Agent Claude Code

Orchestrates TDD GREEN phase by implementing minimal code that passes failing tests, coordinating specialist subagents. Use when making tests pass, implementing features from failing specs, or when user mentions green phase or make tests pass. WHEN NOT: Writing tests (use rspec-agent), refactoring code (use…

658 +2 3mo ago A 78 tokens original MIT

callstackincubator/rozenite

Agent

Rozenite is a development tool. Whatever a project wires into its bundler config, a release build must ship none of our code. @rozenite/test-utils provides the bench that proves it, and every plugin owns a Vitest suite in src/tests/release-bundle.test.ts that uses it.

653 yesterday A 0 tokens original MIT

test-runner

147

membrane/api-gateway

Agent Claude Code

Runs tests in the api-gateway Maven reactor — full/module unit runs, isolating a single core test class, or a single distribution/tutorial example test. Use this whenever tests need to be run, checked, or verified after a change, since naive -Dtest/-Dit.test invocations silently run (or skip) the wrong thing in this…

638 2d ago A 82 tokens original Apache-2.0

CLAUDE

148

TesslateAI/OpenSail

Agent

Unit coverage for the multi-agent orchestration skeleton.

636 4d ago A 0 tokens original Apache-2.0

principal-qa-engineer

149

SixHq/Overture

Agent Claude Code

Use this agent when you need comprehensive end-to-end testing of the Overture UI, when a new feature has been added and you need to verify it doesn't break existing functionality, when you need regression testing across the entire application, or when you want absolute certainty that every feature works flawlessly.…

635 5mo ago A 448 tokens original MIT

comparator

150

KdaiP/EchoBot

Agent

Compare two outputs WITHOUT knowing which skill produced them.

628 17d ago A 0 tokens copy · 100% MIT

grader

151

KdaiP/EchoBot

Agent

Evaluate expectations against an execution transcript and outputs.

628 17d ago A 0 tokens copy · 100% MIT

test-runner

152

LedgerHQ/ledger-live

Agent Codex

Test automation expert for ledger-live. Use proactively to run tests and fix failures after code changes. Handles Jest unit and integration tests with MSW mocking for both mobile and desktop environments.

614 2d ago A 40 tokens original MIT

helper

153

enulus/OpenPackage

Agent

Test helper agent from package-b.

613 3mo ago A 8 tokens original Apache-2.0

eunomia-bpf/agentsight

Agent Claude Code

Use this agent when you need to coordinate end-to-end testing across multiple components, optimize build systems, validate deployments, or ensure proper integration between eBPF programs, Rust collector, and frontend components. Examples: Context: User has made changes to both eBPF programs and Rust collector and…

613 8d ago A 239 tokens original MIT

qa-agents

155

moona3k/macparakeet

Agent

Every quarter someone pitches an "AI does QA" tool. Most are web-first or mobile-first. MacParakeet is a menu-bar macOS app with a non-activating KeylessPanel overlay, global dictation hotkeys, and TCC-gated microphone/screen-recording flows. The general AI-QA frontier doesn't speak our shape yet. This doc tracks…

610 4d ago A 0 tokens

test-runner

156

adaline/gateway

Agent Claude Code

Run tests, analyze failures, suggest fixes. Use after code changes.

605 1mo ago A 18 tokens original MIT

nWave-ai/nWave

Agent

Use for DISTILL wave — designs E2E acceptance tests from user stories and architecture using Given-When-Then format. EXPANDED scope (plan v3 §3.A, 2026-05-19) — exclusive test-expertise owner; authors ATs with maximum PBT + parametrize density, runs self-completeness audit (7-category taxonomy + 15-item checklist)…

602 3d ago A 0 tokens original MIT

fork-verifier-agent

158

FradSer/dotclaude

Agent

You are a read-only verification subagent spawned to check a design deliverable the main agent just built or edited. Your only job: load that deliverable, verify it, and report a single verdict — done or needswork — back to the main agent. You must not modify, create, or delete any file, edit the source, build, or…

588 +1 21d ago A 0 tokens copy · 100% MIT

cos-compliance

159

winstonkoh87/Athena-Public

Agent

Use this agent before shipping, merging, or deploying changes. The Compliance Gate validates that all quality gates are met: tests pass, documentation is updated, breaking changes are communicated, and the change is ready for production. Context: User wants to merge a feature branch user: "I think this PR is ready to…

583 2d ago A 180 tokens original MIT

comparator

160

ZMGID/kivio

Agent

Compare two outputs WITHOUT knowing which skill produced them.

570 4d ago A 0 tokens GPL-3.0

grader

162

ECNU-ICALK/AutoSkill

Agent

Evaluate expectations against an execution transcript and outputs.

570 +2 3mo ago A 0 tokens

evaluator

164

kangarooking/kangarooking-skills

Agent

Use this agent to test implementations against sprint contracts and specifications. Uses Playwright MCP for E2E testing, Chrome DevTools for UI inspection, and visual tools for verification. Grades implementations and provides specific failure reports. Trigger when user says "test", "evaluate", "qa", "verify", or…

553 2d ago A 69 tokens

bug-fix

165

defendend/Claude-ast-index-search

Agent Claude Code

Use this agent when a user reports a concrete bug ("X doesn't work", "crashes on Y", "wrong output for Z", a GitHub issue with steps-to-reproduce). The agent reproduces the bug, locates the root cause, applies a minimal fix, proves the fix works with a regression test, and confirms nothing else broke. Do NOT use for…

552 1mo ago A 100 tokens original MIT

implementer

166

xiaolai/vmark

Agent Claude Code

Implements scoped changes with tests and minimal diffs.

544 3d ago A 14 tokens original ISC

manual-test-author

167

xiaolai/vmark

Agent Claude Code

Writes and maintains comprehensive manual testing guides (incremental + final).

544 3d ago A 17 tokens original ISC