Testing agents

5,604 tagged Testing, measured the same way as everything else here.

Browse within: code-quality 57agent-orchestration 47harness 40spec-driven-development 40agentic-workflow 39Multi-Agent 38playwright 36agentic-coding 32github-copilot 31rtl 31verification 31agentic 29copilot 29context-engineering 27

e2e-runner

169

vibeeval/vibecosystem

Agent

End-to-end testing specialist using Vercel Agent Browser (preferred) with Playwright fallback. Use PROACTIVELY for generating, maintaining, and running E2E tests. Manages test journeys, quarantines flaky tests, uploads artifacts (screenshots, videos, traces), and ensures critical user flows work.

530 24d ago A 69 tokens original MIT

fixture-reviewer

170

phel-lang/phel-lang

Agent

Audits .test integration fixtures under tests/php/Integration/Fixtures for drift against the current compiler output. Use after lexer, parser, analyzer, or emitter changes.

523 9d ago A 37 tokens original MIT

quality-agent

171

vanzan01/claude-code-sub-agent-collective

Agent

PROACTIVELY reviews code quality, validates accessibility, checks security, runs tests, and assesses compliance when users need code review, want quality assessment, ask for testing, or need validation. Use for any quality assurance needs.

521 4mo ago A 48 tokens original MIT

qa-lead-tester

172

openstory-so/openstory

Agent Claude Code

Use this agent when you need comprehensive quality assurance oversight, test strategy development, or code review from a testing perspective. This includes: designing test suites for new features, reviewing pull requests for test coverage, creating e2e tests with Playwright, designing unit tests with Vitest…

519 3d ago A 386 tokens original MIT

test-writer

173

meleantonio/ChernyCode

Agent Cursor

Write comprehensive tests for code changes. Use proactively when implementing features, fixing bugs, or when code coverage is needed.

516 +1 9d ago A 27 tokens

tester

174

ansys/pymapdl

Agent

Expert in Python testing, test coverage, mocking strategies, and PyMAPDL test infrastructure. Use for writing tests, improving coverage, fixing flaky tests, and reviewing test quality.

513 3d ago A 37 tokens original MIT

istvan-ujjmeszaros/touchspin

Agent Claude Code

Use this agent when the user needs to write, debug, fix, or improve coverage for Playwright end-to-end tests. This includes creating new test specifications, implementing test scenarios from Gherkin comments, debugging test failures, fixing timeout errors, resolving locator issues, improving code coverage by targeting…

501 5mo ago A 0 tokens

tester

176

iusztinpaul/designing-real-world-ai-agents-workshop

Agent Codex

Tester role definition for /implement-universal. Loaded by the orchestrator at the start of the Tester phase, on logic tickets only. Trusts the SWE phase's happy-path e2e excerpt (does NOT re-run the Make target), runs at most 1 adversarial break path, walks every Acceptance Criterion with concrete evidence, and emits…

501 +1 3mo ago A 99 tokens copy · 89% MIT

comparator

179

Playa-0v0/Cyrene-Agent

Agent

Compare two outputs WITHOUT knowing which skill produced them.

487 3d ago A 0 tokens copy · 100% MIT

propagate

181

juxt/allium

Agent

Generate tests from Allium specifications. Use when the user wants to propagate tests, generate test files from a spec, write tests for a specification, create property-based tests, produce state machine tests, check test coverage against spec obligations, or understand what tests a specification requires.

478 +1 6d ago A 57 tokens original MIT

session-agent

182

0xnyn/canary

Agent

Record a verifiable Canary QA session — explore a flow step by step against one persistent browser, each script a recorded step capturing trace/video/HAR/console, then render report.html. Use when the user wants to verify or QA a flow, capture a trace or video, or produce a shareable report of a browser run.

476 +1 2mo ago A 69 tokens

verify-agent

183

0xnyn/canary

Agent

Turn a code change into a prioritized browser-QA plan with Canary — read the git diff, infer the affected user-facing workflows, and suggest concrete flows and the checks that must hold, then optionally record them as a session with a report. Use when the user asks what to test for a change, wants to QA a…

476 +1 2mo ago A 81 tokens

saltbo/agent-kanban

Agent Claude Code

Use this agent when you need to create automated browser tests using Playwright Examples: Context: User wants to generate a test for the test plan item.

458 5d ago A 151 tokens

1c-tester

187

comol/ai_rules_1c

Agent

Expert 1C testing agent. Tests code and functions using web browser automation and the /deploy-and-test command. Deploys configuration to test infobase, performs UI testing with human-like interactions, validates functionality. Use when the user asks to run deployment, UI testing, or verification against a test…

448 +4 2d ago A 69 tokens

qa-reviewer

188

platformplatform/PlatformPlatform

Agent Claude Code

QA code reviewer who validates Playwright E2E test implementations against project rules and patterns. Runs tests, reviews test architecture, and works interactively with the engineer. Never modifies code.

438 1mo ago B 41 tokens original MIT

implement

189

Episkey-G/GrokSearch-rs

Agent

Code implementation expert for the Trellis channel runtime. Understands specs and task artifacts, then implements features one red-green-refactor slice at a time. No git commit allowed.

436 6d ago A 37 tokens original MIT

Nexus-Router/nexus

Agent Claude Code

name: integration-test-engineer description: Use this agent when you need to create, modify, or debug integration tests in the crates/integration-tests directory. This includes writing new test scenarios, updating existing tests, working with Docker Compose configurations for test environments, handling authentication…

435 5mo ago A 274 tokens original MPL-2.0

jabrena/plinth

Agent

Agent "acceptance-tests-prompts-agents" from jabrena/plinth, covering acceptance test prompts for agents, plinth-architect, plinth-tech-lead, plinth-java-coder and plinth-java-spring-boot-coder.

435 +1 2d ago A 0 tokens original Apache-2.0