Testing agents

5,338 tagged Testing, measured the same way as everything else here.

Browse within: code-quality 57agent-orchestration 47harness 40spec-driven-development 40agentic-workflow 39Multi-Agent 38playwright 36agentic-coding 32github-copilot 31rtl 31verification 31agentic 29copilot 29context-engineering 27

sw-planner

313

anton-abyzov/specweave

Agent

Part of specweave

Test-Aware Planner for generating tasks.md with BDD test plans. Reads spec.md and plan.md to produce implementation tasks with Given/When/Then scenarios. Use during sw:increment orchestration.

159 3d ago A 44 tokens original MIT

code-reviewer

314

DanielSuo117/velocitai

Agent Claude Code

Part of velocitai

A code-review agent focused on Python and Playwright tests that use the page object model (POM), a way to organize tests by web pages. It looks for bugs and checks rules without changing functionality.

157 3mo ago A 51 tokens

tester

315

joenandez/spectre

Agent

Part of spectre

Write or update behavioral tests, drive a strict RED→GREEN→REFACTOR loop, or diagnose failing tests for a given working set. Use to add coverage, verify a TDD gate, or root-cause a test failure; do not use for non-test implementation (dev), independent code review (reviewer), or scoping/planning. Returns tests…

156 8d ago A 92 tokens original MIT

verifier

316

workos/case

Agent

Fresh-context verification agent for /case. Reads the diff, tests the specific fix with Playwright, creates evidence markers and screenshots. Never implements.

156 1mo ago A 32 tokens

paulpreibisch/AgentVibes

Agent Codex

Act as a document reconstruction specialist. Your purpose is to prove a distillate's completeness by reconstructing the original source documents from the distillate alone.

153 1mo ago A 0 tokens copy · 100% Apache-2.0

red-team-lead

320

irahardianto/awesome-agv

Agent Codex

Delivery validation coordinator. Activated at Tier 2+ only (Tier 1 skips Red Team). Spawned by @overseer after development completes (for structural information isolation — overseer never has development context). Independently verifies the delivered product works correctly by dispatching validators…

153 +2 12d ago A 84 tokens original MIT

e2e-tester

321

DebugBase/glance

Agent

Tests web applications end-to-end using Glance browser MCP. Navigates pages, fills forms, clicks buttons, takes screenshots, runs assertions, and reports bugs. Use when you want to verify an app works correctly — login flows, forms, navigation, responsiveness — with real browser interaction.

151 4mo ago A 63 tokens original MIT

test-runner

322

yezannnnn/agentGroup

Agent

Use this agent when you need to run tests and analyze their results. This agent specializes in executing tests using the optimized test runner script, capturing comprehensive logs, and then performing deep analysis to surface key issues, failures, and actionable insights. The agent should be invoked after code changes…

149 3mo ago A 264 tokens original MIT

e2e-runner

323

cloudnative-co/claude-code-starter-kit

Agent

End-to-end testing specialist for Playwright or equivalent browser tests. Use for critical user journeys, regression checks, and flaky test triage.

147 10d ago A 34 tokens original MIT

mcp-live-tester

325

tacticlaunch/mcp-linear

Agent

Validates mcp-linear changes with local build, test, and optional live Linear smoke checks.

146 +1 5d ago A 25 tokens original MIT

qa-validator

326

systemcrash92/DogSprite

Agent Claude Code

Quality assurance validator. Runs build verification, template gallery export, bundle export, and reports errors. Use after creating new templates or making editor changes.

146 4mo ago A 32 tokens original MIT

pipeline-builder

327

swingerman/engineer

Agent

Part of engineer

Use this agent when generating or updating the acceptance test generator for a project, or when the user asks to "build the pipeline", "generate the test generator", "update the pipeline", "create acceptance test infrastructure", or when the ATDD skill reaches step 3 (pipeline generation). Examples: Context: A spec.md…

146 +2 7d ago A 263 tokens original MIT

spec-guardian

328

swingerman/engineer

Agent

Part of engineer

Use this agent when reviewing GWT acceptance test specs for implementation leakage, or when the user asks to "check specs", "review specs", "audit specs", "clean up specs", or "check for leakage". Also invoked by the /spec-check command and as part of the ATDD workflow. Examples: Context: User has written acceptance…

146 +2 7d ago A 300 tokens original MIT

agent-sdk-verifier-py

330

coleam00/your-claude-engineer

Agent Claude Code

Use this agent to verify that a Python Agent SDK application is properly configured, follows SDK best practices and documentation recommendations, and is ready for deployment or testing. This agent should be invoked after a Python Agent SDK app has been created or modified.

140 +1 7mo ago A 55 tokens copy · 100% MIT

grader

332

idavidov13/agentic-playwright

Agent Claude Code

Evaluate expectations against an execution transcript and outputs.

139 +5 6d ago A 0 tokens copy · 100% MIT

vasu31dev/playwright-ts-template

Agent Claude Code

Use this agent when you need to create automated browser tests using Playwright and vasu-playwright-utils. Examples: Context: User wants to generate a test for the test plan item.

138 5mo ago A 158 tokens

verification

336

AI45Lab/Code

Agent

You are an adversarial verification specialist. Your job is to find evidence, not to reassure.

137 2mo ago A 0 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: