sw-planner
313Agent
Part of specweave
Test-Aware Planner for generating tasks.md with BDD test plans. Reads spec.md and plan.md to produce implementation tasks with Given/When/Then scenarios. Use during sw:increment orchestration.
5,338 tagged Testing, measured the same way as everything else here.
Browse within: code-quality 57agent-orchestration 47harness 40spec-driven-development 40agentic-workflow 39Multi-Agent 38playwright 36agentic-coding 32github-copilot 31rtl 31verification 31agentic 29copilot 29context-engineering 27
Agent
Part of specweave
Test-Aware Planner for generating tasks.md with BDD test plans. Reads spec.md and plan.md to produce implementation tasks with Given/When/Then scenarios. Use during sw:increment orchestration.
Agent Claude Code
Part of velocitai
A code-review agent focused on Python and Playwright tests that use the page object model (POM), a way to organize tests by web pages. It looks for bugs and checks rules without changing functionality.
Agent
Part of spectre
Write or update behavioral tests, drive a strict RED→GREEN→REFACTOR loop, or diagnose failing tests for a given working set. Use to add coverage, verify a TDD gate, or root-cause a test failure; do not use for non-test implementation (dev), independent code review (reviewer), or scoping/planning. Returns tests…
Agent
Fresh-context verification agent for /case. Reads the diff, tests the specific fix with Playwright, creates evidence markers and screenshots. Never implements.
Agent
Compare two outputs WITHOUT knowing which skill produced them.
Agent Claude Code
Part of map-framework
Evaluates solution quality and completeness (MAP).
Agent Codex
Act as a document reconstruction specialist. Your purpose is to prove a distillate's completeness by reconstructing the original source documents from the distillate alone.
Agent Codex
Delivery validation coordinator. Activated at Tier 2+ only (Tier 1 skips Red Team). Spawned by @overseer after development completes (for structural information isolation — overseer never has development context). Independently verifies the delivered product works correctly by dispatching validators…
Agent
Tests web applications end-to-end using Glance browser MCP. Navigates pages, fills forms, clicks buttons, takes screenshots, runs assertions, and reports bugs. Use when you want to verify an app works correctly — login flows, forms, navigation, responsiveness — with real browser interaction.
Agent
Use this agent when you need to run tests and analyze their results. This agent specializes in executing tests using the optimized test runner script, capturing comprehensive logs, and then performing deep analysis to surface key issues, failures, and actionable insights. The agent should be invoked after code changes…
cloudnative-co/claude-code-starter-kit
Agent
End-to-end testing specialist for Playwright or equivalent browser tests. Use for critical user journeys, regression checks, and flaky test triage.
Agent
Part of claude-code-agents
API endpoint testing. Discovery, validation, auth flows, error handling.
Agent
Validates mcp-linear changes with local build, test, and optional live Linear smoke checks.
Agent Claude Code
Quality assurance validator. Runs build verification, template gallery export, bundle export, and reports errors. Use after creating new templates or making editor changes.
Agent
Part of engineer
Use this agent when generating or updating the acceptance test generator for a project, or when the user asks to "build the pipeline", "generate the test generator", "update the pipeline", "create acceptance test infrastructure", or when the ATDD skill reaches step 3 (pipeline generation). Examples: Context: A spec.md…
Agent
Part of engineer
Use this agent when reviewing GWT acceptance test specs for implementation leakage, or when the user asks to "check specs", "review specs", "audit specs", "clean up specs", or "check for leakage". Also invoked by the /spec-check command and as part of the ATDD workflow. Examples: Context: User has written acceptance…
Agent ✓ vendor
Validate an improvement worktree in read-only mode and return a pass or fail verdict.
Agent Claude Code
Use this agent to verify that a Python Agent SDK application is properly configured, follows SDK best practices and documentation recommendations, and is ready for deployment or testing. This agent should be invoked after a Python Agent SDK app has been created or modified.
Agent Claude Code
Compare two outputs WITHOUT knowing which skill produced them.
Agent Claude Code
Evaluate expectations against an execution transcript and outputs.
vasu31dev/playwright-ts-template
Agent Claude Code
Use this agent when you need to create automated browser tests using Playwright and vasu-playwright-utils. Examples: Context: User wants to generate a test for the test plan item.
vasu31dev/playwright-ts-template
Agent Claude Code
Use this agent when you need to debug and fix failing Playwright tests.
vasu31dev/playwright-ts-template
Agent Claude Code
Use this agent when you need to create a comprehensive test plan for a web application or website.
Agent
You are an adversarial verification specialist. Your job is to find evidence, not to reassure.
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: