verifier
25Agent
Verify that implemented features actually work by executing realistic functional scenarios against a running application.
1,071 tagged E2E testing, measured the same way as everything else here.
Browse within: playwright 27agent-orchestration 24code-quality 20agentic-workflow 18Multi-Agent 16harness 16agentic 15claude-ai 15Autonomous Agents 14browser-automation 13agentic-coding 12claude-code-skills 12copilot 11ai-skills 10
Agent
Verify that implemented features actually work by executing realistic functional scenarios against a running application.
Agent Claude Code
Use this agent when you need to test recent code changes using Playwright automation. Examples: Context: The user has just implemented a new login feature and wants to test it. user: "I just added a new login validation feature, can you test it?" assistant: "I'll use the qa agent to test your recent changes with…
Agent Claude Code
An end-to-end testing agent for Flutter apps, meaning tests that operate the app as a user would on a simulator.
Agent
Audience. Future-me (or any agent) the next time a make agent-test / pytest -n auto run in this repo hangs without finishing. The common causes here are xdist worker crash-and-replace cycles and fixture-teardown hangs; the iteration loop below generalizes to any hanging suite.
Agent Claude Code
Orchestrate F1 test drives to validate the Cyrus agent system end-to-end. Use this agent to run comprehensive test drives that verify issue-tracker, EdgeWorker, and renderer components.
Auriti-Labs/geo-optimizer-skill
Agent Claude Code
Performs independent QA, regression testing, build verification, contract validation, smoke checks, and final PASS/PASS WITH ISSUES/FAIL reports across GEO Optimizer and GeoReady.
Agent
Generates integration/E2E test skeletons from Design Doc ACs using ROI-based selection and journey-based E2E reservation. Use when Design Doc is complete and test design is needed, or when "test skeleton/AC/acceptance criteria" is mentioned. Behavior-first approach for minimal tests with maximum coverage.
HazAT/pi-interactive-subagents
Agent
Visual QA tester — navigates web UIs via Chrome CDP, spots visual issues, tests interactions, produces structured reports.
Agent Claude Code
Use this agent when you need comprehensive end-to-end testing of the Overture UI, when a new feature has been added and you need to verify it doesn't break existing functionality, when you need regression testing across the entire application, or when you want absolute certainty that every feature works flawlessly.…
Agent Claude Code
Use this agent when you need to coordinate end-to-end testing across multiple components, optimize build systems, validate deployments, or ensure proper integration between eBPF programs, Rust collector, and frontend components. Examples: Context: User has made changes to both eBPF programs and Rust collector and…
Agent
Use for DISTILL wave — designs E2E acceptance tests from user stories and architecture using Given-When-Then format. EXPANDED scope (plan v3 §3.A, 2026-05-19) — exclusive test-expertise owner; authors ATs with maximum PBT + parametrize density, runs self-completeness audit (7-category taxonomy + 15-item checklist)…
Agent
You are a read-only verification subagent spawned to check a design deliverable the main agent just built or edited. Your only job: load that deliverable, verify it, and report a single verdict — done or needswork — back to the main agent. You must not modify, create, or delete any file, edit the source, build, or…
kangarooking/kangarooking-skills
Agent
Use this agent to test implementations against sprint contracts and specifications. Uses Playwright MCP for E2E testing, Chrome DevTools for UI inspection, and visual tools for verification. Grades implementations and provides specific failure reports. Trigger when user says "test", "evaluate", "qa", "verify", or…
Agent
End-to-end testing specialist using Vercel Agent Browser (preferred) with Playwright fallback. Use PROACTIVELY for generating, maintaining, and running E2E tests. Manages test journeys, quarantines flaky tests, uploads artifacts (screenshots, videos, traces), and ensures critical user flows work.
Agent
Audits .test integration fixtures under tests/php/Integration/Fixtures for drift against the current compiler output. Use after lexer, parser, analyzer, or emitter changes.
Agent Claude Code
Use this agent when the user needs to write, debug, fix, or improve coverage for Playwright end-to-end tests. This includes creating new test specifications, implementing test scenarios from Gherkin comments, debugging test failures, fixing timeout errors, resolving locator issues, improving code coverage by targeting…
iusztinpaul/designing-real-world-ai-agents-workshop
Agent Codex
Tester role definition for /implement-universal. Loaded by the orchestrator at the start of the Tester phase, on logic tickets only. Trusts the SWE phase's happy-path e2e excerpt (does NOT re-run the Make target), runs at most 1 adversarial break path, walks every Acceptance Criterion with concrete evidence, and emits…
Agent
Record a verifiable Canary QA session — explore a flow step by step against one persistent browser, each script a recorded step capturing trace/video/HAR/console, then render report.html. Use when the user wants to verify or QA a flow, capture a trace or video, or produce a shareable report of a browser run.
Agent
Turn a code change into a prioritized browser-QA plan with Canary — read the git diff, infer the affected user-facing workflows, and suggest concrete flows and the checks that must hold, then optionally record them as a session with a report. Use when the user asks what to test for a change, wants to QA a…
Agent Claude Code
Use this agent when you need to create automated browser tests using Playwright Examples: Context: User wants to generate a test for the test plan item.
Agent Claude Code
Use this agent when you need to debug and fix failing Playwright tests.
Agent Claude Code
Use this agent when you need to create comprehensive test plan for a web application or website.
Agent
Expert 1C testing agent. Tests code and functions using web browser automation and the /deploy-and-test command. Deploys configuration to test infobase, performs UI testing with human-like interactions, validates functionality. Use when the user asks to run deployment, UI testing, or verification against a test…
platformplatform/PlatformPlatform
Agent Claude Code
QA code reviewer who validates Playwright E2E test implementations against project rules and patterns. Runs tests, reviews test architecture, and works interactively with the engineer. Never modifies code.