E2E testing agents

1,071 tagged E2E testing, measured the same way as everything else here.

Browse within: playwright 27agent-orchestration 24code-quality 20agentic-workflow 18Multi-Agent 16harness 16agentic 15claude-ai 15Autonomous Agents 14browser-automation 13agentic-coding 12claude-code-skills 12copilot 11ai-skills 10

verifier

25

higress-group/himarket

Agent

Verify that implemented features actually work by executing realistic functional scenarios against a running application.

1.3k +1 13d ago A 0 tokens

qa

26

sheshbabu/zen

Agent Claude Code

Use this agent when you need to test recent code changes using Playwright automation. Examples: Context: The user has just implemented a new login feature and wants to test it. user: "I just added a new login validation feature, can you test it?" assistant: "I'll use the qa agent to test your recent changes with…

1.2k +3 4d ago A 0 tokens AGPL-3.0

e2e-verifier

27

K9i-0/ccpocket

Agent Claude Code

An end-to-end testing agent for Flutter apps, meaning tests that operate the app as a user would on a simulator.

1.0k 3d ago A 65 tokens original MIT

Pipelex/pipelex

Agent

Audience. Future-me (or any agent) the next time a make agent-test / pytest -n auto run in this repo hangs without finishing. The common causes here are xdist worker crash-and-replace cycles and fixture-teardown hangs; the iteration loop below generalizes to any hanging suite.

852 3d ago A 0 tokens original MIT

f1-test-drive

29

cyrusagents/cyrus

Agent Claude Code

Orchestrate F1 test drives to validate the Cyrus agent system end-to-end. Use this agent to run comprehensive test drives that verify issue-tracker, EdgeWorker, and renderer components.

792 +1 4d ago A 43 tokens original Apache-2.0

geo-qa-verifier

30

Auriti-Labs/geo-optimizer-skill

Agent Claude Code

Performs independent QA, regression testing, build verification, contract validation, smoke checks, and final PASS/PASS WITH ISSUES/FAIL reports across GEO Optimizer and GeoReady.

744 7d ago A 42 tokens original MIT

shinpr/claude-code-workflows

Agent

Generates integration/E2E test skeletons from Design Doc ACs using ROI-based selection and journey-based E2E reservation. Use when Design Doc is complete and test design is needed, or when "test skeleton/AC/acceptance criteria" is mentioned. Behavior-first approach for minimal tests with maximum coverage.

675 5d ago A 68 tokens original MIT

visual-tester

32

HazAT/pi-interactive-subagents

Agent

Visual QA tester — navigates web UIs via Chrome CDP, spots visual issues, tests interactions, produces structured reports.

662 +9 3mo ago A 28 tokens original MIT

SixHq/Overture

Agent Claude Code

Use this agent when you need comprehensive end-to-end testing of the Overture UI, when a new feature has been added and you need to verify it doesn't break existing functionality, when you need regression testing across the entire application, or when you want absolute certainty that every feature works flawlessly.…

635 5mo ago A 448 tokens original MIT

eunomia-bpf/agentsight

Agent Claude Code

Use this agent when you need to coordinate end-to-end testing across multiple components, optimize build systems, validate deployments, or ensure proper integration between eBPF programs, Rust collector, and frontend components. Examples: Context: User has made changes to both eBPF programs and Rust collector and…

613 8d ago A 239 tokens original MIT

nWave-ai/nWave

Agent

Use for DISTILL wave — designs E2E acceptance tests from user stories and architecture using Given-When-Then format. EXPANDED scope (plan v3 §3.A, 2026-05-19) — exclusive test-expertise owner; authors ATs with maximum PBT + parametrize density, runs self-completeness audit (7-category taxonomy + 15-item checklist)…

602 3d ago A 0 tokens original MIT

fork-verifier-agent

36

FradSer/dotclaude

Agent

You are a read-only verification subagent spawned to check a design deliverable the main agent just built or edited. Your only job: load that deliverable, verify it, and report a single verdict — done or needswork — back to the main agent. You must not modify, create, or delete any file, edit the source, build, or…

588 +1 21d ago A 0 tokens copy · 100% MIT

evaluator

37

kangarooking/kangarooking-skills

Agent

Use this agent to test implementations against sprint contracts and specifications. Uses Playwright MCP for E2E testing, Chrome DevTools for UI inspection, and visual tools for verification. Grades implementations and provides specific failure reports. Trigger when user says "test", "evaluate", "qa", "verify", or…

553 3d ago A 69 tokens

e2e-runner

38

vibeeval/vibecosystem

Agent

End-to-end testing specialist using Vercel Agent Browser (preferred) with Playwright fallback. Use PROACTIVELY for generating, maintaining, and running E2E tests. Manages test journeys, quarantines flaky tests, uploads artifacts (screenshots, videos, traces), and ensures critical user flows work.

530 24d ago A 69 tokens original MIT

fixture-reviewer

39

phel-lang/phel-lang

Agent

Audits .test integration fixtures under tests/php/Integration/Fixtures for drift against the current compiler output. Use after lexer, parser, analyzer, or emitter changes.

523 9d ago A 37 tokens original MIT

istvan-ujjmeszaros/touchspin

Agent Claude Code

Use this agent when the user needs to write, debug, fix, or improve coverage for Playwright end-to-end tests. This includes creating new test specifications, implementing test scenarios from Gherkin comments, debugging test failures, fixing timeout errors, resolving locator issues, improving code coverage by targeting…

501 5mo ago A 0 tokens

tester

41

iusztinpaul/designing-real-world-ai-agents-workshop

Agent Codex

Tester role definition for /implement-universal. Loaded by the orchestrator at the start of the Tester phase, on logic tickets only. Trusts the SWE phase's happy-path e2e excerpt (does NOT re-run the Make target), runs at most 1 adversarial break path, walks every Acceptance Criterion with concrete evidence, and emits…

501 +1 3mo ago A 99 tokens copy · 89% MIT

session-agent

42

0xnyn/canary

Agent

Record a verifiable Canary QA session — explore a flow step by step against one persistent browser, each script a recorded step capturing trace/video/HAR/console, then render report.html. Use when the user wants to verify or QA a flow, capture a trace or video, or produce a shareable report of a browser run.

476 +1 2mo ago A 69 tokens

verify-agent

43

0xnyn/canary

Agent

Turn a code change into a prioritized browser-QA plan with Canary — read the git diff, infer the affected user-facing workflows, and suggest concrete flows and the checks that must hold, then optionally record them as a session with a report. Use when the user asks what to test for a change, wants to QA a…

476 +1 2mo ago A 81 tokens

saltbo/agent-kanban

Agent Claude Code

Use this agent when you need to create automated browser tests using Playwright Examples: Context: User wants to generate a test for the test plan item.

458 5d ago A 151 tokens

1c-tester

47

comol/ai_rules_1c

Agent

Expert 1C testing agent. Tests code and functions using web browser automation and the /deploy-and-test command. Deploys configuration to test infobase, performs UI testing with human-like interactions, validates functionality. Use when the user asks to run deployment, UI testing, or verification against a test…

448 +4 2d ago A 69 tokens

qa-reviewer

48

platformplatform/PlatformPlatform

Agent Claude Code

QA code reviewer who validates Playwright E2E test implementations against project rules and patterns. Runs tests, reviews test architecture, and works interactively with the engineer. Never modifies code.

438 1mo ago B 41 tokens original MIT