E2E testing agents

1,038 tagged E2E testing, measured the same way as everything else here.

Browse within: ai-coding 41playwright 41agent-orchestration 28Multi-Agent 20agentic-workflow 19Autonomous Agents 16copilot 16claude-ai 15harness 15agentic-coding 14agentic 13antigravity 13browser-automation 13ai-development 12

Demonstrate

217

malwarebo/nyrve

Agent

Agent for demonstrating VS Code features.

5 2mo ago A 10 tokens copy · 100% MIT

adversarial-qa

218

cunhaax/ai-workflow

Agent

Part of ai-workflow

Runs an exploratory, adversarial pass over a feature after the code-critic passes, through whichever surface(s) it exposes (UI via Playwright, API via curl/Bash). Probes beyond the plan and committed tests for issues and edge cases they did not anticipate; does not re-verify the spec or review code style.

5 6d ago A 72 tokens original MIT

harness-evaluator

219

uppifyagency/claude-harness

Agent

Part of claude-harness

Use this agent when the harness needs to evaluate a generator's output. The evaluator interacts with the running application like a user — clicking, navigating, testing features — and produces a structured eval report with pass/fail verdicts per criterion. Uses Playwright MCP when available.

5 5mo ago A 59 tokens original MIT

tester

220

ecorkran/context-forge

Agent Claude Code

Use this agent when you need to create, maintain, or improve tests for your codebase. This includes writing unit tests for new functions or components, creating integration tests for API endpoints or user workflows, updating existing tests after code changes, reviewing test coverage, or establishing testing patterns…

5 22d ago A 252 tokens original MIT

e2e-runner

221

Wania-Kazmi/claude-code-autonomous-agent-workflow

Agent Claude Code

End-to-end testing specialist using Playwright. Use PROACTIVELY for generating, maintaining, and running E2E tests. Manages test journeys, quarantines flaky tests, and ensures critical user flows work.

5 7mo ago A 49 tokens

tui-tester

222

uwuclxdy/agenticat

Agent

Drives a TUI or CLI program like a real user in a local tmux pty. Read-only on source; builds or launches the target itself, reports pass/fail with captured screens. Spawn one per program or flow.

5 3d ago A 52 tokens original MIT

maestro-tester

223

aleksandr-chaika/flutter-clean-arch-skills

Agent

QE E2E testing agent for Flutter apps using Maestro. Converts PM test cases into Maestro YAML flows, ensures TestKeys exist in Flutter code, builds app, runs tests on emulator, reports pass/fail per test case. Use after unit/widget tests pass to verify user journeys.

5 5mo ago A 63 tokens original MIT

flow-verifier

224

vonzelle-vzt/vzt-flow

Agent Claude Code

Runs the VZT Flow end-to-end verification ladder (build, tests, TTS-transcribe checks, clean-test latency, paste-test, daemon socket checks, overlay states) and reports real measured numbers — never estimates. Use before claiming a change works, before a release, or when asked to verify VZT Flow.

5 27d ago A 68 tokens original MIT

test-writer

225

stoa-platform/stoa

Agent Claude Code

Specialiste en ecriture de tests unitaires et integration. Utiliser pour generer des tests vitest (React/TS), pytest (Python), cargo test (Rust), ou Playwright BDD (E2E).

5 1mo ago A 49 tokens original Apache-2.0

Harness

226

sabbour/agentweaver

Agent

Drive persona scenarios dynamically (live API/UI/tool calls guided by a persona brief) and structural conformance checks, returning integrity-protected evidence.

5 3d ago A 30 tokens original MIT

PersonaActor

227

sabbour/agentweaver

Agent

Fully impersonate one Agentweaver persona and drive the real target API one live turn at a time via direct curl calls against the live OpenAPI spec — deciding each next action from actual API responses, never a pre-written script. Returns the resulting transcript to Harness.

5 3d ago A 56 tokens original MIT

qa-auditor

228

smallTechOrg/zero-shot-claude-boilerplate

Agent Claude Code

Read-only quality gate. REVIEWS the new code (logic, security, spec-fidelity, style) AND RUNS the phase gate tests against the real LLM/API (keys from .env), the golden-path/live-server smoke, and the UI tests — exercising the EXACT path the user will test so it works first time — and also performs the whole-tree…

5 1mo ago A 164 tokens

visual-verifier

229

Mineru98/imagine

Agent

Part of imagine

A visual checking agent for image-to-code work. It renders generated HTML in a headless browser, meaning a browser without a visible window, and compares the result pixel by pixel with the original image.

5 4d ago A 113 tokens original MIT

verify-e2e

230

pablomarin/claude-codex-forge

Agent

E2E verification — executes user-journey use cases through user-facing interfaces (API, UI via Playwright MCP, CLI) and produces a markdown report. Read-only: cannot modify code or write files.

5 5d ago A 48 tokens original MIT

web-qa

234

revfactory/sk-hynix-report

Agent Claude Code

A web-page quality reviewer that checks visual consistency, data accuracy, semantic HTML, responsive behaviour, accessibility, performance, motion, and browser compatibility. It verifies each module as it is completed and checks connections between related parts.

5 3mo ago A 131 tokens

qa-engineer

236

ljojua1998/skills

Agent Claude Code

Part of devflow

Senior QA engineer. Verifies built work against acceptance criteria by actually executing it — runs tests and the app, probes edge cases, and reports severity-ranked findings. Use after implementation, before anything is called done.

5 1mo ago A 46 tokens original MIT

e2e-runner

237

mh2-lee/everything-claude-code

Agent

End-to-end testing specialist using Playwright. Use PROACTIVELY for generating, maintaining, and running E2E tests. Manages test journeys, quarantines flaky tests, uploads artifacts (screenshots, videos, traces), and ensures critical user flows work.

5 7mo ago A 59 tokens

swe-qa

238

ameer00/cloud-workstations

Agent Claude Code

You are the QA Engineer (SWE-QA) for the Cloud Workstation project. You think in terms of Critical User Journeys (CUJs) — the end-to-end flows that define whether the product works for real users. Your job is to validate that every critical path through the application works correctly, performs well, and is accessible.

5 1mo ago A 0 tokens

feature-verifier

239

SeanningTatum/cf-saas-starter-react-router

Agent Claude Code

Drives a feature's golden path + one error path through the LIVE app using the Playwright CLI (a throwaway Node script run headless), screenshots each step, and writes a verification doc to .brain/features/ /verifications/ .md. Replaces writing per-feature e2e specs. Use for any user-visible feature/flow before…

5 27d ago A 131 tokens

tester-agent

240

thaitype/thaitype-stack-spa-react-elysia-prisma-template

Agent Claude Code

Long-running verification agent responsible for non-deterministic, integration, API, UI, and external acceptance testing. Executes medium-to-long duration validation tasks that go beyond local deterministic verification. Does NOT implement features. Does NOT refactor code. Reports findings back to chief-agent.…

5 5mo ago A 74 tokens

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: