E2E testing agents

1,133 tagged E2E testing, measured the same way as everything else here.

Browse within: playwright 41agent-orchestration 27Multi-Agent 21code-quality 20agentic-workflow 19claude-ai 17Autonomous Agents 16copilot 16harness 16agentic 15browser-automation 15agentic-coding 14antigravity 14ai-development 12

auth-route-tester

98

blencorp/claude-code-kit

Agent

Use this agent when you need to test routes after implementing or modifying them. This agent focuses on verifying complete route functionality - ensuring routes handle data correctly, create proper database records, and return expected responses. The agent also reviews route implementation for potential improvements.…

101 9mo ago A 374 tokens original MIT

e2e-runner

99

krishnakanthb13/everything-antigravity

Agent

End-to-end testing specialist using Vercel Agent Browser (preferred) with Playwright fallback. Use PROACTIVELY for generating, maintaining, and running E2E tests. Manages test journeys, quarantines flaky tests, uploads artifacts (screenshots, videos, traces), and ensures critical user flows work.

90 6mo ago A 69 tokens

feature-tester-e2e

100

AIBiz-Automatyzacje/claude-code-starter

Agent Claude Code

Weryfikuje scenariusze E2E w przeglądarce przez agent-browser. Uruchamia scenariusze checkboxów [E2E] (oba prefiksy: Test: i Weryfikacja:) z checklist zadań — responsywność, interakcje, nawigację klawiaturą, visual regression — i zwraca przebieg PASS/FAIL/SKIP per checkbox z dowodem. Nie pisze seedów, nie modyfikuje…

82 9d ago A 136 tokens

dom-extraction-tester

101

WebMCP-org/npm-packages

Agent Claude Code

Use this agent when you need to test progressive DOM reading implementations by navigating to websites and extracting specific information. This agent works as a driver that receives instructions from a navigator AI about what website to visit and what data to extract, then attempts the extraction and reports back on…

80 3d ago A 275 tokens original MIT

harness-implementer

102

panayiotism/claude-harness

Agent

Part of claude-harness

Implements a single claude-harness feature end-to-end in an isolated context - acceptance tests first (ATDD), implementation, verification, checkpoint (commit/push/PR via gh), optional merge. Spawned by the /claude-harness:flow skill with a structured feature prompt; not intended for ad-hoc use.

79 1mo ago A 73 tokens original MIT

browser-tester-v2

103

lipas-liikuntapaikat/lipas

Agent Claude Code

Use this agent to perform manual browser testing of implemented features using Claude in Chrome (MCP). Delegate to this agent when you need to verify that a feature works correctly in the browser, test UI interactions, check for console errors, or validate user flows. Provide context about what was implemented and…

78 3d ago A 69 tokens original MIT

integ-test-runner

104

Apra-Labs/apra-fleet

Agent

Runs integ-test-playbook.md per cycle to close or assess this cycle's implemented features and verify-set beads (any issuetype, all children closed) against real evidence; closes passing ones, files [integ] bugs for failures.

72 5d ago A 53 tokens

test-engineer

105

xvirobotics/metaskill

Agent Claude Code

Use this agent when tests need to be written, debugged, or improved. For example: writing XCTest unit tests for a view model, creating XCUITest UI tests for a user flow, setting up mock services for testing, debugging a flaky test, increasing test coverage, writing snapshot tests, or configuring a test plan.

67 6mo ago A 69 tokens original MIT

api-skill-tester

106

Bria-AI/bria-skill

Agent Claude Code

Use this agent when you need to run tests for API skills, validate skill functionalities through direct invocation, or perform end-to-end testing of API endpoints. This includes running existing test suites, exercising API skills manually to verify behavior, and validating that skill functionalities work as…

65 3d ago A 471 tokens

e2e-tester

107

laguagu/claude-code-nextjs-skills

Agent

Part of claude-code-nextjs-skills

Tests web applications end-to-end by exercising real user flows and fixing verified code-level issues. Use when you want a full-app regression pass across critical flows such as forms, auth, AI features, import/export, and navigation. Reports infrastructure, environment, and product-level issues that require manual…

61 6d ago A 84 tokens original MIT

Fighter90/career-ops-ui

Agent Claude Code

Use this agent after adding or modifying any file in tests/ to verify the test will pass in CI (no parent-project dependency, no live network, no port collision). Invoke proactively after every test diff.

58 3d ago A 49 tokens original MIT

test-analysis-design

110

ZTE-AICloud/Co-OmniSpec

Agent

Part of omni-dsdd

A test-analysis and test-case design workflow that turns requirements into test points and black-box tests. It is used after `/specify` and `/clarify` commands.

54 1mo ago A 109 tokens original MIT

migration-planner

111

adriannoes/awesome-agentic-ai

Agent

Part of pw

Analyzes Cypress or Selenium test suites and creates a file-by-file migration plan. Invoked by /pw:migrate before conversion starts.

53 5d ago A 31 tokens original MIT

test-debugger

112

adriannoes/awesome-agentic-ai

Agent

Part of pw

Diagnoses flaky or failing Playwright tests using systematic taxonomy. Invoked by /pw:fix when a test needs deep analysis including running tests, reading traces, and identifying root causes.

53 5d ago A 41 tokens original MIT

product-manager

113

ATTCKDigital/smith

Agent

Part of smith

Product manager who owns E2E acceptance testing, requirements clarity, and business impact analysis. Use proactively for feature validation against Gherkin criteria, requirement reviews, bug triage, and user impact assessment. MANDATORY reviewer after /smith.specify, during /smith.clarify, and owns E2E testing after…

52 1mo ago A 78 tokens original MIT

staff-frontend

114

ATTCKDigital/smith

Agent

Part of smith

Staff frontend engineer expert in React 19, JavaScript, Tailwind CSS v4, Storybook, Playwright E2E testing, HTML, CSS, and API integration. Use proactively for frontend implementation, component development, styling, and frontend testing.

52 1mo ago A 55 tokens original MIT

stagehand-expert

115

chongdashu/browserbase-claude-code-stagehand

Agent Claude Code

Use this agent when you need executable Stagehand test files for TDD workflow. ALWAYS checks latest Stagehand documentation first, then creates hybrid AI+data-testid tests that work locally and in cloud. Expert in LOCAL vs BROWSERBASE modes, proper API usage (stagehand.page.act/observe), and fallback strategies for…

52 1y ago A 0 tokens

gotalab/uxaudit

Agent

Part of uxaudit

Journey Compiler for the uxaudit pipeline. Translates each natural-language primaryjourney from project-context.json into an executable JSON action script that capturejourney.mjs can replay against the running app. Reads project-context.json + curls the running app for selector hints, writes one journey-scripts/ .json…

52 4mo ago A 101 tokens original Apache-2.0

uxaudit-l4-judge

117

gotalab/uxaudit

Agent

Part of uxaudit

L4-journey Judge for the uxaudit pipeline. Reads the per-journey capture directory (screenshots + steps.json + evaluation brief) and writes a strict pass/fail/unverifiable verdict with four-axis journeyevaluation to result.json. Has Read, Write, Glob — no Bash, no WebFetch — physically cannot drive a browser or curl…

52 4mo ago A 111 tokens original Apache-2.0

test-writer

118

alleneubank/claude-code

Agent

Writes unit, integration, and e2e tests that verify correctness without gaming assertions. Use when adding test coverage or writing new tests.

52 2mo ago A 31 tokens original Apache-2.0 archived

e2e-runner

119

weny911/AI-coding-almighty

Agent Claude Code

End-to-end testing specialist using Playwright. Use PROACTIVELY for generating, maintaining, and running E2E tests. Manages test journeys, quarantines flaky tests, uploads artifacts (screenshots, videos, traces), and ensures critical user flows work.

51 6mo ago A 59 tokens

Rtur2003/Claude-Code-Promts-Skills

Agent

You are a testing specialist agent. Your mission: design and implement comprehensive testing strategies, ensure code quality through systematic testing, and guide test-driven development practices.

49 1mo ago A 0 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: