E2E testing skills

3,650 tagged E2E testing, measured the same way as everything else here.

Browse within: ai-tools 77skills 76ai-coding 73claude-code-plugin 63agent-orchestration 50agentic-workflow 47ai-testing 47playwright 47agentic 41android 40antigravity 36ai-skills 35agent-browser 32test-automation 29

agent-inbox

121

gsd-build/agent-inbox

Skill Claude CodeCodex

Create temporary email inboxes and receive emails for testing auth flows, email verification, account confirmation, and any scenario where an AI agent needs to receive an email. Uses the agent-inbox MCP server with mail.tm + 1secmail fallback.

not rated 61 4mo ago A 54 tokens original MIT

DamianEdwards/copilotd

Skill Claude CodeCodex

Perform real end-to-end validation of copilotd issue and pull request orchestration using live GitHub artifacts.

not rated 59 8d ago A 31 tokens original MIT

testing-rustyclaw-cli

123

rexlunae/RustyClaw

Skill Claude CodeCodex

Test RustyClaw CLI commands end-to-end. Use when verifying CLI changes, swarm commands, new subcommands, or gateway error handling.

not rated 59 11d ago A 36 tokens original MIT

forcedotcom/SalesforceMobileSDK-Templates

Skill Claude CodeCodex

End-to-end test harness for the consolidated SDK consumer skills. Creates 10 apps (5 iOS, 5 Android) in parallel, one per scenario within the ios-mobile-sdk and android-mobile-sdk skills, builds each, and reports results.

not rated 58 2d ago A 56 tokens original BSD-3-Clause

deviludo-test

125

asssaver97/DeviLudo

Skill Claude CodeCodex

Plan and judge complete cross-platform DeviLudo E2E evidence and hand product failures back to Development.

not rated 58 changed yesterday A 26 tokens

usage-pattern-testing

126

Dark-Alex-17/coyote

Skill Claude CodeCodex

Verify a change from the consumer's perspective - exercise the changed surface (HTTP API, RPC, CLI) black-box against a locally running instance with clean, isolated state. Run existing usage suites first for regressions, derive new tests from the spec (never the implementation), and classify every failure as bug /…

not rated 57 today A 126 tokens AGPL-3.0

workshop-testing

127

dotnet-presentations/ai-workshop

Skill Claude CodeCodex

Walk through the .NET AI Workshop as an attendee to validate that the READMEs, commands, and code snapshots still work. USE FOR: testing the workshop, testing a specific Part, dry-running the labs, verifying a README against its snapshot, reconciling or refreshing code snapshots, producing a workshop test report. DO…

not rated 55 +1 10d ago A 95 tokens original MIT

using-axis

128

netlify/axis

Skill Claude CodeCodex

Run AXIS, read its reports, navigate its project layout, and interpret scores. Use when the user asks to run AXIS, invoke the CLI, compare runs, explain a score, find a regression, manage baselines, or understand where AXIS writes its files.

not rated 55 +1 yesterday A 58 tokens original MIT

write-test-plan

129

andresharpe/dotbot

Skill Claude CodeCodex

Generate a QA/UAT test plan from product specifications and task definitions, covering acceptance testing, integration flows, and exploratory testing. Unit tests are out of scope (handled by write-unit-tests skill).

not rated 54 9d ago A 43 tokens original MIT

scout

130

tester-army/scout

Skill Claude CodeCodex

Safely explore and adversarially test an authorized HTTP API using the scout CLI, with or without an OpenAPI spec. Use when asked to test, probe, validate, or explore an API, whether or not an OpenAPI/Swagger spec is available. Scout is the harness; you are the operator.

not rated 51 1mo ago A 65 tokens original MIT

e2e-test

131

coleam00/ai-coding-summit-workshop-2

Skill Claude CodeCodex

Comprehensive end-to-end testing command. Launches parallel sub-agents to research the codebase (structure, database schema, potential bugs), then uses the Vercel Agent Browser CLI to test every user journey — taking screenshots, validating UI/UX, and querying the database to verify records. Run after implementation…

not rated 50 6mo ago A 74 tokens

testing-strategy

132

shinpr/agentic-code

Skill Claude CodeCodex

Selects the narrowest sufficient test boundary from requirements, repository evidence, and maintenance cost. Use when deciding integration or E2E coverage.

not rated 49 7d ago A 32 tokens original MIT

verify-ui

133

vikshana/vikshana-graft-app

Skill Claude CodeCodex

Use when verifying that a frontend code change, LLM harness edit, or system prompt update produces the expected result in the running Graft plugin UI (http://localhost:3000/a/vikshana-graft-app). Drives a real headed Chrome session via Chrome DevTools MCP — navigate, click, inspect console errors, inspect network…

not rated 48 +1 yesterday A 105 tokens AGPL-3.0

vindicate

134

OpenEvident/vindicate

Skill Claude CodeCodex

Use when the user wants to write, add, fix, stabilize (flaky), refactor, run, or audit Playwright browser tests, draft requirements/stories from a recording (no tests), find test-coverage gaps, scaffold a Playwright project, or set up Playwright CI. Vindicate's guided workflow for grounded, conformant Playwright test…

not rated 46 +2 6d ago A 77 tokens original Apache-2.0

e2e-test

136

bun913/mcp-testrail

Skill Claude CodeCodex

Run E2E regression tests against TestRail using MCP tools. Builds the project, then exercises Projects, Suites, Sections, Cases (CRUD), Runs, Tests, Results, Plans, Milestones, and SharedSteps.

not rated 44 2mo ago A 49 tokens original MIT archived

mobile-automation

139

congwa/mobile-agent

Skill Claude CodeCodex

A mobile testing skill that combines a phone-control tool with Feishu, a Chinese collaboration platform, to manage test cases and operate Android or iOS apps.

not rated 44 6mo ago A 0 tokens original Apache-2.0

playwright-cli

140

ErkanBarin/playwright-agent-mcp-starter

Skill Claude CodeCodex

Automates browser interactions for web testing, form filling, screenshots, and data extraction. Use when the user needs to navigate websites, interact with web pages, fill forms, take screenshots, test web applications, or extract information from web pages.

not rated 43 6mo ago A 52 tokens copy · 83% MIT

e2e-loop

141

TOMOSIA-VIETNAM/open-pr

Skill Claude CodeCodex

Run the open-pr e2e fixture through a real review in a fresh subagent, grade it against e2e/checklist.md with a second independent subagent, diagnose each failure to the prompt file that owns it, fix, and repeat until clean or the round budget runs out. Use when changing anything under src/ and you want evidence the…

not rated 42 +1 2d ago A 85 tokens original MIT

tui-uat

142

utensils/mold

Skill Claude CodeCodex

Run acceptance tests on the mold TUI. Use when asked to test, verify, or UAT the TUI, or after making TUI changes that need visual verification.

not rated 42 +3 changed yesterday C 40 tokens original MIT

mewbo-cli-smoketest

143

bearlike/Assistant

Skill Claude CodeCodex

End-to-end smoke testing of the Mewbo CLI via tmux. Use this skill when asked to test the CLI, verify CLI behavior after changes, smoke-test the agent loop, check for regressions, or validate MCP/plugin/session features work correctly through the terminal interface. Also use when validating the live-streaming…

not rated 41 2d ago B 111 tokens original MIT

FTShare-Lab/agent-claim-network

Skill Claude CodeCodex

Verify and debug this project's ratatui/crossterm TUI by running the ACN agent CLI inside tmux, capturing terminal screens as text, sending scripted keys, checking stderr, and writing focused tmux regressions for requested TUI flows. Use after changes to src/sessiontui, src/bin/acn.rs interactive mode, terminal…

not rated 41 today A 108 tokens original Apache-2.0

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: