E2E testing skills

3,596 tagged E2E testing, measured the same way as everything else here.

Browse within: ai-coding 75claude-code-plugin 65agent-orchestration 51ai-testing 47playwright 46agentic-workflow 45agentic 42android 42ai-skills 33antigravity 33agent-browser 32agentic-coding 29ai-assistant 29copilot 26

obsidian-e2e

49

uphy/obsidian-reminder

Skill Claude CodeCodex

An automated end-to-end test workflow for a running Obsidian application, controlled through Chrome DevTools Protocol. Obsidian is a note-taking app, and end-to-end testing checks behavior across the real application rather than only individual functions.

not rated 656 3d ago A 140 tokens original MIT

pr-reviewer

50

pyuvm/pyuvm

Skill Claude CodeCodex

Evaluates GitHub Pull Requests against a Test Sufficiency Matrix and Intent Realization Alignment, or provides a high-level summary of all open PRs in the repository.

not rated 569 3d ago A 37 tokens

gm-evaluate

51

RandallLiuXin/GodotMaker

Skill Claude CodeCodex

Evaluate the current tag's quality: enforce the playable-closed-loop gate, maintain a single cross-tag e2e/ suite that always reflects the current game (add tests for new mechanics, prune tests for mechanics this tag deliberately removed), and reason about gameplay quality. Independent from the build process — fresh…

not rated 532 2d ago A 84 tokens

plugin-test

52

allenhutchison/obsidian-gemini

Skill Claude CodeCodex

Three-pass acceptance test for the obsidian-gemini plugin — unit tests, then UI/state via the Obsidian CLI (cheap pass), then API-spending verification (only with explicit user authorization). Driven by the user-facing docs as the source of truth for what should work, with extra focus on functionality shipped since…

not rated 521 yesterday A 174 tokens original MIT

moav-e2e

53

MotherofallVPNs/MoaV

Skill Claude CodeCodex

Run and debug MoaV's end-to-end tests — real protocol connectivity (client-test.sh) and the moav CLI smoke test — against a LIVE server, via the self-hosted e2e workflow or a local test VPS. Use when validating a branch before release, diagnosing a protocol that won't connect, or checking that moav CLI commands still…

not rated 427 yesterday B 122 tokens original MIT

senpi-qa

54

code-yeongyu/senpi

Skill Claude CodeCodex

Manual QA harness for the senpi coding agent itself. MUST USE after changing packages/ai, packages/agent, packages/coding-agent, or packages/tui — a green typecheck and npm test are NOT QA. Drives the real CLI from source in an isolated sandbox (never touches /.senpi or real credentials) across four channels: remote…

not rated 419 yesterday A 201 tokens original MIT

reticle

55

reticlehq/reticle

Skill Claude CodeCodex

Reticle embeds a dev-only SDK in the user's running app and exposes it to you as reticle MCP tools. You look, act, observe, and assert against the real app. No screenshots, and no browser download for the verify loop: it drives the tab the user already has open.

not rated 417 changed yesterday A 0 tokens

verify

56

kenn-io/kata

Skill Claude CodeCodex

Build and drive the kata CLI/TUI to verify changes at the real terminal surface.

not rated 412 +4 yesterday A 18 tokens original MIT

analyze-and-plan

57

marcellourbani/vscode_abap_remote_fs

Skill Claude CodeCodex

Standalone Phase 1 of SAP UI testing. Discovers the configured test folder and target system, downloads one complete ABAP source snapshot, then READS that source and captures the full picture in three reference artifacts — flow.md (functional flow), units.md (per-unit input/output inventory), and findings.md (the…

not rated 381 6d ago A 125 tokens original MIT

currents-dev/playwright-best-practices-skill

Skill Claude CodeCodex

Use when writing Playwright tests, fixing flaky tests, debugging failures, implementing Page Object Model, configuring CI/CD, optimizing performance, mocking APIs, handling authentication or OAuth, testing accessibility (axe-core), file uploads/downloads, date/time mocking, WebSockets, geolocation, permissions…

not rated 371 +1 1mo ago A 214 tokens original MIT

appium-skill

59

LambdaTest/agent-skills

Skill Claude CodeCodex

Generates production-grade Appium mobile automation scripts for Android and iOS in Java, Python, or JavaScript. Supports real device and emulator testing locally and on TestMu AI cloud with 100+ real devices. Use when the user asks to automate mobile apps, test on Android/iOS, write Appium tests, or mentions "Appium"…

not rated 366 +1 1mo ago A 135 tokens original MIT

e2e-testing

60

ai-dashboad/flutter-skill

Skill Claude CodeCodex

AI-powered E2E testing for any app — Flutter, React Native, iOS, Android, Electron, Tauri, KMP, .NET MAUI. Connects via MCP to running apps so the agent can take screenshots, tap elements, enter text, scroll, inspect UI trees, and verify state with natural language. Use when the user wants to test an app's UI…

not rated 363 3d ago A 107 tokens original MIT

playwright-cli

61

testdino-hq/playwright-skill

Skill Claude CodeCodex

Automates browser interactions for testing and validating your own web applications using playwright-cli. Use when you need terminal-first browser control for navigation, form filling, screenshots, tracing, bound browser sessions, debugging, or generating Playwright test code. Only use against applications you own or…

not rated 356 +3 2mo ago A 64 tokens original MIT

make-e2e-live

62

receptron/mulmoclaude

Skill Claude CodeCodex

A development guide for adding or improving real-LLM end-to-end tests. End-to-end tests exercise a complete feature path, while a real-LLM test uses an actual language model rather than a fake response.

not rated 344 yesterday A 103 tokens original MIT

testing-webui

63

AnnenkovLabs/girl-agent

Skill Claude CodeCodex

Test the girl-agent WebUI end-to-end locally. Use when verifying profile setup, profile selection, config, assistant, addons, logs, memory, and runtime controls.

not rated 342 +1 1mo ago D 38 tokens

Asymptote-Labs/agent-beacon

Skill Claude CodeCodex

Verify a Beacon change end to end by running a real Claude Code session inside a disposable Linux cloud sandbox and checking that Beacon captured what the agent actually did. Use when asked to verify, validate, test, or prove that a Beacon change works for real rather than just compiling; when asked whether telemetry…

not rated 325 +4 2d ago A 114 tokens original MIT

testing-preview

65

boringcomputers/nehemiah

Skill Claude CodeCodex

Test the preview proxy feature end-to-end. Use when verifying preview URL changes, auth changes on the web proxy route, or networking-related fixes.

not rated 320 +2 24d ago B 32 tokens original Apache-2.0

stove

66

Trendyol/stove

Skill Claude CodeCodex

Use when configuring, writing, or debugging Stove end-to-end tests; choosing JVM, process, container, or provided-application runners; wiring Stove systems; enabling tracing, dashboard, or MCP; or extending Stove with custom systems.

not rated 310 changed yesterday A 49 tokens original Apache-2.0

trailblaze-author

67

block/trailblaze

Skill Claude CodeCodex ✓ vendor

Use when turning a captured human demonstration (a Trail Runner demonstration bundle: demo.yaml + actions.ndjson + per-action screenshots and view hierarchies) into a durable, independently runnable Trailblaze trail. Trigger when a prompt hands you a demonstration bundle directory and asks you to author, generate, or…

not rated 310 +2 2d ago A 99 tokens original Apache-2.0

droid-ash/finalrun-agent

Skill Claude CodeCodex

Generate test and suite specifications in the strict FinalRun YAML format. Handles automated test planning, folder grouping by feature, repo app configuration, environment-specific overrides in .finalrun/env/.yaml, and validation via finalrun check.

not rated 305 1mo ago A 51 tokens original Apache-2.0

Fu-Jie/openwebui-extensions

Skill Claude CodeCodex

Use when inspecting UI bugs in OpenWebUI plugins, taking screenshots of plugin output, capturing console errors, testing Action/Filter/Pipe plugins in the chat interface, or verifying plugin installation in the Admin panel. Triggered by: plugin UI bug, Action HTML output, screenshot, console error, plugin not working…

not rated 302 +1 1mo ago A 80 tokens original MIT

n9e/fe

Skill Claude CodeCodex

Maintain config-driven Nightingale E2E tests that convert JSON config data into UI-readable normalized values, drive Playwright + Midscene interactions, and verify persistence through APIs.

not rated 303 today A 43 tokens original Apache-2.0

playwright-bowser

71

disler/bowser

Skill Claude CodeCodex

Headless browser automation using Playwright CLI. Use when you need headless browsing, parallel browser sessions, UI testing, screenshots, web scraping, or browser automation that can run in the background. Keywords - playwright, headless, browser, test, screenshot, scrape, parallel.

not rated 262 +2 6mo ago A 62 tokens

edt-mcp-e2e-testing

72

DitriXNew/EDT-MCP

Skill Claude CodeCodex

How to write and run black-box end-to-end (e2e) tests for the EDT-MCP server. Covers the architecture (shared harness + one file per tool + orchestrator), the git-fixture isolation protocol, happy-path AND negative coverage, error-quality assertions, and the anti-cheat rules. An agent that has never seen this project…

not rated 265 changed yesterday A 94 tokens AGPL-3.0

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: