E2E testing

2,791 mods in this category, of every kind an agent can take. Each one carries what it costs per session, what the scan found, and whether it is the original.

regression-tests

25

telepresenceio/telepresence

Skill Claude CodeCodex

Run, scope, or debug telepresence regression tests under regressiontest/ — the integration-level suite. Use when the user wants to run an area, suite, or single test, debug a failure, or says "/regression-tests". Runs go test ./regressiontest scoped with -run, in the background, writing to a log file so heavy output…

not rated 7.3k 10d ago A 82 tokens original Apache-2.0

monke-testing-guide

26

airweave-ai/airweave

Cursor rule Cursor

Monke is Airweave's end-to-end testing framework for source connectors. It creates real test data in external systems, triggers syncs, and verifies data appears correctly in the search index.

not rated 6.6k +4 3mo ago A 7,424 tokens original MIT archived

solopi-ai

27

alipay/SoloPi

Skill Claude CodeCodex

A command-line framework for testing Android apps and devices with SoloPi, including on-device or cloud AI decision models. It manages devices, test cases, recorded interactions, replays, performance history, and evidence.

not rated 6.2k +10 15d ago A 127 tokens original Apache-2.0

test-writer

28

shareAI-lab/Kode-CLI

Agent

Specialized in writing comprehensive test suites. Use for creating unit tests, integration tests, and test documentation.

not rated 5.2k +2 7d ago A 25 tokens original Apache-2.0

99 AGENTS.md

29

ThePrimeagen/99

Instructions file CodexOpenCode

AGENTS.md instructions for ThePrimeagen/99, covering testing and e2e / integration style testing.

not rated 4.8k +1 2mo ago A 374 tokens

codex-qa-tester

30

Waishnav/devspace

Agent

Manual QA profile for browser testing, workflow verification, and regression checks.

not rated 4.4k +119 3d ago A 21 tokens original MIT

android-emulator

31

callstack/agent-device

Skill Claude CodeCodex

Verify and debug native, React Native, Expo, or Flutter apps on an Android Emulator with agent-device. Use when an agent needs to launch an app, inspect its live UI, tap, type, scroll, validate a code change, collect failure evidence, or reproduce a workflow on an Android virtual device.

not rated 4.3k +28 2d ago A 65 tokens original MIT

agent-device

32

callstack/agent-device

MCP server Claude CodeCodexCursor +2

MCP server for mobile app automation: verify, control, and debug iOS, Android, TV, and desktop apps. Runs locally from the agent-device npm package.

not rated 4.3k +28 2d ago A tokens not measured original MIT

httprunner/httprunner

Instructions file

Instructions for httprunner/httprunner, covering claude.md, project overview, development commands, building and testing.

not rated 4.3k +1 8mo ago A 1,029 tokens original Apache-2.0

feishu-e2e-test

34

m1heng/clawdbot-feishu

Skill Claude CodeCodex

Local E2E debug and test framework for clawd-feishu plugin development. Use when debugging message flow, testing bot responses, verifying Feishu web UI interactions, or performing end-to-end validation of the OpenClaw-Feishu integration during development.

not rated 4.2k +1 5mo ago A 58 tokens original MIT

Atmosphere/atmosphere

Skill Claude CodeCodex

Run the pre-release end-to-end sweep of every user-facing surface — the 33 samples under samples/ (booted from their packaged artifacts and driven in a real browser via chrome-devtools MCP), the Expo/React Native client, and the atmosphere CLI. Use before cutting a release, and after any change to the Console bundle…

not rated 3.8k +8 3d ago A 143 tokens original Apache-2.0

stress-test-webhook

36

TracecatHQ/tracecat

Command Claude Code

Send large JSON payloads to a webhook endpoint to verify that result externalization to MinIO is working correctly. This tests that payloads exceeding Temporal's 2MB blob limit are properly externalized to object storage.

not rated 3.8k +1 2d ago A 0 tokens AGPL-3.0

ZTools AGENTS.md

37

ZToolsCenter/ZTools

Instructions file CodexOpenCode

Project instructions for ZTools covering JavaScript documentation comments and end-to-end testing in Electron, a desktop app framework.

not rated 3.7k +28 6d ago A 542 tokens original MIT

tutti-os/tutti

Skill Claude CodeCodex

From a Tutti checkout, run, audit, freshly replay, publish, or diagnose Session Replay cassettes that are driven by case-repository scenario scripts (CDP), not by interactive UI recording. Use for real-Provider capture while a scenario.mjs executes, cassette transport or semantic-state mismatches, fresh replay…

not rated 3.7k +31 2d ago A 127 tokens original Apache-2.0

browser-testing

39

millionco/expect

Cursor rule Cursor

Use when performing browser-based testing, visual verification, accessibility checks, or performance profiling using the browser MCP.

not rated 3.6k +2 4mo ago A 21 tokens

expect

40

millionco/expect

MCP server Claude CodeCodexCursor +2

Let agents test your code in a real browser. Runs locally from the expect-cli npm package.

not rated 3.6k +2 4mo ago A tokens not measured

expect CLAUDE.md

41

millionco/expect

Instructions file

Claude Code instructions for millionco/expect, a project described as: Expect tests your agent's code in a real browser.

not rated 3.6k +2 4mo ago A 3 tokens

playwright

42

langwatch/langwatch

MCP server Claude CodeCodexCursor +2

One of 2 in .mcp.json

MCP server "playwright" as configured in langwatch/langwatch. Launched with bash -c d=$PWD; while [ "$d" != / ] && [ ! -f "$d/dev/scripts/playwright-mcp.sh".

not rated 3.5k +5 2d ago A tokens not measured original Apache-2.0

connect-agent

43

langwatch/langwatch

Skill Claude CodeCodex

Connect the codebase's AI agent to LangWatch agent simulations, so test suites run against the real agent process. Adds a small connect function beside the service startup that calls the agent already in the codebase, which opens an outbound connection and registers the agent with its environment and its run…

not rated 3.5k +5 changed 2d ago A 102 tokens original Apache-2.0

qawolf-cli

44

qawolf/cli

Skill Claude CodeCodex

Manage QA Wolf through the qawolf CLI. Use when asked to create, update, or list coverage requests, bug reports, or maintenance reports; start a run of flows or tags on the QA Wolf platform or read a run's results; list, set, or delete environment variables; manage environments, flows, or tags; request automation of…

not rated 3.4k +1 changed 2d ago C 118 tokens original Apache-2.0

qawolf/cli

Command

How to use what qawolf run get --run-id --json returns, and how to read the Playwright trace it links to.

not rated 3.4k +1 2d ago A 0 tokens original Apache-2.0

qawolf-cli

46

qawolf/cli

Command

Manage QA Wolf through the qawolf CLI. Use when asked to create, update, or list coverage requests, bug reports, or maintenance reports; start a run of flows or tags on the QA Wolf platform or read a run's results; list, set, or delete environment variables; manage environments, flows, or tags; request automation of…

not rated 3.4k +1 2d ago A 118 tokens original Apache-2.0

agent-browser

47

superagent-ai/grok-cli

Skill Claude CodeCodex

Use the host-side agent-browser CLI for local browser smoke tests, screenshots, snapshots, and simple UI validation against forwarded localhost URLs.

not rated 3.4k +4 1mo ago A 31 tokens original MIT

VOICEVOX/voicevox

Skill Claude CodeCodex

A skill for generating Playwright end-to-end tests, which test a complete user flow in a browser. Generated tests use Japanese step names with test.step and do not include comments.

not rated 3.2k +2 2d ago A 52 tokens

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: