Command Claude Code
You are a post-release testing agent. Your job is to verify that a sandbox-agent release works correctly.
6,258 tagged E2E testing, measured the same way as everything else here.
Browse within: awesome-list 11cursor-ai 11ai-coding-assistant 10cursorrules 10cursor-rules 9playwright 9cursor-ide 8github-copilot 7ai-testing 6copilot 6access 5hpc 5metrics 5performance-analysis 5
Command Claude Code
You are a post-release testing agent. Your job is to verify that a sandbox-agent release works correctly.
Open-Source-Legal/OpenContracts
Cursor rule Cursor
How to run frontend component tests.
Skill Claude CodeCodex
Run a live-instance verification of traceway-cli that goes beyond the Go smoke suite — exercises real-data detail endpoints, TTY-default rendering, adaptive metric-name discovery, and emits a human-readable coverage report. Invoke ONLY when the user explicitly asks (e.g. "run integration tests", "verify the CLI…
Agent Claude Code
An end-to-end testing assistant built around Playwright, a tool that controls real browsers to test complete user journeys.
Command Claude Code
A command for creating and running end-to-end tests with Playwright. End-to-end tests check a complete user journey across the application, such as logging in, searching, or paying.
Skill Claude CodeCodex
A test-driven development workflow, where tests are written before the code they check. TDD means using failing tests to define a feature, then implementing and refining the code until the tests pass.
Agent
You are a team leader for batch-barrier E2E testing.
Agent
You are a team leader for worker-pool E2E testing.
Agent
You are a team leader for E2E testing. Your job is to decompose a task into independent subtasks.
Skill Claude CodeCodex
Build, launch, and drive the mindwalk web UI end-to-end for verification.
Skill Claude CodeCodex
Prepare the environment and run the LLM-driven agent evals (e2e/agent-evals/) against a chosen sim-use binary. Use when the user runs /run-evals or asks to "run the agent evals", "run the LLM-driven tests", "eval the skill", or wants pre-release confidence that an agent reading the bundled skill still picks the right…
Agent
Verify that implemented features actually work by executing realistic functional scenarios against a running application.
conorluddy/ios-simulator-skill
Instructions file
Instructions for conorluddy/ios-simulator-skill, covering claude.md - developer guide, project overview, project structure, architecture patterns and pattern 1: class-based script design.
conorluddy/ios-simulator-skill
Plugin Claude Code
29 production-ready scripts for iOS app testing, building, and automation. Provides semantic UI navigation, build automation, accessibility testing, and simulator lifecycle management.
conorluddy/ios-simulator-skill
Skill Claude CodeCodex
29 production-ready scripts for iOS app testing, building, and automation. Provides semantic UI navigation, build automation, accessibility testing, and simulator lifecycle management. Optimized for AI agents with minimal token output.
Skill Claude CodeCodex
Smoke test the full running Centaur system from a new Slack thread. Use when asked to QA the stack, run a smoke test, verify a deployment, check stack health, check deploy readiness, or prove Slack tools, file upload/download, company context, logs, metrics, and tool loading work end to end.
Agent Claude Code
Use this agent when you need to test recent code changes using Playwright automation. Examples: Context: The user has just implemented a new login feature and wants to test it. user: "I just added a new login validation feature, can you test it?" assistant: "I'll use the qa agent to test your recent changes with…
Skill Claude CodeCodex
Run tests. Use after code changes to validate. Arguments: unit (default, no GPU), e2e (with models), filter name, or all.
jerry-ai-dev/MODULAR-RAG-MCP-SERVER
Skill Claude CodeCodex
Fully autonomous QA testing agent for Modular RAG MCP Server. Reads test cases from QATESTPLAN.md, executes ALL test types automatically without human intervention — CLI commands, Dashboard UI via Streamlit AppTest headless rendering, MCP protocol via subprocess JSON-RPC, provider switches, and data lifecycle checks.…
Skill Claude CodeCodex
Crabbox/Testbox remote proof: portable provider routing, untrusted isolation, Linux/macOS/Windows/WSL2, live E2E, diagnostics, cleanup.
Skill Claude CodeCodex
Set up a local agent replay server for Raindrop Workshop. Use when the user wants Workshop to replay a captured trace against their real local agent code and tools. Creates/updates .raindrop/agents.yaml, scaffolds a language-appropriate replay server, registers the project with raindrop replay register, and verifies…
Agent Claude Code
An end-to-end testing agent for Flutter apps, meaning tests that operate the app as a user would on a simulator.
Skill Claude CodeCodex
A guide for automated end-to-end and interface checks of a Flutter app using a Dart tool connection and Marionette. End-to-end testing checks a complete user flow, while a simulator runs the app without a physical phone.
Skill Claude CodeCodex
Use when working with Roborazzi screenshot tests on Android/JVM — setting up the Roborazzi Gradle plugin, running record/compare/verify tasks, writing tests with captureRoboImage or RoborazziRule, Compose Preview screenshot testing (ComposablePreviewScanner), Compose Multiplatform (iOS/desktop) screenshots, AI-powered…
At most 3 mods per repository are shown here — the rest are on their repository pages: