Demonstrate
217Agent
Agent for demonstrating VS Code features.
1,038 tagged E2E testing, measured the same way as everything else here.
Browse within: ai-coding 41playwright 41agent-orchestration 28Multi-Agent 20agentic-workflow 19Autonomous Agents 16copilot 16claude-ai 15harness 15agentic-coding 14agentic 13antigravity 13browser-automation 13ai-development 12
Agent
Agent for demonstrating VS Code features.
Agent
Part of ai-workflow
Runs an exploratory, adversarial pass over a feature after the code-critic passes, through whichever surface(s) it exposes (UI via Playwright, API via curl/Bash). Probes beyond the plan and committed tests for issues and edge cases they did not anticipate; does not re-verify the spec or review code style.
Agent
Part of claude-harness
Use this agent when the harness needs to evaluate a generator's output. The evaluator interacts with the running application like a user — clicking, navigating, testing features — and produces a structured eval report with pass/fail verdicts per criterion. Uses Playwright MCP when available.
Agent Claude Code
Use this agent when you need to create, maintain, or improve tests for your codebase. This includes writing unit tests for new functions or components, creating integration tests for API endpoints or user workflows, updating existing tests after code changes, reviewing test coverage, or establishing testing patterns…
Wania-Kazmi/claude-code-autonomous-agent-workflow
Agent Claude Code
End-to-end testing specialist using Playwright. Use PROACTIVELY for generating, maintaining, and running E2E tests. Manages test journeys, quarantines flaky tests, and ensures critical user flows work.
Agent
Drives a TUI or CLI program like a real user in a local tmux pty. Read-only on source; builds or launches the target itself, reports pass/fail with captured screens. Spawn one per program or flow.
aleksandr-chaika/flutter-clean-arch-skills
Agent
QE E2E testing agent for Flutter apps using Maestro. Converts PM test cases into Maestro YAML flows, ensures TestKeys exist in Flutter code, builds app, runs tests on emulator, reports pass/fail per test case. Use after unit/widget tests pass to verify user journeys.
Agent Claude Code
Runs the VZT Flow end-to-end verification ladder (build, tests, TTS-transcribe checks, clean-test latency, paste-test, daemon socket checks, overlay states) and reports real measured numbers — never estimates. Use before claiming a change works, before a release, or when asked to verify VZT Flow.
Agent Claude Code
Specialiste en ecriture de tests unitaires et integration. Utiliser pour generer des tests vitest (React/TS), pytest (Python), cargo test (Rust), ou Playwright BDD (E2E).
Agent
Drive persona scenarios dynamically (live API/UI/tool calls guided by a persona brief) and structural conformance checks, returning integrity-protected evidence.
Agent
Fully impersonate one Agentweaver persona and drive the real target API one live turn at a time via direct curl calls against the live OpenAPI spec — deciding each next action from actual API responses, never a pre-written script. Returns the resulting transcript to Harness.
smallTechOrg/zero-shot-claude-boilerplate
Agent Claude Code
Read-only quality gate. REVIEWS the new code (logic, security, spec-fidelity, style) AND RUNS the phase gate tests against the real LLM/API (keys from .env), the golden-path/live-server smoke, and the UI tests — exercising the EXACT path the user will test so it works first time — and also performs the whole-tree…
Agent
Part of imagine
A visual checking agent for image-to-code work. It renders generated HTML in a headless browser, meaning a browser without a visible window, and compares the result pixel by pixel with the original image.
Agent
E2E verification — executes user-journey use cases through user-facing interfaces (API, UI via Playwright MCP, CLI) and produces a markdown report. Read-only: cannot modify code or write files.
LinkedInLearning/playwright-ai-automating-tests-with-mcp-copilot-and-chatgpt-7263000
Agent
Use this agent when you need to create automated browser tests using Playwright Examples: Context: User wants to generate a test for the test plan item.
LinkedInLearning/playwright-ai-automating-tests-with-mcp-copilot-and-chatgpt-7263000
Agent
Use this agent when you need to debug and fix failing Playwright tests.
LinkedInLearning/playwright-ai-automating-tests-with-mcp-copilot-and-chatgpt-7263000
Agent
Use this agent when you need to create comprehensive test plan for a web application or website.
Agent Claude Code
A web-page quality reviewer that checks visual consistency, data accuracy, semantic HTML, responsive behaviour, accessibility, performance, motion, and browser compatibility. It verifies each module as it is completed and checks connections between related parts.
RedHatInsights/platform-frontend-ai-toolkit
Agent
Part of infrastructure-plugin
Configures Konflux pipeline YAML for E2E testing.
Agent Claude Code
Part of devflow
Senior QA engineer. Verifies built work against acceptance criteria by actually executing it — runs tests and the app, probes edge cases, and reports severity-ranked findings. Use after implementation, before anything is called done.
mh2-lee/everything-claude-code
Agent
End-to-end testing specialist using Playwright. Use PROACTIVELY for generating, maintaining, and running E2E tests. Manages test journeys, quarantines flaky tests, uploads artifacts (screenshots, videos, traces), and ensures critical user flows work.
Agent Claude Code
You are the QA Engineer (SWE-QA) for the Cloud Workstation project. You think in terms of Critical User Journeys (CUJs) — the end-to-end flows that define whether the product works for real users. Your job is to validate that every critical path through the application works correctly, performs well, and is accessible.
SeanningTatum/cf-saas-starter-react-router
Agent Claude Code
Drives a feature's golden path + one error path through the LIVE app using the Playwright CLI (a throwaway Node script run headless), screenshots each step, and writes a verification doc to .brain/features/ /verifications/ .md. Replaces writing per-feature e2e specs. Use for any user-visible feature/flow before…
thaitype/thaitype-stack-spa-react-elysia-prisma-template
Agent Claude Code
Long-running verification agent responsible for non-deterministic, integration, API, UI, and external acceptance testing. Executes medium-to-long duration validation tasks that go beyond local deterministic verification. Does NOT implement features. Does NOT refactor code. Reports findings back to chief-agent.…
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: