Testing

38,352 mods in this category, of every kind an agent can take. Each one carries what it costs per session, what the scan found, and whether it is the original.

microsoft/playwright

Skill Claude CodeCodex ✓ vendor

Query Playwright CI test results from the aggregated DuckDB database. Answers questions about flaky tests, failure rates, slow tests, and per-run/SHA/PR results without hunting through GitHub artifacts.

not rated 96k +115 today A 45 tokens original Apache-2.0

microsoft/playwright

Instructions file GitHub Copilot ✓ vendor

Copilot instructions for microsoft/playwright, a project described as: Playwright is a framework for Web Testing and Automation. It allows testing Chromium, Firefox and WebKit with a single API.

not rated 96k +115 today A 121 tokens original Apache-2.0

microsoft/playwright

Instructions file ✓ vendor

Claude Code instructions for microsoft/playwright, covering monorepo packages, browser packages, tooling packages, key directories and build.

not rated 96k +115 today A 1,902 tokens original Apache-2.0

microsoft/playwright

Agent Claude Code ✓ vendor

Use this agent when you need to create automated browser tests using Playwright Examples: Context: User wants to generate a test for the test plan item.

not rated 96k +115 today A 151 tokens original Apache-2.0

microsoft/playwright

Agent Claude Code ✓ vendor

Use this agent when you need to debug and fix failing Playwright tests.

not rated 96k +115 today A 20 tokens original Apache-2.0

microsoft/playwright

Agent Claude Code ✓ vendor

Use this agent when you need to create comprehensive test plan for a web application or website.

not rated 96k +115 today A 23 tokens original Apache-2.0

playwright-test

31

microsoft/playwright

MCP server Claude CodeCodexCursor +2 ✓ vendor

A high-level API to automate web browsers. Runs locally from the playwright npm package.

not rated 96k +115 today A tokens not measured original Apache-2.0

playwright-cli

32

microsoft/playwright

Skill Claude CodeCodex ✓ vendor

Automate browser interactions, test web pages and work with Playwright tests.

not rated 96k +115 today A 19 tokens original Apache-2.0

microsoft/playwright

Skill Claude CodeCodex ✓ vendor

Set up component testing with Playwright using a story gallery — scaffold stories and a gallery dev page driven by the built-in mount fixture, no dedicated component-testing runtime. Use when asked to test React or Vue components in isolation with Playwright, or to migrate off @playwright/experimental-ct-react / -vue.

not rated 96k +115 changed today A 69 tokens original Apache-2.0

servers CLAUDE.md

34

modelcontextprotocol/servers

Instructions file ✓ vendor

Claude Code instructions for modelcontextprotocol/servers, covering claude.md, project overview, monorepo structure, build & test commands and typescript servers.

not rated 90k +27 today A 953 tokens

gpui-test

35

zed-industries/zed

Skill Claude CodeCodex

Use when writing, debugging, or reproducing GPUI tests in Zed, including gpui::test arguments, TestAppContext parameters, scheduler seeds, ITERATIONS/SEED reproduction, parking failures, and pending task traces.

not rated 89k 4d ago A 50 tokens

verify-worldmonitor

36

koala73/worldmonitor

Skill Claude CodeCodex

Launch and drive the WorldMonitor browser dashboard (Vite app at /dashboard) to prove user-facing behavior with screenshots and transcripts. Use when a change needs proof in the real app rather than only unit tests — panels, map layers, settings, search, country briefs, boot — or when asked to run, screenshot, or…

not rated 85k +305 today A 73 tokens AGPL-3.0

bytedance/deer-flow

Instructions file GitHub Copilot ✓ vendor

Copilot instructions for bytedance/deer-flow, covering copilot onboarding instructions for deerflow, 1) repository summary, 2) runtime and toolchain requirements, 3) build/test/lint/run - verified command sequences and a. bootstrap and install.

not rated 81k +104 today A 1,714 tokens original MIT

skill-creator

38

bytedance/deer-flow

Skill Claude CodeCodex ✓ vendor

Create new skills, modify and improve existing skills, and measure skill performance. Use when users want to create a skill from scratch, edit, or optimize an existing skill, run evals to test a skill, benchmark skill performance with variance analysis, or optimize a skill's description for better triggering accuracy.

not rated 81k +104 today A 64 tokens original MIT

c4ai-check

39

unclecode/crawl4ai

Command Claude Code

Test current changes with adversarial tests, then run full regression suite.

not rated 81k +766 2d ago A 13 tokens original Apache-2.0

coder

40

paperclipai/paperclip

Agent

Use this template when hiring software engineers who implement code, debug issues, write tests, and coordinate with QA or engineering leadership.

not rated 80k +145 today A 0 tokens original MIT

debugger

41

rtk-ai/rtk

Agent Claude Code

Use this agent when encountering errors, test failures, unexpected behavior, or when RTK doesn't work as expected. This agent should be used proactively whenever you encounter issues during development or testing.\n\nExamples:\n\n \nContext: User encounters filter parsing error.\nuser: "The git log filter is crashing…

not rated 78k +233 today A 0 tokens original Apache-2.0

rtk-ai/rtk

Agent Claude Code

RTK testing expert - snapshot tests, token accuracy, cross-platform validation.

not rated 78k +233 today A 20 tokens original Apache-2.0

rtk-ai/rtk

Instructions file GitHub Copilot

Copilot instructions for rtk-ai/rtk, covering copilot instructions for rtk, using rtk in this session, instead of: use, build, test & lint and pre-commit gate (must all pass before any pr).

not rated 78k +233 today A 846 tokens original Apache-2.0

openai/openai-cookbook

Skill Claude CodeCodex ✓ vendor

Bootstrap a new realtime eval folder inside this cookbook repo by choosing the right harness from examples/evals/realtimeevals, scaffolding prompt/tools/data files, generating a useful README, and validating it with smoke, full eval, and test runs. Use when a user wants to start a new crawl, walk, or run realtime eval…

not rated 76k 6d ago A 76 tokens original MIT

agent-eval

45

colbymchenry/codegraph

Skill Claude CodeCodex

Benchmark CodeGraph retrieval quality on a real codebase by comparing agent behavior with vs without CodeGraph. Use when the user runs /agent-eval or asks to test, benchmark, audit, or validate a codegraph version (the local dev build or a published npm version) against a language's repo.

not rated 69k +401 2d ago A 65 tokens original MIT

review-work

46

code-yeongyu/oh-my-openagent

Skill Claude CodeCodex

Post-implementation review orchestrator. Launches 5 parallel background sub-agents: Oracle (goal/constraint verification), Oracle (code quality), Oracle (security), unspecified-high (hands-on QA execution), unspecified-high (context mining from GitHub/git/Slack/Notion). All must pass for review to pass. MUST USE…

not rated 69k +71 today A 125 tokens

qa-testing

47

openinterpreter/openinterpreter

Skill Claude CodeCodex

Verify your work by actually operating the app or website you changed, instead of assuming it works. Strongly recommended whenever you build, modify, or debug a web app, website, or desktop GUI app. Drive real browsers with the agent-browser CLI and native desktop apps with the cua-driver CLI. These are installed on…

not rated 68k +20 14d ago A 77 tokens original Apache-2.0

cline/cline

Instructions file GitHub Copilot

Copilot instructions for cline/cline, covering copilot instructions for cline, architecture, build & test (critical — non-obvious commands), protobuf rpc workflow (4 steps) and adding api providers (silent failure risk).

not rated 67k +139 today A 976 tokens original Apache-2.0

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: