Testing skills

11,702 tagged Testing, measured the same way as everything else here.

Browse within: LLM 188agentic-ai 140agents 134cli 107ai-coding 94agent 81skills 66javascript 56openai 52agentic-workflow 41agent-orchestration 40claude-code-plugin 37agent-browser 36software-architecture 35

webapp-testing

793

rockexe0000/my-awesome-copilot

Skill Claude CodeCodex

Toolkit for interacting with and testing local web applications using Playwright. Supports verifying frontend functionality, debugging UI behavior, capturing browser screenshots, and viewing browser logs.

not rated 1 8mo ago A 35 tokens copy · 100% MIT

test-execution

794

zavora-ai/skill-test-execution

Skill Claude Code

Orchestrate test execution — run unit, integration, and E2E tests, collect coverage reports, and analyze failures. Use when running tests, checking coverage, debugging test failures, or validating code changes before merge.

not rated 1 3mo ago A 48 tokens

ai-workflow

796

vindm/dotclaude

Skill Claude CodeCodex

Set up LLM workflow discipline for projects that use AI / LLM calls in production or eval suites. Authors an eval-cost-watcher agent that projects token cost BEFORE regression evals run, plus an AI-workflow-discipline rule covering mock-mode placement, fixture freshness, and multi-stage cost accumulation. Optionally a…

not rated 1 yesterday A 88 tokens original MIT

blind-ab-verify

797

aaronartistzhang-afk/DailyWork

Skill Claude CodeCodex

A workflow for comparing two versions of a prompt, skill, or piece of text without revealing which one is the new version. Separate agents create the versions, and a person scores them without seeing their labels.

not rated 1 10d ago A 305 tokens original MIT

proctor

798

catfish-1234/proctor

Skill Claude CodeCodex

Honest-completion ruleset for changes that touch tests or the code they cover. Use before deleting, skipping, renaming, or rewriting a test, before weakening an assertion, and before hardcoding or stubbing an implementation to make a test pass. Also covers what to do when a test looks genuinely wrong.

not rated 1 6d ago A 66 tokens original MIT

msw

799

anivar/msw-skill

Skill Claude Code

MSW (Mock Service Worker) v2 best practices, patterns, and API guidance for API mocking in JavaScript/TypeScript tests and development. Covers handler design, server setup, response construction, testing patterns, GraphQL, and v1-to-v2 migration. Baseline: msw ^2.15.0. Triggers on: msw imports, http.get, http.post…

not rated 1 1mo ago A 123 tokens original MIT

implement

800

orin-dx/agent-plugins

Skill Claude CodeCodex

Execute an approved plan or specification with targeted implementation, tests, evidence, review, and independent verification. Use for “implement this”, “execute the plan”, “build this feature”, “write the code”, or “refactor without changing behavior”.

not rated 1 changed 2d ago A 51 tokens original MIT

agentforge-protocol

801

Yat-mo/agentforge-protocol

Skill Claude CodeCodex

Use when doing non-trivial coding with Hermes, OpenClaw, Claude Code, Codex CLI, or similar autonomous coding agents. Orchestrates Karpathy-style minimal-change discipline, grill-plan intake, TDD, systematic debugging, subagent-driven implementation, spikes, and pre-commit review into one end-to-end workflow.

not rated 1 4mo ago A 71 tokens original MIT

code-writing

802

stepanenkoviktor0110-boop/ai-dev-methodology-codex

Skill Claude CodeCodex

A coding workflow for planning changes, writing tests before or alongside code, and reviewing the result. TDD, or test-driven development, means using tests to guide the implementation.

not rated 1 4mo ago A 77 tokens

jp-harness-tune

803

Sora-bluesky/ja-output-harness

Skill Codex

A Japanese-language workflow for tuning the rules used by ja-output-harness, a checker for unwanted words or patterns in agent output.

not rated 1 4mo ago A 86 tokens original MIT

project-test-report

804

trsiddiqui/skill-recorder-codex

Skill Claude CodeCodex

Use when the user asks to check a repository's current branch, run its documented test command, and write a concise local test report.

not rated 1 1mo ago A 31 tokens original MIT

audio-verification

805

AkshitIreddy/agent-skills

Skill Claude CodeCodex

Verify and debug synthesized audio (Web Audio API) by rendering it offline and measuring the samples, instead of guessing from code or claiming it works untested. Use whenever building, changing, or reviewing UI sound effects, tones, synths, or any AudioContext graph — and especially when sound is reported as…

not rated 1 11d ago A 117 tokens original MIT

loop-engineering

806

imMamdouhaboammar/loop-engineering-skill

Skill Claude CodeCodex

Use when designing, orchestrating, or executing autonomous multi-agent coding loops with financial budget caps, minimal fixes, and continuous dev-QA gates.

not rated 1 21d ago A 33 tokens original MIT

codexskills/agent-forge

Skill Claude CodeCodex

Test web applications in real browsers via Chrome DevTools Protocol including DOM inspection, console analysis, network profiling, performance auditing, Lighthouse CI, and visual regression.

not rated 1 3mo ago A 37 tokens

ui-apca-contrast

808

Yugoge/awesome-claude-harness

Skill Claude CodeCodex

Run APCA Lc text-contrast measurement on a Playwright page in BOTH light and dark color schemes. Returns deterministic apca. findings against rule-map.json. Use during ui-specialist Phase 6 (Accessibility).

not rated 1 yesterday A 50 tokens

verify

809

edible999999999-jpg/humhum

Skill Claude CodeCodex

Run the HUMHUM quality gates that are actually touched by the current diff — frontend typecheck + vitest, and Rust fmt/clippy/test — mirroring CI. Use before marking work done, opening a PR, or when asked to "verify", "check", or "run the gates".

not rated 1 26d ago A 62 tokens original MIT

spec-writer-setup

810

ivan-szz/spec-writer-skill

Skill Claude CodeCodex

Scans a target project to extract idiomatic patterns (test framework, code style, architecture, conventions) and generates customized TDD agents that use the project's actual patterns instead of generic pseudocode. Use when the user wants to adapt the TDD agents to a specific codebase.

not rated 1 2mo ago A 62 tokens original Apache-2.0

jrobelia/inventree-plugin-ai-toolkit

Skill Claude CodeCodex needs its repo

Run the InvenTree Plugin AI Toolkit devcontainer and execute a plugin's test-all.sh. Covers Docker setup, postCreateCommand, server startup, and the lint/unit/integration/frontend-build/E2E layers.

not rated 1 21d ago A 49 tokens original MIT

qa-runner

812

adbarc92/mcp-browser-bridge

Skill Claude CodeCodex

Use when asked to "run QA", "qa check", "test checklist", or execute QA checklists against a running app via browser automation. Requires the browser-bridge MCP server and Chrome extension.

not rated 1 10d ago A 44 tokens original MIT

adrianmikula/JakartaMigrationMCP

Skill Claude CodeCodex

This skill guides scanning the codebase for memory and performance code smells, debugging performance problems, writing performance/memory tests that follow established project conventions, and updating docs/patterns/ when new categories of issues are discovered and fixed.

not rated 1 29d ago A 0 tokens

agentation-qa-skill

814

WtecHtec/agentation-qa-skill

Skill Claude CodeCodex

A QA workflow for testing software changes and fixing reported bugs through an MCP server. QA means checking that software works as expected before release.

not rated 1 3mo ago A 67 tokens

tacoda/keystone-mcp

Skill Claude CodeCodex

Run every applicable sensor and produce a unified PASS/FAIL report.

not rated 1 2mo ago A 13 tokens original MIT archived

ptywright-testing

816

kingsword09/ptywright

Skill Codex

Build, run, record, replay, debug, and maintain deterministic terminal, TUI, PTY cassette, and browser-terminal agent regression tests with ptywright. Use when an agent needs to drive CLI/TUI apps, create ptywright scripts, configure ptywright.config., record or replay PTY output, solidify browser terminal agent flows…

not rated 1 3mo ago A 93 tokens original Apache-2.0

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: