Testing skills

11,673 tagged Testing, measured the same way as everything else here.

Browse within: LLM 179agentic-ai 140agents 130cli 103ai-coding 92agent 82skills 71javascript 57openai 52agent-browser 49agentic-workflow 41ai-testing 41agent-orchestration 40claude-code-plugin 37

rigor

673

olimxonuz0-lab/rigor

Skill Claude CodeCodex

Use for any non-trivial coding, engineering, or deliverable-producing task — building a feature, fixing a bug, refactoring, designing an architecture, writing a script someone will run, or drafting a document someone will use. Enforces upfront planning before acting, rejects placeholder/stub/TODO code and unhandled…

not rated 2 20d ago A 181 tokens original MIT

browser-test-executor

675

dangnhit/qa-tester

Skill Claude CodeCodex

Execute approved bounded browser Test DSL cases with fresh isolated contexts and auditable attempts. Use when running browser tests, reruns, regression checks, or blocked execution diagnostics.

not rated 2 28d ago A 38 tokens original Apache-2.0

Coverage Guard

676

rolecraft-sh/skills

Skill Claude CodeCodex

Use when the user wants to check test coverage, enforce 100% coverage, find uncovered code, add missing tests, or increase code coverage. Works with vitest, jest, react-scripts, and other test runners. Also for 'coverage', 'test coverage', 'cover', 'untested', 'uncovered', 'add tests for', 'increase coverage'…

not rated 2 21d ago A 90 tokens original MIT

pg-skill-forge

677

gaoguo/pg-skill-forge

Skill Claude Code

A toolkit for creating, testing, and improving skills for the opencode coding agent. A skill is a reusable set of instructions for a particular workflow.

not rated 2 2mo ago A 111 tokens original MIT

tcr

678

xpepper/tcr-skill

Skill Claude CodeCodex

Guide users through TCR (Test && Commit || Revert), TCRDD, and git-gamble workflows. ALWAYS trigger when a user mentions TCR, TCRDD, "test commit revert", "git gamble", or "git-gamble". Trigger when a user wants to combine TDD with automatic commits/reverts, enforce baby steps via a commit-or-revert loop, or asks…

not rated 2 5mo ago A 164 tokens original MIT

qa-testing-kit

679

SoftwareOneHN/qa-testing-kit

Skill Claude CodeCodex

Lifecycle-driven QA workflow — from requirements analysis to test reports, with state tracking and impact analysis.

not rated 2 2mo ago A 5 tokens original MIT

xorcise-playbooks

680

xorcise-ai/xorcise-skills

Skill Claude CodeCodex

Run a standardized XORCISE agent benchmark — pick a ready-made playbook (or build your own from existing missions), point it at any OpenHands-supported model(s), and get a polished eval-card HTML report of how each model performed across the missions. Warns you about cost before spending anything.

not rated 2 1mo ago B 66 tokens

jpbaking/playwright-fieldkit

Skill Codex

Explore, debug, audit, compare, record, and test live websites with deterministic Playwright scripts and QE workflows. Use for requests to map a site, find broken pages or links, reproduce browser bugs, discover hidden or role-gated features, audit accessibility/performance, compare crawls, design or review test cases…

not rated 2 1mo ago A 143 tokens original MIT

mx-generate

682

Claritune/mutantx

Skill Claude Code

MutantX Phase 2 — Generate realistic code mutants as unified diff patches.

not rated 2 1mo ago A 20 tokens original MIT

ios-dev

683

AlphaSquadTech/ios-dev

Skill Claude Code

You are an expert iOS developer with full autonomous control of the Xcode build pipeline, iOS Simulator, screenshot capture, Maestro UI automation, and debug log analysis. Follow these procedures exactly.

not rated 2 6mo ago B 95 tokens original MIT

cloud

684

ShiplightAI/agent-skills

Skill Claude CodeCodex

Sync local tests with Shiplight cloud — push and pull YAML test cases, templates, and functions between your repo and the cloud. Requires a Shiplight cloud subscription.

not rated 2 2mo ago A Socket: warnSnyk: pass 35 tokens original MIT archived

playwright-cli

685

sonofmagic/skills

Skill Claude Code

Automate browser interactions, test web pages and work with Playwright tests.

not rated 2 changed 3d ago A 19 tokens copy · 100% MIT

wp-mutate

686

soderlind/wordpress-agent-plugin

Skill Claude CodeCodex

Run mutation testing on WordPress plugins and themes to find weak tests — Pest --mutate or Infection for PHP, StrykerJS for JavaScript — then triage surviving mutants into concrete test improvements.

not rated 2 changed 3d ago A 45 tokens

audit-then-fanout-fix

687

AllanWessels/Bratan

Skill Claude CodeCodex

When a class of bug keeps slipping through tests, run a coverage-matrix audit first, then fan out fix agents per gap — don't whack-a-mole individual failures.

not rated 2 2mo ago A 44 tokens

using-shakespii

688

ai-creed/ai-shakespii

Skill Claude CodeCodex

Use when the user asks to lint, audit, test, benchmark, validate, or fix an agent skill — from a single SKILL.md frontmatter check to trigger-accuracy measurement or a corpus-wide audit of installed skills for duplication — driving the shakespii CLI (init, lint --json, test --run, bench) to resolve findings until…

not rated 2 1mo ago A 78 tokens original MIT

direct

689

hraness/direct

Skill Codex

Use Hraness Direct to install, adopt, test, audit, or troubleshoot deterministic frontend and UI testing workbenches for web, React, React Native, and Expo. Trigger for repeatable signed-in, empty, loading, and error states; frontend fixtures and scenario URLs; strict JSON worlds; product-owned ports and adapters…

not rated 2 changed 2d ago A 111 tokens original MIT

ax-verification

690

shenli/devtool-ax-kit

Skill Claude CodeCodex

Build independent, deterministic verifiers for agent-tool tasks.

not rated 2 8d ago A 15 tokens original Apache-2.0

aipdlc

691

mbmd/AIPDLC

Skill Claude CodeCodex

AIFLC (AI Full Life Cycle) PDLC (Product Development Life Cycle) Family — 11 injectable workflow packages that guide AI coding agents through professional software delivery: idea evaluation → project initiation → portfolio governance → product ownership → UX design → architecture → workspace generation → compliance →…

not rated 2 changed 2d ago C 64 tokens

qa

692

0xherve/qa-skill

Skill Claude Code

Assist a QA engineer by navigating the app, drafting test cases from requirements, executing them, and recording results — with the human in the loop for judgment.

not rated 2 2mo ago A 33 tokens original MIT

supervisor

693

ma-nucho-pro/supervisorLLM-plugin

Skill Claude CodeCodex

Universal multi-agent quality supervisor and release gate for Claude Code, Cursor, Codex, Gemini CLI, ChatGPT/Codex Agent Skills, and other Agent Skills-compatible harnesses. Use for rigorous research, coding, frontend design, testing, anti-hallucination review, independent judges, adversarial review, evidence-based…

not rated 2 24d ago A 82 tokens

verify

694

yusufalikync/ccs

Skill Claude Code needs its repo

Run full project verification — smoke tests, CLI checks, and code quality.

not rated 2 5mo ago A 16 tokens original MIT

test-ui

695

rahulcvwebsitehosting/wayfinder

Skill Claude Code

Test the Wayfinder agent extension UI by starting the dev environment and visually verifying changes via CDP. Covers the new tab page (left sidebar — Home, Scheduled Tasks, Settings, etc.) and the right side panel (chat interface). Use after making UI changes to apps/agent/.

not rated 2 2mo ago A 60 tokens AGPL-3.0

pursr

696

0xheycat/pursr

Skill Claude CodeCodex

Use Pursr for browser screenshots, scripted visual operation, visual regression, accessibility audits, DOM inspection, and MCP-driven browser sessions. Use when a user asks an agent to inspect a site, operate an existing browser session, fill or draft UI content, record a tutorial, compare visuals, debug layout, or…

not rated 2 1mo ago A 85 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: