Testing skills

11,690 tagged Testing, measured the same way as everything else here.

Browse within: LLM 179agentic-ai 140agents 130cli 103ai-coding 92agent 82skills 71javascript 57openai 52agent-browser 49agentic-workflow 41ai-testing 41agent-orchestration 40claude-code-plugin 37

autoevolve-worker

649

RightNow-AI/autoevolve

Skill Claude CodeCodex

Use this skill whenever asked to join an autoevolve run, evolve code toward a measured target, work an evolution population, mutate a candidate under EVOLVE-BLOCK rules, or report measured autoevolve progress and artifacts.

not rated 3 1mo ago A 50 tokens original Apache-2.0

traceknot

650

Jin-Doh/traceknot

Skill Claude CodeCodex

Apply Traceknot's ISTQB-aligned, evidence-bound QA process to repository changes across OMP, Codex, GajaeCode, Claude Code, and OpenCode, including session-scoped QA Board publication. Use for implementation verification, bug fixes, release checks, repository audits, defect confirmation, and residual-risk decisions…

not rated 3 7d ago A 80 tokens original MIT

rabee-elkholy/android-harness-kit

Skill Claude CodeCodex needs its repo

Use when developing business logic, UseCases, Repositories, ViewModels, or reproducing and fixing bugs using strict Red-Green-Refactor cycles. Requires writing and proving a failing test before writing implementation code.

not rated 3 8d ago A 47 tokens original MIT

api-deep-analyzer

652

oumaimah-QA/QIOS

Skill Claude CodeCodex

Deeply analyze an API endpoint and generate complete, structured test coverage. Use this skill whenever the user provides an API endpoint, a Swagger/OpenAPI spec, or describes an API to test. Triggers on: "analyze this API", "generate test cases for this endpoint", "what should I test on this API", "test coverage for…

not rated 3 1mo ago A 117 tokens original MIT

hid-bridge

653

PageMastr/AIID

Skill Claude CodeCodex

Simulate real keyboard/mouse input (Windows SendInput) and capture screenshots or short GIFs of any window — built for playtesting and validating desktop apps/games (e.g. a raylib/GLFW or similar native game window) from the CLI. Use this whenever a task requires actually driving a running Windows app with input and…

not rated 3 +1 1mo ago A 110 tokens original MIT

test-integration

654

Affitor/agent-skills

Skill Claude CodeCodex

Verify Affitor tracking pipeline end-to-end with CLI test commands. Triggers on "test tracking", "verify integration", "test affitor".

not rated 3 4mo ago A 33 tokens original MIT

hive-test

656

mattmre/EVOKORE-MCP-PUBLIC

Skill Claude CodeCodex

Iterative agent testing with session recovery. Execute, analyze, fix, resume from checkpoints. Use when testing an agent, debugging test failures, or verifying fixes without re-running from scratch.

not rated 3 3mo ago A 41 tokens original MIT

code-review

657

contentstack/contentstack-ios

Skill Claude CodeCodex

Use when reviewing PRs or before opening a PR – API design, errors, memory/threading, backward compatibility, dependencies, security, XCTest quality.

not rated 3 1mo ago A 32 tokens original MIT

crowd-test

658

anhhuyn411-alt/crowd-test

Skill Claude CodeCodex

Unleash a mob of AI virtual users (impatient shoppers, confused seniors, keyboard-only users, chaos monkeys) on a website to stress-test its UX and hunt bugs. Use when the user asks to "crowd test", "mob test", "send virtual users", or wants persona-based UX/QA feedback on a URL before launch.

not rated 2 1mo ago A 74 tokens original MIT

verify

659

BuyWhere/buywhere

Skill Claude Code

Verify the Next.js web surface from a deploy-like scratch copy when the repo's root Python app/ directory masks src/app locally.

not rated 2 today A 1 tokens

nido-vm-testing

660

Josepavese/nido

Skill Codex

Use when designing, writing, or running Nido-backed VM tests, disposable production-like QA environments, template-accelerated test labs, or isolated multi-agent AI sandboxes. Provides workflows for Nido CLI spawn/provision/upload/test/delete automation, port forwarding, templates, cleanup hygiene, and VM isolation…

not rated 2 yesterday A 69 tokens original MIT

blankfiles

661

filearchitect/blankfiles-website

Skill Claude CodeCodex

Use blankfiles.com as a binary test-file gateway: discover formats, filter by type/category, and return direct download URLs from the public API.

not rated 2 3mo ago A 32 tokens

inkcheck

662

chaoz23/inkcheck

Skill Claude CodeCodex

CI for ink interactive-fiction stories. Use it whenever a .ink file changes hands or changes state: "does my story compile?", "can any path crash it?", "are all my endings reachable?", "is there content nobody can ever see?", pre-commit checks, reviewing a story PR, or validating a generated/edited ink file before…

not rated 2 yesterday A 104 tokens original MIT

shanmukhaditya/agent-skills

Skill Codex

Operate a hierarchical software-development team of up to ten agents to understand a codebase, turn product requirements and bug reports into safe production-ready changes, and verify the result end to end. Use for multi-file feature implementation, bug fixing, refactoring, migrations, integrations, performance…

not rated 2 15d ago A 110 tokens original MIT

test-execution

664

CoffeeCheese/easy-prd-testing

Skill Claude CodeCodex

Use when confirmed planning artifacts and final execution plan are available and page verification, field comparison, evidence capture, or priority execution is requested.

not rated 2 26d ago A 31 tokens original MIT

testforge

665

whaojie797-design/Novera-AI-skills

Skill Claude CodeCodex

A test-generation tool for Python that creates runnable pytest tests. Pytest is a Python tool for writing and running automated tests.

not rated 2 1mo ago A 73 tokens

portable-skill

666

daijx-ai/skillwitness

Skill Claude CodeCodex

Validate a synthetic Agent Skill when testing SkillWitness locally or in CI.

not rated 2 1mo ago A 18 tokens original MIT

caliper

667

zhengbowenai-cmd/caliper

Skill Claude Code

Statistically test whether a prompt or SKILL.md change is actually better than the old version. Use when the user asks to compare two prompts, A/B test a prompt change, check if a recent edit really improved things, find rule conflicts in a long prompt, or identify which sections of a prompt are pulling weight.…

not rated 2 2mo ago A 93 tokens original MIT

tester

668

erikfiala/e2e-tester

Skill Cursor

Run end-to-end web quality audits with Playwright and Lighthouse using existing project scripts first. Use when the user says /tester, e2e test, smoke test, Playwright audit, Lighthouse audit, performance audit, accessibility audit, dark mode audit, mobile audit, or asks for 100/100 Lighthouse improvement guidance.

not rated 2 5mo ago A 67 tokens original MIT

mcp-com-ai/mcp-server-evaluations-skills

Skill Claude Code

Test MCP servers for quality and reliability. Verify tool functionality, test error handling, generate tests, and assess response quality with no dependencies other than curl. Use this when validating MCP server implementations, testing OpenAPI-to-MCP conversions, or assessing API tool quality.

not rated 2 7mo ago A 59 tokens original MIT

evolve-skill

671

taneltaluri/evolve-skill

Skill Claude CodeCodex

Evolve Skill: measurement-first skill optimizer. Evaluates SKILL.md files against an anchored 9-dimension rubric, validates that the rubric itself is stable (test-retest), optimizes with a hill-climbing loop that only accepts improvements larger than measurement noise, protects against overfitting with train/holdout…

not rated 2 4mo ago A 128 tokens original MIT

skill-check

672

JckJhns/skill-check

Skill Claude Code

Comprehensive testing and validation of Claude skills. Use this skill whenever the user wants to test, validate, audit, or quality-check a skill — whether they say "test my skill", "check this skill works", "validate my skill", "run skill-check", or anything similar. Also trigger when the user asks things like "does…

not rated 2 4mo ago A 192 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: