Testing skills

16,346 tagged Testing, measured the same way as everything else here.

Browse within: agent-browser 89ai-coding 82ai-testing 80javascript 52agentic-coding 50openclaw 49agentic-workflow 47openai 47android 44claude-code-plugin 44cypress 41static-analysis 38skill-scanner 35agentic-framework 34

netbox-mcp-testing

241

netboxlabs/netbox-mcp-server

Skill Claude CodeCodex

This skill should be used when systematically testing the NetBox MCP server after code changes or to validate tool functionality. Provides structured protocol for discovering and testing MCP tools with a live NetBox instance, with emphasis on performance monitoring and comprehensive reporting.

not rated 221 +2 7d ago A 53 tokens original Apache-2.0

Accessibility Auditor

242

PramodDutta/qaskills

Skill Claude CodeCodex

Comprehensive WCAG 2.1 AA compliance testing combining automated axe-core scans with manual keyboard navigation, screen reader compatibility, and focus management verification.

not rated 214 5d ago A 32 tokens original MIT

cashubtc/cashu.me

Skill Claude CodeCodex

Record, assemble, validate, and present watchable Playwright E2E demo videos for Cashu.me. Use this skill whenever a user asks to watch, record, share, or make a video or montage of wallet browser tests, payment flows, mint or melt operations, ecash transfers, or existing Playwright recordings, even if they do not…

not rated 213 8d ago A 83 tokens original MIT

go

244

inference-gateway/inference-gateway

Skill Claude CodeCodex

Idiomatic Go - package and interface design, error wrapping, table-driven tests, generics, the modern standard library (slices/maps/cmp/errors.Join), current syntax, and logging discipline. Use when writing, reviewing, or refactoring any Go code, especially code drifting toward Java/Spring shapes (deep layer trees…

not rated 210 5d ago A 99 tokens original Apache-2.0

ffuf-web-fuzzing

245

jthack/ffuf_claude_skill

Skill Claude CodeCodex

Expert guidance for ffuf web fuzzing during penetration testing, including authenticated fuzzing with raw requests, auto-calibration, and result analysis.

not rated 210 10mo ago A 34 tokens

quality-control

246

aronprins/paperclip-company-playbook

Skill Claude CodeCodex

A mandatory verification skill for CEO agents. Ensures every deliverable is personally verified before reporting to the Founder.

not rated 202 5mo ago A 0 tokens

vectorbt-expert

247

marketcalls/vectorbt-backtesting-skills

Skill Claude CodeCodex

VectorBT backtesting expert. Use when user asks to backtest strategies, create entry/exit signals, analyze portfolio performance, optimize parameters, fetch historical data, use VectorBT/vectorbt, compare strategies, position sizing, equity curves, drawdown charts, or trade analysis. Also triggers for openalgo.ta…

not rated 201 +1 1mo ago A 85 tokens

prove-feature

248

searlsco/prove_it

Skill Claude CodeCodex

Create a temporary real project and prove a proveit feature works (or doesn't) end-to-end. Builds a disposable git repo, writes a focused config, runs real dispatches through the installed or local proveit, and produces a human-readable session transcript. Use when you need to prove a feature, reproduce a bug, or…

not rated 199 4mo ago A 81 tokens original MIT

evaluate-run

249

adrianco/retort

Skill Claude CodeCodex

Evaluate a single retort experiment run. Score the generated code against the task's TASK.md requirements, run its build and tests, compute metrics, and emit a structured evaluation report plus a machine-readable findings file.

not rated 199 5d ago A 45 tokens original Apache-2.0

mirroir-onboard

250

jfarcand/mirroir-mcp

Skill Claude CodeCodex

Onboard a consumer web app to mirroir's .mirroir/ dotfile by EXPLORING the running app (chrome-devtools-mcp) — derive real selectors from the accessibility tree, exercise each surface's primary action, emit the .mirroir/ tree, and validate by LIVE REPLAY with a self-heal loop. Reject shallow "page renders" coverage.

not rated 199 +1 5d ago A 85 tokens original Apache-2.0

e2e-driving

251

SystemSculpt/obsidian-systemsculpt-ai

Skill Claude CodeCodex

Drive the real Obsidian GUI from the CLI for SystemSculpt plugin testing — click, type, submit, attach files, change settings, run scripted scenarios, and read live UI state deterministically via data-testid targets. Use whenever testing plugin UI behavior, verifying a UI change in the live app, reproducing a user…

not rated 193 18d ago A 96 tokens original MIT

test-with-gt

252

gollem-dev/gollem

Skill Claude CodeCodex

Write Go test code using the gt library. Use when writing tests, creating test files, or when the user asks to add tests for Go code.

not rated 193 +1 5d ago A 35 tokens original Apache-2.0

tdd-methodoly-expert

253

thedaviddias/skill-check

Skill Claude CodeCodex

Use when implementing features or fixing bugs with strict Test-Driven Development (TDD). Activate for coding tasks that require new functionality, refactoring, or comprehensive test coverage, especially when the user mentions TDD or Test Driven Development.

not rated 188 3mo ago A 53 tokens original MIT

noqa-testing

254

noqa-ai/noqa

Skill Claude CodeCodex

Use this skill when the user wants to boot and interact with iOS or Android devices/simulators — inspect the screen, execute actions, generate or edit test cases, or run UI tests via the noqa platform.

not rated 185 1mo ago A 47 tokens original MIT

mattgierhart/PRD-driven-context-engineering

Skill Claude CodeCodex

Execute implementation within EPICs following test-first development, continuous SoT updates, and code traceability during PRD v0.7 Build Execution. Triggers on requests to start building, implement an epic, begin coding, or when user asks "start building", "implement epic", "coding", "development", "build execution"…

not rated 182 4d ago A 113 tokens original MIT

qa

256

arcee-ai/nac

Skill Claude CodeCodex

Run scalable, isolated live QA for nac development. The top-level local orchestrator must parse n (default 4), dispatch one setup worker with this skill, copy its n assignment contracts verbatim into exactly n parallel test workers with this skill, then dispatch one aggregate worker with this skill using all test…

not rated 201 yesterday A 100 tokens original Apache-2.0

android

257

yang1ming/android-harness

Skill Claude CodeCodex

Direct Android device control through ADB. Use for authorized device automation, testing, screenshots, UI inspection, and app interaction.

not rated 174 1mo ago A 27 tokens original MIT

test-fixer

258

mitchdenny/hex1b

Skill Claude CodeCodex

Agent for diagnosing and fixing flaky terminal UI tests in the Hex1b test suite. Use when tests pass locally but fail in CI, or when tests exhibit timing-sensitive behavior.

not rated 173 6d ago A 39 tokens original MIT

tdd

259

wbern/agent-instructions

Skill Claude CodeCodex

Remind agent about TDD approach and continue conversation.

not rated 168 3mo ago A 13 tokens original MIT

device-test

260

Mahdi-mortazavi/relay

Skill Claude CodeCodex

Run Relay against the phone plugged into this laptop and the Windows app installed on it, find what breaks, and fix it. Use when the user asks to test on real hardware, test on their phone, check pairing on their own Wi-Fi, or verify a release on their own machine. Only meaningful in a session running locally on the…

not rated 171 2d ago A 83 tokens GPL-3.0

implement-feature

261

tddworks/SkillsManager

Skill Claude CodeCodex

Guide for implementing features following architecture-first design, TDD, rich domain models, and Swift 6.2 patterns. Use this skill when: (1) Adding new functionality to a Swift app (2) Creating domain models that follow user's mental model (3) Building SwiftUI views that consume domain models directly (4) User asks…

not rated 165 3mo ago A 103 tokens

dev-cycle

262

xberg-io/crawlberg

Skill Claude CodeCodex

Crawlberg iteration loops codified as Taskfile tasks — alef install/generate/format/bump, core and binding builds, e2e generate/build/test cycles, cleanup tiers, and the mock-server / stale-.so / precompiled-NIF / generated-e2e gotchas. Load when running or debugging crawlberg build, alef regeneration, or e2e…

not rated 168 yesterday A 89 tokens original MIT

dogfood

263

paiml/paiml-mcp-agent-toolkit

Skill Claude CodeCodex

Dogfood pmat — rebuild, install, exercise every CLI command against pmat's own repo, check output integrity + self-quality, find next work. Read-only audit; files issues for bugs.

not rated 164 yesterday A 40 tokens original MIT

qa

264

eric-tramel/slop-guard

Skill Claude CodeCodex

Black-box QA audit of slop-guard across MCP, CLI, fit, docs, agent workflows, and writing-effectiveness. Files GitHub issues for real problems found.

not rated 163 1mo ago A 37 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: