Testing skills

11,742 tagged Testing, measured the same way as everything else here.

Browse within: LLM 180agents 133agentic-ai 129cli 101ai-coding 99agent 80skills 72javascript 57agent-browser 54openai 51ai-testing 46agentic-workflow 41agent-orchestration 40claude-code-plugin 40

dxkit-action

457

vyuh-labs/dxkit

Skill Claude Code

Read a dxkit report and execute fixes — prioritize findings by severity, plan the fix sequence, run the fix, verify the score moved, re-baseline if appropriate. Supports a SCOPED pass to burn down one category at a time (dependency/BOM vulnerabilities, security, code quality, tests, docs), and a BASELINE-CLEANUP pass…

not rated 10 8d ago A 201 tokens original MIT

gptadmin-mcp-testing

458

megamen32/gptadmin_opensource

Skill Claude CodeCodex

Verify a Memos-style MCP integration through GPTADMIN without duplicating the backend, leaking credentials, or confusing the native ShellMCP supervisor with legacy per-agent services.

not rated 10 changed today A 40 tokens AGPL-3.0

periscope

459

segentic-lab/periscope-mcp

Skill Claude CodeCodex

Drive the Periscope MCP server (74 Playwright + headless-Chrome tools) to test, audit, and debug websites and web apps — static sites, SPAs, and apps behind a login. Use this skill whenever the user wants to test a website or web app, run an E2E or interactive flow, audit accessibility/SEO/GEO/ performance, measure…

not rated 10 1mo ago A 183 tokens AGPL-3.0

pyats

460

automateyournetwork/pyATS_skills

Skill Claude CodeCodex

Use this skill for any pyATS/Genie/Unicon network test-automation task — testbeds, secrets, learn/parse/execute/configure, parallel operations (pcall), test authoring (AEtest, Blitz, Robot Framework, Genie triggers/verifications), device reset (Clean), mock/recorded devices, REST connectivity, containers, Webex…

not rated 10 16d ago C 102 tokens original Apache-2.0

sdd-verify

461

Alan-TheGentleman/gentle-ai-enterprise

Skill Claude CodeCodex

Validate that implementation matches specs, design, and tasks. Trigger: When the orchestrator launches you to verify a completed (or partially completed) change.

not rated 10 6mo ago A 35 tokens

testing-kit

462

fermonterom/claude-testing-kit

Skill Claude Code

Unit + E2E testing + quality gate para Next.js con Claude Code. Combina TDD workflow, Vitest patterns, Playwright E2E, validacion de build, env vars, y security check en una sola skill. Se activa al crear/modificar route.ts, page.tsx, o archivos de test. Para cualquier proyecto Next.js con App Router.

not rated 10 5mo ago A 78 tokens original MIT

drupal-behat-test

463

trebormc/drupal-ai-agents

Skill Claude Code needs its repo

Generates Behat tests for Drupal 10/11 using the Drupal Extension for Behat. Use this skill when the project already has Behat configured (look for behat.yml), when acceptance tests written in natural language (Gherkin) are needed, E2E flows, or when the client needs to read and validate test scenarios. Trigger…

not rated 10 1mo ago A 158 tokens original Apache-2.0

unit-tests

464

WordPress/core-contributor-skills

Skill Codex

Test quality auditor that reviews existing test suites—audits, deletes, rewrites, and fills genuine gaps. Use when asked to review tests, improve test quality, or audit a test suite. Does NOT blindly add tests.

not rated 10 10d ago A 48 tokens GPL-2.0

add-minimum-job

465

henryiii/skills

Skill Claude CodeCodex

Add a minimum version test job to a noxfile.

not rated 10 +1 24d ago A 16 tokens original MIT

bug-hunting

466

m4vic/bug-hunting

Skill Claude CodeCodex

Bug bounty hunting and penetration testing skills for Claude, Codex, and other agentic AI tools.

not rated 10 +2 11d ago A 24 tokens original MIT

ifBars/blender-agent-studio

Skill Codex

Benchmark Blender modeling agents, skills, prompts, scripts, or MCP tools with paired isolated runs. Use for baseline-versus-plugin comparisons, regression suites, skill forward-testing, MCP usefulness evaluation, score calibration, or claims that a Blender workflow improves mesh, visual, structural, animation, cost…

not rated 12 +4 19d ago A 68 tokens original MIT

e2e-acceptance

468

Justinmendezai/The-Adam-Repo

Skill Claude CodeCodex

Final end-to-end acceptance pass after all slices are merged: full tests-first suite, Playwright CLI when the packet defines e2e, extra verification.layers commands (API/integration/contract), code-review-graph integrity, and browser MCP walks for UI success criteria — without using Playwright for backend depth. Use…

not rated 11 +1 12d ago A 87 tokens original Apache-2.0

qa-runner

469

dynatrace-oss/dynatrace-snowflake-observability-agent

Skill OpenCode needs its repo

AI-guided QA walkthrough for DSOA releases. Automates version detection, deployment commands, notebook deployment, and interactive test walkthrough. Use when a QA engineer needs to execute the DSOA release test suite.

not rated 10 3d ago A 47 tokens original MIT

writing-skills

470

axiomantic/spellbook

Skill Claude CodeCodex

Use when creating new skills, editing existing skills, or verifying skills work before deployment. Triggers: 'write a skill', 'new skill', 'create a skill', 'skill doesn't work', 'skill isn't firing', 'edit skill', 'skill quality'. NOT for: general prompt improvement (use instruction-engineering) or command creation…

not rated 10 2d ago A 77 tokens original MIT

Forge

471

ikennaokpala/forge

Skill Claude CodeCodex

Autonomous quality engineering swarm that forges production-ready code through continuous behavioral verification, exhaustive E2E testing, and self-healing fix loops. Combines DDD+ADR+TDD methodology with BDD/Gherkin specifications, 7 quality gates, defect prediction, chaos testing, and cross-context dependency…

not rated 9 5mo ago A 91 tokens original MIT

manta-e2e-smoke

472

antoinedc/MantaUI

Skill Claude CodeCodex

Drive MANTA's built Electron app in a real renderer context (Playwright's electron launcher) and assert that key UI surfaces render correctly — no crash, no blank screen, sidebar/chat/terminal present. Load BEFORE marking any frontend/UI task done, and when manta-pr-workflow or manta-handle-reviewer-return verifies a…

not rated 10 3d ago A 102 tokens original MIT

pr-test-recommender

473

srinidhis05/agentura

Skill Claude CodeCodex

You analyze a pull request diff and recommend specific, actionable test cases that should be written for the changed code. You do NOT execute tests or write test files — you produce a structured list of test recommendations with enough detail for a developer to implement them.

not rated 9 1mo ago A 5 tokens original Apache-2.0

btt

474

sablier-labs/agent-skills

Skill Claude CodeCodex

Write bulloak tree specifications (.tree files) for smart contract integration tests. Trigger phrases - write a tree, create test tree, BTT spec, bulloak tree, Branching Tree Technique, or when writing integration tests for contract functions.

not rated 9 29d ago A 53 tokens original MIT

science-integrity

475

leventilo/mobius

Skill Claude CodeCodex

Parallel critic agent that enforces physical invariants on Mobius SimSpecs and simulation outputs. Runs five bundled deterministic checks — units, CFL stability, conservation laws, figure-diff, and claim-match — and blocks or warns upstream generators. Use whenever a SimSpec or simulation artifact is produced.

not rated 9 4mo ago A 62 tokens original MIT

courtneyr-dev/wp-release-audit-method

Skill Claude CodeCodex needs its repo

Hand off WordPress audit or testing work as ONE coworker-ready or SRE-ready markdown document, with privately-disclosed findings scrubbed to a bare acknowledgment. Use when asked to "hand off" a WordPress audit, make a handoff/shareable doc from testing findings, package audit results for a coworker/SRE/stakeholder…

not rated 9 today A 84 tokens GPL-2.0

verify-feature

479

exiao/pm-skills

Skill Claude CodeCodex

Verify a code change actually works by building/running the app and observing it at its real surface (CLI, API, UI, library, agent), capturing runtime evidence rather than trusting tests. Make sure to use this skill whenever the user has changed code and wants to know it works, is about to merge/push and wants…

not rated 9 1mo ago A 163 tokens original Apache-2.0

e2e

480

prkharueh12/playwright-cli-custom-agents

Skill Claude Code

E2E test generation skill using Playwright CLI with Page Object Model pattern and visual regression testing via @visual tag.

not rated 9 2mo ago A 28 tokens

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: