Testing skills

11,763 tagged Testing, measured the same way as everything else here.

Browse within: LLM 180agents 133agentic-ai 129cli 101ai-coding 99agent 80skills 72javascript 57agent-browser 54openai 51ai-testing 46agentic-workflow 41agent-orchestration 40claude-code-plugin 40

901-feature-loop

433

ajelinek/SpecFlow

Skill Claude Code

Use 901 to run one feature end-to-end through SpecFlow's full pipeline as a single git-tracked run: high-level design, happy-path specs, scenario-by-scenario test-driven build-out validated and committed per module, cleanup, an optional expanded-coverage round, and a merge back to the base branch. Every design/spec…

not rated 11 9d ago A 160 tokens

y3-game-test

434

pirronewantlux529-coder/y3autohotfreshskill

Skill Claude CodeCodex

A testing workflow for y3 games, with three mandatory rules: review code before sending it, treat only [run-ok] as success, and ensure every run or Lua command reaches the command-ID success marker.

not rated 11 6mo ago A 45 tokens

flowforge-testing

435

Remon-16/flow-forge

Skill Claude CodeCodex

Generate, modify, validate, execute and triage Flow Forge API test cases. Use when the user wants to turn requirement documents, API documents, table structures or business rules into executable YAML/Excel test cases, revise existing cases after requirement/API changes, run cases with the Flow Forge executor and…

not rated 11 1mo ago A 90 tokens original MIT

refactor

436

marcoemrich/agentic_coding_lab

Skill Claude CodeOpenCode

TDD Refactor Phase - Improve code while keeping tests green using Simple Design Rules and APP mass.

not rated 11 today A 22 tokens original MIT

simlease

437

alexissan/simlease

Skill Claude CodeCodex

Use when building, running, or testing an iOS app on a simulator — claims a leased simulator through simlease so parallel AI agents never collide, labels the simulator window with the task name, and releases it when done.

not rated 11 1mo ago A 48 tokens original MIT

agora-bug-detection

438

lebronlambert/Agora

Skill Claude CodeCodex

Hunt for bugs in distributed consensus protocol implementations using a 2-agent strategy+testgen pipeline. Use when targeting consensus protocols (Raft, PBFT, HotStuff, Tendermint, etc.) for automated vulnerability discovery.

not rated 11 3mo ago A 53 tokens

praman-sap-cli

440

mrkanitkar/playwright-praman

Skill Claude Code

SAP UI5 test automation via Playwright CLI. Use when testing SAP Fiori apps, discovering UI5 controls, debugging Praman tests, or automating SAP workflows. Extends playwright-cli with SAP/UI5 awareness.

not rated 11 2d ago A 50 tokens original Apache-2.0

test-skill

441

Bin-hy/EasyCoding

Skill Claude CodeCodex

An end-to-end test skill for checking whether the skill system works.

not rated 11 1mo ago A 15 tokens

verify

442

intervalrain/OpenClaw.Net

Skill Claude CodeCodex

Adversarial verification agent that tests implementations by running builds, tests, and probing edge cases. Returns a structured VERDICT (PASS/FAIL/PARTIAL) with evidence.

not rated 11 5mo ago A 37 tokens original MIT

palm-dev

443

clintonthegeek/pose64

Skill Claude CodeCodex

Use when developing or testing Palm OS applications in the POSE64 emulator, including using MCP tools, automating emulator actions, or writing Palm app test scripts.

not rated 11 2mo ago A 36 tokens GPL-3.0

mass_verify

444

intimatep/PenTestClaw

Skill Claude CodeCodex

A batch verification process that checks proof-of-concept items against a list of assets. It runs the checks in parallel and produces a results matrix and Markdown report.

not rated 11 5mo ago A 27 tokens

samurai

445

zerosixty/samurai

Skill Claude CodeCodex

Samurai scoped testing framework for Go (github.com/zerosixty/samurai). MUST use when writing, modifying, reviewing, or debugging Go tests that import samurai, or reference samurai.Run, samurai.RunWith, samurai.Scope, samurai.TestScope, samurai.W, or samurai.BaseContext.

not rated 11 3mo ago A 69 tokens original MIT

babysit-pr

446

chemany/Mente

Skill Codex

Babysit a GitHub pull request after creation by continuously polling review comments, CI checks/workflow runs, and mergeability state until the PR is merged/closed or user help is required. Diagnose failures, retry likely flaky failures up to 3 times, auto-fix/push branch-related issues when appropriate, and keep…

not rated 11 3mo ago A 114 tokens original MIT

tdd

447

chriswritescode-dev/opencode-forge

Skill Claude CodeCodex

Test-driven development with red-green-refactor loop. Use when user wants to build features or fix bugs using TDD, mentions "red-green-refactor", wants integration tests, or asks for test-first development.

not rated 11 today A 45 tokens copy · 86% MIT

code-assist

448

brazil-bench/pourpoise

Skill Claude CodeCodex

This sop guides the implementation of code tasks using test-driven development principles, following a structured Explore, Plan, Code, Commit workflow. It balances automation with user collaboration while adhering to existing package patterns and prioritizing readability and extensibility. The agent acts as a…

not rated 11 7mo ago A 105 tokens original Apache-2.0

performance-testing

449

felipepg22/FPGSkills

Skill Codex

Plan, generate, execute, analyze, and report k6 performance tests for REST/HTTP and gRPC endpoints, journeys, or whole applications. Use for load, stress, spike, soak, scalability, latency, throughput, and service performance, including approved mutations and remote non-production targets. Excludes production…

not rated 11 +2 changed yesterday A 85 tokens original MIT

rcampos09/performance-testing-skills

Skill Claude CodeCodex

Guides developers and testers in writing, fixing, and structuring Gatling load test scenarios. Use this skill whenever the user mentions load testing, performance testing, stress testing, Gatling, virtual users, VUs, ramp-up, injection profiles, simulations, JMeter migration, k6 migration, throughput, response time…

not rated 10 5mo ago A 99 tokens

bf-test-workflow

451

bryan-gu/E2E_TestSKILL

Skill Claude CodeCodex

BF project testing workflow for Claude Code. Use when generating or maintaining UI automation testing assets from requirements documents or UI exploration, including BF project initialization, sprint0 full test generation, sprintN incremental updates, test case Excel generation, Playwright E2E script generation…

not rated 10 2mo ago A 74 tokens

does-it-work

452

TsakunovR/does-it-work

Skill Claude CodeCodex

A testing and quality-checking toolkit for running an application, finding bugs, and creating automated tests. It supports API tests and browser-based UI tests in Python or Java.

not rated 10 6d ago A 295 tokens original MIT

eval

453

julienamorgan/signal-prospecting-kit

Skill Claude Code

Test harness for the Signal Prospecting Kit v4.1. Runs deterministic checks and rubric-based grading against all 6 skills including tool connections, multi-batch flow, LinkedIn outreach, channel selection, self-improvement loop, enrichment tiers, bridge/CTA variation, variation planning, and fallback tool detection.…

not rated 10 3mo ago A 99 tokens

quality-assurance

454

LetheChen/openclaw-longmen-inn

Skill Claude CodeCodex

A software quality and release-checking guide for reviewing code, testing applications, checking performance, and verifying delivery materials.

not rated 10 5mo ago A 23 tokens original MIT

release

455

dynobox/dynobox

Skill Claude CodeCodex

Prepare dynobox packages for release to npm. Use this skill whenever the user asks to release, publish, ship, bump, or cut a version of any dynobox package — including dry runs, version bumps, changelog updates, and git tagging. Also trigger when the user asks about the release process or wants to verify publish…

not rated 10 4d ago A 71 tokens original Apache-2.0

cc-port-live-e2e

456

Ling-ye/cc-port

Skill Codex

Run and audit CC Port's opt-in live Windows package E2E and broader native remaining-scope validation. Use when validating that a packaged installer or release candidate can enable AI automation, upload/download a harmless Skill through packaged MCP plus desktop approval, verify Registry v1 and Git blob bytes, or when…

not rated 10 18d ago A 136 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: