Testing

18,481 mods in this category, of every kind an agent can take. Each one carries what it costs per session, what the scan found, and whether it is the original.

hetzner-ansible-lab

1297

Bitbull-Ideas/hermes.skills

Skill Claude CodeCodex

Provision temporary Hetzner Cloud VMs for Ansible role QA, bootstrap OS prerequisites such as Rocky 8 Python 3.9, run verified work, clean up cloud resources, and report sanitized evidence.

not rated 2 2mo ago A 51 tokens AGPL-3.0

eval-skills

1298

JarvixGaby/eval-skill

Skill Claude CodeCodex

Evaluate, benchmark, or test-drive an unfamiliar AI agent skill, tool bundle, prompt workflow, or capability package. Use when the user asks whether a downloaded/shared .skill, SKILL.md, agent workflow, or reusable AI capability actually helps, is worth installing, works as advertised, performs better than asking an…

not rated 2 1mo ago A 114 tokens original MIT

alltest

1299

Epsilondelta-ai/alltest

Skill Claude CodeCodex

Full project test coverage. Analyzes the entire codebase, identifies coverage gaps, and writes tests to achieve 100% coverage (minimum 80%). Ensures every source file with exportable logic has at least one test. Use when the user mentions 'alltest', wants comprehensive testing, or asks for full test coverage.

not rated 2 6mo ago A 69 tokens original MIT

mvp-test-strategist

1300

mkashyap00/business-idea-claude-skills

Skill Claude Code

Selects the optimal testing mechanism, designs the minimum viable test, and actively executes the creation of the test assets (e.g., landing pages, scripts) using tool calling. Reads the GTM plan and outputs a live test environment and execution timeline.

not rated 2 4mo ago A 57 tokens

scenario-qa

1301

zzangmin2/product_skills

Skill Claude CodeCodex

An automated review of a user flow shown in a PNG image or Figma design. It creates a Markdown document covering the understood flow, missing scenarios, edge cases, and UI/UX improvements.

not rated 2 3mo ago A 136 tokens

dstest

1303

bxrne/dstest

Skill Claude CodeCodex

Deterministic simulation testing for containerized services. Write Lua scripts to inject chaos (pause, kill, resource deprivation) into Docker containers with reproducible, seeded fault injection. Use when writing chaos experiments, testing service resilience, or debugging distributed systems.

not rated 2 changed 8d ago A 53 tokens original MIT

aoa-evals-skills

1304

8Dionysus/aoa-evals

Skill Codex

Route the aoa-evals skill family for central proof selection, review, evolution, named results or verdicts, source-linked reports, Eval Forge owner review, and proof lifecycle. Hand repository-local eval selection, application, intake/design, or session-hit classification to aoa-eval. Candidates, readiness checks…

not rated 2 6d ago A 84 tokens original Apache-2.0

tdd

1305

JustinThomas2/agentrc

Skill Claude CodeCodex

How to write good tests - behavior-focused assertions, the AAA structure, red/green/refactor, small vertical slices, and what to mock versus leave real. Use whenever writing a new test, modifying or fixing an existing test, reviewing someone else's tests, deciding whether something needs a test, choosing what to mock…

not rated 2 1mo ago A 76 tokens original MIT

refactor-py

1306

meshulga/agents-doc

Skill Claude Code

Apply standard Python refactoring patterns and re-run pytest.

not rated 2 4mo ago A 16 tokens original MIT

andrejkaxz/Vanessa_for_AI

Skill Codex

A skill for creating, editing, checking, and diagnosing Russian-language Turbo Gherkin `.feature` files for Vanessa Automation, a tool for testing 1C applications. It covers UI and export scenarios without MCP.

not rated 2 16d ago A 105 tokens

careerchain-ys/stdd

Skill Claude Code

A guide for documenting and testing an existing software feature or page by studying how its code already works. It creates requirement, design, and test documents that describe the current behavior rather than an imagined future version.

not rated 2 1mo ago A 137 tokens original Apache-2.0

e2e-tests

1309

GlamgarOnDiscord/claude-saas-blueprint

Skill Claude Code needs its repo

Tests E2E Playwright pour SaaS Next.js : setup, Page Object Model, auth state, flows critiques (login, billing, onboarding), CI GitHub Actions.

not rated 2 3mo ago A 40 tokens original MIT

ping-pong-tdd

1310

danethurber/.dotfiles

Skill Claude CodeCodex

Pair on an implementation via ping-pong TDD — alternating red/green rounds between the agent and the user. Use when the user says "ping-pong" or asks to alternate writing failing tests and making them pass.

not rated 2 8d ago A 51 tokens

browser-tester

1311

eliottvalette/Anthropic-Voodoo-Hackathon

Agent Claude Code

Headless browser tester and root-cause diagnostician for the proto-pipeline-m single-file HTML playables. Runs Playwright on the playable, captures the state trajectory and console errors, cross-references symptoms with the source sketches in 04sketches.json and the 04plan.json contract, and returns a structured…

not rated 2 4mo ago A 147 tokens

test-writer

1312

fschifone/claude-setup-wizard

Agent Claude Code

Writes unit and integration tests for new or modified code. Use this agent when the user asks to add test coverage, test-drive a feature, or when a PR lacks tests.

not rated 2 4mo ago A 39 tokens original MIT

matchers

1313

romilly/claude-code-helpers

Command Claude Code

Command "matchers" from romilly/claude-code-helpers, covering pyhamcrest custom matchers guide, when to create a custom matcher, spotting opportunities, before - verbose and redundant length check and after - expressive, no length check needed.

not rated 2 7mo ago A 0 tokens original MIT

atlas-benchmark

1315

madaeroblade/atlas

Skill Claude CodeCodex

Benchmark whether Atlas measurably improves coding-agent behavior, by running every scenario twice — with and without Atlas — in clean contexts and scoring the pair blind.

not rated 2 1mo ago A 35 tokens original MIT

qa-recorder

1316

azam-sdet/10xquality

Cursor rule Cursor

QA Recorder - Records user's manual browser interactions and generates natural language test scripts with rich locator metadata.

not rated 2 6mo ago A 18 tokens original MIT

n-test

1317

nitra/7n-rules

Cursor rule Cursor

A test-configuration rule for JavaScript and Rust projects. It places JavaScript tests in tests/ and controls Vitest, Stryker, coverage, mutation testing, and Rust mutation-testing configuration.

not rated 2 11d ago A 0 tokens archived

qa-tester

1318

karkranikhil/sf-ai-toolkit

Agent Claude Code

Creates test strategy, Apex test classes, LWC Jest tests, validation checklists, and regression checklists. Ensures code coverage, edge cases, and security scenarios are covered.

not rated 2 4mo ago A 40 tokens original MIT

code-verifier

1320

1466094598lilye-byte/cursor-coding-workflow

Skill Claude CodeCodex

An independent check of code written by a code executor against success criteria defined by a task decomposer. It runs test commands and produces a structured pass-or-fail report, but does not write or repair code.

not rated 2 6mo ago A 72 tokens

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: