Testing

18,481 mods in this category, of every kind an agent can take. Each one carries what it costs per session, what the scan found, and whether it is the original.

flowproof-config

1009

automators-com/flowproof

Skill Claude CodeCodex

Configure flowproof's SAP GUI, Fiori, and AI authoring credentials by walking the user through flowproof config sap / flowproof config fiori / flowproof config ai. Use when the user wants to set up, change, or check their SAP/Fiori login or model authoring key, or when a flow run fails because SAPUSER, FIORIPASSWORD…

not rated 5 yesterday A 106 tokens original Apache-2.0

prestashop-pr-qa

1010

PrestaShop/skills

Skill Claude CodeCodex

QAs a PrestaShop pull request against an environment that is already running, in a real browser, on the command line or over HTTP, and writes an HTML report stating whether it is approved, with the recording as proof. Use when the user says "QA this PR", "test this pull request", "check that this fix works"…

not rated 5 4d ago A 90 tokens AFL-3.0

autonomous-build

1011

bnet47/codexicon

Skill Claude CodeCodex needs its repo

Execute a local task register with requirement tracing, verification, review, and refinement.

not rated 5 changed today A 20 tokens original MIT

kpubdata-dataset

1012

yeongseon/kpubdata

Skill Claude CodeCodex

A set of instructions for adding and repairing datasets in kpubdata, a project that describes public data APIs with specification files. It covers recording real API responses, replaying examples, validating schemas, and updating documentation.

not rated 5 today A 42 tokens original MIT

wopee-mcp

1013

Wopee-io/wopee-mcp

Plugin Cursor

Bundles 1 MCP server

MCP server for autonomous end-to-end testing with Wopee.io. Analyzes web applications, generates and executes Playwright-based functional tests, and validates results.

not rated 5 yesterday A tokens not measured original MIT

dsh-chaos-test

1014

cyanseek/dsh-tool-chaos

Skill Codex

Design, install, run, and report deterministic DeepSeek Harness tool-failure experiments. Use when a user wants to prove retry or fallback behavior, timeout or cooperative cancellation, policy-denial handling, blocked-result recovery, Code Mode nested-call resilience, or CI evidence for a DSH agent/plugin. Complete…

not rated 4 +1 15d ago B 91 tokens original MIT

live-audit

1015

Grinv/steam-games-mcp

Skill Claude CodeCodex

Audit steam-games-mcp — build/test/lint gate, live MCP tool edge-case sweep (input validation, SteamID64/vanity/appid edge cases, key-gating), and source-level code review. Use when asked to test/audit the published or just-fixed steam-games-mcp package, hunt for bugs/edge cases, or repeat "the same kind of testing as…

not rated 4 18d ago A 83 tokens original MIT

sniff

1016

Aboudjem/sniff

Plugin Claude Code

Bundles 3 skills, 1 MCP server · 151 tokens together

Autonomous QA scanner via sniff-qa: walks your running app in a real browser and reports real bugs (broken pages/links, console/network errors, broken forms, state-loss, empty/placeholder data, bad loading/error states, responsive + a11y) with reproduction proof, severity, confidence, and a fix. No API key.

not rated 4 changed 8d ago A tokens not measured original Apache-2.0

qai-marketing

1017

gvasile29/qai-consultant

Skill Claude CodeCodex

Acts as the QAI Consultant marketing/PR specialist. Use whenever Gabi asks to write, draft, or plan a social media post (LinkedIn, Facebook, Instagram) promoting the QAI Consultant app or MCP server - release announcements, educational QA tips, milestones, case studies, or behind-the-scenes. Covers Romanian and…

not rated 4 2d ago A 96 tokens

test-intel

1018

barissozudogru/test-intel-mcp

MCP server Claude CodeCodexCursor +2

Coverage analysis, untested function detection and complexity scoring for TS and JS. Runs locally from the @barissozudogru/test-intel-mcp npm package.

not rated 4 16d ago A tokens not measured original MIT

goodeye

1019

Goodeye-Labs/goodeye-cli

MCP server Claude CodeCodexCursor +2

Design, save, and run outcome-aligned AI workflows and verifiers, with reliable image output. Remote server at mcp.goodeye.dev.

not rated 4 1mo ago A tokens not measured original MIT

five46

1020

sekharsdet/five46

MCP server Claude CodeCodexCursor +2

BYOK, fully local AI agent that tests your app/API and writes a real Playwright spec on success. Runs locally from the five46 npm package. Needs 5 environment variables to run.

not rated 4 14d ago A tokens not measured original MIT

temper-skills

1021

CyrilLeMat/temper-skills

Skill Claude Code

Give an agent skill's decision logic a test suite and a deterministic implementation — a labeled validation dataset plus versionable Python — via an adversarial multi-persona loop that runs natively in Claude Code: proposer plus persona subagents through the Task tool, entirely on your Claude Code subscription, no…

not rated 4 1mo ago A 173 tokens original Apache-2.0

figma-design-skills

1022

jeltehomminga/figma-design-skills

Plugin Claude Code

Bundles 2 skills · 193 tokens together

Two composable skills for design-to-code fidelity: figma-design-extract pulls exact specs from Figma into a build-ready spec table; design-fidelity-verify proves the running web or mobile app matches them by measuring rendered values.

not rated 4 +1 2mo ago A tokens not measured original MIT

craft-skills

1023

mikestangdevs/craft-skills

Plugin Claude Code

Bundles 17 skills · 2,250 tokens together

Skills for code you'll have to live with. Two collections: craft (code that's understandable, named well, lean, loud about failure, and testable — without turning every edit into a rewrite) and rigor (claims you can believe — adversarial testing, full-suite banking, cited numbers, population sweeps, evidence-backed…

not rated 4 3mo ago A tokens not measured original MIT

akirtok/preflight-security-audit

Plugin Claude Code

Bundles 1 skill, 1 command, 6 agents · 979 tokens together

Comprehensive 360 pre-ship audit: 61 checks across security, reliability, performance, AI/LLM, privacy/compliance, and launch readiness — including live-database, serverless/edge, GraphQL, client-side, and modern-attack (Trojan Source) checks. Reports a 0–100 security score with severity-ranked findings and proposes…

not rated 4 1mo ago A tokens not measured original MIT

basteez/java-skills

Skill Claude Code

Generate integration tests for a Java class using Testcontainers. Supports Spring Boot (3.1+), Quarkus (3.0+), and Micronaut (4.0+). Detects class type (repository, controller, service) and applies framework-appropriate test patterns. Requires Java 17+.

not rated 4 5mo ago A 69 tokens

off-by-none

1026

NTCHz/off-by-none

Plugin Claude Code

Bundles 1 skill · 47 tokens together

Spec-faithful implementation: spec-derived tests, mandatory boundary coverage, and an evidence-gated done-report. Every rule TDD-tested against baseline failures on Opus, Sonnet, and Haiku.

not rated 4 1mo ago A tokens not measured original MIT

harness-design

1027

zanwei/harness-design

Skill Claude CodeCodex

Quality harness for design-dna Phase 3 output with browser-based visual verification. After an agent generates a design from a Design DNA JSON + user content, this skill acts as a verification and scoring layer — collecting all page resources via console/network inspection, performing section-by-section screenshot…

not rated 4 5mo ago A 232 tokens original MIT

skill-testing

1028

wangwei1237/agent-skills

Skill Codex

Test Agent Skills from tests/skillcases.yaml files. Use when validating whether a Skill triggers correctly, produces a dry-run plan, satisfies output contracts, handles edge/failure cases, or when asked to run Skill Test cases, judge Agent responses against YAML expectations, or produce Skill test reports.

not rated 4 2mo ago A 61 tokens original Apache-2.0

emmanuelperu/microcks-skills

Skill Claude CodeCodex needs its repo

Write OpenAPI examples that work with Microcks dispatchers for API mocking. Covers example pairing rules, JSON body dispatching, Groovy script dispatching, and dispatcher configuration via the Microcks API.

not rated 4 5mo ago A 47 tokens original Apache-2.0

genesis

1030

gabrieldabbah/genesis

Plugin Claude Code

Bundles 10 skills, 1 hook · 1,280 tokens together

Three modes for Claude Code projects: create a new one from an empty folder (research and design before stack, then build and test it), transition an existing repository to a stated standard, and check the user-level /.claude layer that loads in every session. English-first, with secret denials and human gates on anyt.

not rated 4 1mo ago A tokens not measured original MIT

minottobot

1031

EmanueleMinotto/minottobot

Plugin Claude Code

Bundles 8 skills · 1,258 tokens together

Use whenever the user asks about QA, testing strategy, CI/CD health, team processes, developer experience, code review practices, test coverage, flaky tests, monitoring, or any audit of an engineering team's quality practices.

not rated 4 13d ago A tokens not measured original MIT

e2e-alertmanager-test

1032

conallob/o11y-analysis-tools

Skill Claude CodeCodex needs its repo

Render end-to-end previews of what an alert notification will actually look like (plain-text email, HTML email, Slack attachment JSON, raw webhook JSON) by replaying a Prometheus unit-test file's expected alerts through a live Alertmanager. Effectively a "print preview" for alerts. Use to review notification…

not rated 4 1mo ago A 153 tokens original BSD-3-Clause

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: