Testing skills

11,794 tagged Testing, measured the same way as everything else here.

Browse within: LLM 186agentic-ai 140agents 133cli 109ai-coding 94agent 81skills 65javascript 56openai 52agent-orchestration 41agentic-workflow 41claude-code-plugin 35software-architecture 35dotnet 32

kanban-review-tester

865

jopa79/kanban-mcp

Skill Claude CodeCodex

Testet alle Kanban-Tasks im Status "Review" automatisch mit Chrome/Playwright und VHS (Terminal). Erkennt automatisch ob ein Task Browser- oder Terminal-Tests braucht. Liest Task-Notes, startet Server wenn nötig, testet visuell und funktional, und dokumentiert Ergebnisse in den Task-Notes. Nutze diesen Skill wenn JoPa…

not rated 0 5d ago A 104 tokens original MIT

qa_audit_pro

866

0xC1pher/MCP-HUB-V11

Skill Claude CodeCodex

Esta skill transforma al agente en un Ingeniero de QA especializado en Yari Medic. Su misión es auditar código contra los documentos de grounding en data/projectcontext/.

not rated 0 3mo ago A 0 tokens

songofhawk/playwright-mcp-tabbed

Skill Claude CodeCodex

Parent Agent uses playwright-tabbed to allocate browser tabs and partition tasks, then spawns child agents in parallel — each operating on its assigned tabid — for browser acceptance testing or automation, with a final aggregated summary. Suitable for multi-page validation, parallel module checks, and…

not rated 0 5mo ago A 96 tokens original MIT

browser-test

868

l0s3r-Q/browser-test-mcp

Skill Claude CodeCodex

A browser automation testing skill combining Playwright's precise controls with browser-use's AI-guided exploration. It operates a shared Chromium browser session.

not rated 0 1mo ago A 119 tokens original MIT

glassbox

869

adihebbalae/glassbox

Skill Claude CodeCodex

Verify and debug your own web UI on localhost. Use when you changed frontend code and need to check it actually works — console/network errors, layout/contrast/a11y problems, dead buttons, wrong colors, broken handlers — or when testing a dev server, reproducing a UI bug, or debugging why an element behaves wrong.…

not rated 0 4d ago A 101 tokens original MIT

prespec

870

emretheus/prespec

Skill Claude CodeCodex

Write the behaviour specification for a feature BEFORE implementing it, as test cases. Use whenever you are about to build a new endpoint, screen, flow, or behaviour — especially anything touching auth, sessions, tokens, lists, or pagination. Also use when asked "how should I test this", "what should this actually…

not rated 0 1mo ago A 91 tokens original MIT

json-comparator

871

EE-WILL-I/simple-data-comparator-mcp

Skill Claude CodeCodex

Compares JSON actual vs expected template via the validate-json MCP tool. Use when validating API responses, JSON configs, nested objects, JSON Schema contracts, or when the user asks to compare, diff, or validate JSON data.

not rated 0 28d ago A 49 tokens original MIT

security-qa-tester

872

PurrlyDigital/owasp-wstg-mcp

Skill Claude CodeCodex

Security QA tester workflow for the owasp-wstg MCP server — scope, search, neighbor-walk, and citation over the WSTG corpus using five MCP tools. Load this file first; then load the model-variant file that matches your model family.

not rated 0 1mo ago A 54 tokens CC-BY-SA-4.0

api-test-design

873

mdkulkarni2005/osmos-marketing-mcp

Skill Claude CodeCodex

Generic, service-agnostic API test scenario design skill. Consumes an API contract from a Documentation MCP (or any authoritative contract source) and produces traceable, prioritized, coverage-reported test scenarios. Works for any REST/HTTP endpoint of any service/version — no hardcoded API knowledge.

not rated 0 10d ago A 63 tokens

mobile-testing

874

mobile-agent-platform/mobile-agent-platform

Skill Claude CodeCodex

A skill for running step-by-step tests on a real Android phone, using screen-reading and device-control tools while checking the result after each action.

not rated 0 29d ago A 83 tokens original Apache-2.0

hyprland-tester

875

Keylessboi/hyprland-mcp

Skill Claude CodeCodex

Automated desktop UI testing through the Hyprland MCP server. Captures screenshots, analyzes them via the vision skill, clicks and types, and re-verifies. Use when asked to test a desktop app, verify UI behavior, run a visual regression check, or drive a background app without interrupting the user. Triggers: "test…

not rated 0 25d ago A 0 tokens

telegraph-miner-tests

876

ashishwarkhade/Telegraph-MCP

Skill Claude CodeCodex

Use when running, extending, or diagnosing batch tests against the Telegraph miner or chatbot API, including JSON query validation, resumable JSONL collection, request failures, and coverage summaries.

not rated 0 12d ago A 42 tokens original MIT

generate-test-files

877

Zulelee/test-files-mcp

Skill Claude CodeCodex

Generate local test fixtures via Test Files MCP. Use when the user needs dummy files, images, PDFs, CSVs, JSON, ZIPs, WAV audio, corrupted files, or edge-case filenames for upload testing, QA, parser testing, or validation.

not rated 0 20d ago A 55 tokens original MIT

mcp-eval

878

GSA-TTS/mcp-hackathon-template

Skill Claude CodeCodex

Guide for building a Phoenix-based evaluation harness for an MCP server. Use when you need to create or extend LLM-as-judge evaluations that measure how well an agent can accomplish realistic tasks using only the MCP server's tools. Covers the eval/phoenix module layout, dataset creation, the LangChain agent, judges…

not rated 0 14d ago A 76 tokens original MIT

razor-dvara-evals

879

Utkarsh-Sinha0/razor-dvara

Skill Claude CodeCodex

Run the eval harness over 50+ synthetic COD orders and produce batch metrics with honest assumptions. Use when validating gate decisions or preparing evidence for the buildathon submission.

not rated 0 18d ago A 40 tokens original MIT

vunit-mcp

880

ru551n/vunit-mcp

Skill Claude CodeCodex

Drive a VUnit (HDL unit-testing) project through the vunit-mcp MCP server (vunitstatus, vunitlisttests, vunitlistfiles, vunitcompile, vunitelaborate, vunitruntests, vunitgetreport, vunitgettestlog, vunitgettestwaveform, vunitexportjson, vunittestdependencies). Use when the user asks to run or compile VUnit tests, find…

not rated 0 changed yesterday A 0 tokens original MPL-2.0

agent-eval-gate

881

dxwss/agent-eval-gate

Skill Claude CodeCodex

Inspect agent JSONL traces and run deterministic evaluation gates.

not rated 0 18d ago A 17 tokens original MIT

spec-code-review

882

ronronner02/codepilot-agent

Skill Claude Code

Structured code review for bugs, regressions, tests, and standards. Report-only by default; apply fixes only when the current user or upstream caller explicitly requests review-and-fix. mode:agent is always report-only.

not rated 0 8d ago A 48 tokens original MIT

BongSuCHOI/memex

Skill Claude CodeCodex

Produce a coverage-checked report of the entire Memex conversation corpus when the user asks to analyze, organize, or summarize all Codex history. Use deterministic totals first and label unfinished backfill honestly.

not rated 0 2d ago A 47 tokens original MIT

jterratsdev/ableton-live-mcp

Skill Claude CodeCodex

Design deterministic failure scenarios that prove workflows, APIs, providers, gates, budgets, and regulated flows degrade safely.

not rated 0 15d ago A 0 tokens original MIT

verify-before-pr

886

thechrisgrey/thechrisgrey

Skill Claude CodeCodex

Run the full verification suite required before finishing any change to the thechrisgrey codebase. Covers lint, typecheck, tests, formatting, and AGENTS.md validation.

not rated 0 yesterday A 39 tokens

ticket-test-runner

887

hcarrillo001/retrieval-mcp

Skill Claude CodeCodex

Runs an end-to-end test from a Jira ticket number. Reads the ticket, executes the browser test steps it contains with Playwright, optionally scores any AI or text output with the RetriEval MCP, and writes a pass/fail result back to the ticket. Use this WHENEVER the user gives a ticket key and asks to test, verify, QA…

not rated 0 13d ago A 139 tokens original Apache-2.0

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: