Testing skills

11,704 tagged Testing, measured the same way as everything else here.

Browse within: LLM 188agentic-ai 140agents 134cli 107ai-coding 94agent 81skills 65javascript 57openai 52agentic-workflow 41agent-orchestration 40agent-browser 36claude-code-plugin 36software-architecture 35

danish54/kiwi-tcms-mcp-server

Skill Claude CodeCodex

Generate business-level test scenarios from a feature branch and push them to Kiwi TCMS. Use when asked to "generate tests for TICKET-XXXX", "push tests to Kiwi", or "update tests after PRD change".

not rated 1 2mo ago A 52 tokens

code-implementation

818

w693847022/memory_service

Skill Claude Code needs its repo

A coding workflow for implementing an existing feature or bug fix from a written plan, then adding unit tests and running integration tests.

not rated 1 3mo ago A 32 tokens original MIT

ie-migration-vrt

819

rayven122/mcp-ie-migration-vrt

Skill Codex

Compare a legacy page rendered in Microsoft Edge IE mode with its migrated Chromium Edge page, diagnose visual differences, fix the migrated implementation, and repeat VRT. Use when working on IE-to-Edge migrations with the mcp-ie-migration-vrt MCP server, especially when both browser sessions must be operated into…

not rated 1 8d ago A 74 tokens original MIT

Nolane-x/Nolane-habitat

Skill Codex

Maintain, debug, test, document, package, or release the Nolane Habitat repository. Trigger when work changes Habitat workspace lifecycle, storage, semantic providers, mutation safety, MCP, Observatory, tests, Codex integration, or release artifacts.

not rated 1 8d ago A 56 tokens

git-asmt-repo

821

Canvas-LMS-MCP/canvas-teacher-mcp

Skill Claude CodeCodex

GLOBAL, language-agnostic L1 abstract for building + testing a GitHub coding-assignment repo through the org-hub autograder. Owns the WHAT (audit → instruction-fidelity harness → prove 100/100 → ship the Starter → gates) with NO language mechanics. The HOW-to-compile/run/assert is delegated to an L2 language skill…

not rated 1 24d ago A 126 tokens original MIT

aar-operations

822

phenomenoner/adaptive-agent-harness

Skill Codex

Operate the Adaptive Agent Runtime through its public MCP tools. For tasks that need tool use or analysis, consider AAR MCP early when bounded stateful computation, brokered jobs, operations, contracts, or artifacts can materially help; software planning, development, testing, and troubleshooting belong to this class…

not rated 1 7d ago A 145 tokens original MIT

ras-end-to-end

823

Zhonghao1995/Agentic-HEC-RAS

Skill Claude CodeCodex

Standard operating procedure for an auditable, headless HEC-RAS workflow via the hec-ras MCP server — how to plan, copy, edit boundaries, run, QA-gate, read results, compare, plot, audit, and safely stop. Use FIRST whenever an agent is handed a HEC-RAS project (.prj folder) or a results file (.p##.hdf) plus a…

not rated 1 23d ago A 113 tokens original MIT

bongo-test

824

emicyx/bongocat-mcp

Skill Claude CodeCodex

A skill for testing the BongoCat MCP connection by checking its response, reading status, changing an expression by name, and sending a speech bubble.

not rated 1 14d ago A 44 tokens

agent-mailbox

825

stumct/agent-mailbox

Skill Codex

Create and use disposable, token-scoped email addresses through a configured Agent Mailbox MCP server. Use for testing signups, email verification, magic links, password resets, one-time codes, inbound messages, attachments, and transactional email flows.

not rated 1 5d ago A 52 tokens original MIT

capture-failure

826

jiangkoumo/agenttape

Skill Claude CodeCodex

Turn a captured Codex tool failure into reviewed .tape evidence and an offline regression test. Use when the user asks to inspect a failed run, fork captured evidence, save a regression, or prepare AgentTape evidence for CI.

not rated 1 9d ago A 50 tokens original MIT

ilkuru/mcp-server-antigravity-hyperv

Skill Claude CodeCodex

Standard operating procedure for safely deploying, testing, validating, and debugging PowerShell and CMD scripts inside isolated Hyper-V virtual machines using MCP tools.

not rated 1 17d ago A 34 tokens original Apache-2.0

hronaut

828

hronaut/hronaut

Skill Claude CodeCodex

Use the Hronaut desktop Browser MCP for visible, persistent web workflows in isolated agent workspaces. Use for authenticated browser QA with human handoff, localhost or responsive testing, accessibility and performance diagnosis, or tasks that should survive one agent session while the user watches or takes over.

not rated 1 changed today A 59 tokens

improve-test-coverage

829

Hyeonu-Cha/dotnet-coverage-mcp

Skill Claude CodeCodex

Iteratively write .NET unit tests to improve code coverage toward 80% line and branch targets. USE FOR: autonomously running tests, finding gaps, writing targeted tests, and repeating until 80% coverage or plateau. Calls all 7 dotnet-coverage-mcp MCP tools: GetSourceFiles, RunTestsWithCoverage, GetCoverageSummary…

not rated 1 2d ago A 139 tokens original MIT

invarianteval

830

AlpharomeroJL/invarianteval

Skill Claude CodeCodex

Before finalizing structured or extracted output in a domain with declared safety invariants, verify the output against your invariant suite and refuse to ship when a locked field was model-auto-filled. Invoke when the user or task involves high-stakes structured extraction, compliance fields, or pass/fail results…

not rated 1 2mo ago A 67 tokens original MIT

device-farms

831

almasumdev/awesome-mobile-testing-agent-skills

Skill Claude CodeCodex

Expert guidance on running mobile tests on Firebase Test Lab, AWS Device Farm, BrowserStack App Automate, and Sauce Labs. Use when asked to set up device-farm coverage, design a device matrix, or compare vendors.

not rated 1 4mo ago A 49 tokens

scrcpy-mcp

832

1999AZZAR/scrcpy-mcp

Skill Claude CodeCodex

Professional Android automation agent for device interaction, UI testing, and app management. Use for: (1) interacting with Android apps, (2) filling inputs, (3) structural UI inspection via XML/axtree, (4) complex multi-step mobile tasks. Prefer XML/axtree over screenshots — images are expensive tokens.

not rated 0 3d ago A 71 tokens

mcp-tester-helper

833

Pstasov/mcp-tester-helper

Skill Claude CodeCodex

Rules for AI-assisted infrastructure testing via mcp-tester-helper MCP server. Activate when the user asks to make HTTP requests to services (REST API), search logs (OpenSearch), or run SQL queries against databases (PostgreSQL) across different stands (environments).

not rated 0 1mo ago A 60 tokens original MIT

gauntlet

834

oscarsterling/clelp-skills

Skill Claude CodeCodex

Adversarial hardening loop for security-critical code, especially hooks and guards that must never fail open. Two models from different labs attack the same code for concrete, reproducible failures; an orchestrator adjudicates; a separate fixer closes the enumerable class; repeat until both models return zero…

not rated 0 2d ago A 81 tokens original MIT

skills-eval

835

shipeasy-ai/shipeasy

Skill Claude Code

How to run the behavioural skills-eval — drive headless claude -p with prompts and assert the right Shipeasy skill fires, the right MCP tools get called, and the resulting server state is correct. Covers the self-contained eval:fresh run (wipes + seeds its own isolated DB per run), the manual-backend eval, the…

not rated 0 25d ago A 163 tokens

verified-run

836

runvouch/runvouch

Skill Claude Code

Wrap a scheduled OpenClaw task in a RunVouch run. Reports start and end, attaches the output file as evidence, and alerts on Telegram, Slack or e-mail when the task is missed, fails, or ends without the file it was supposed to write.

not rated 0 yesterday B 57 tokens original MIT

artifact-verifier

837

Dailyaiagents/daily-ai-agent-toolkit

Skill Claude CodeCodex

Verify that declared deliverable files exist, are non-empty, remain inside an allowed root, and avoid declared placeholder terms.

not rated 0 4d ago A 28 tokens original Apache-2.0

hwcontract

838

MohibShaikh/hwcontract

Skill Claude CodeCodex

Use when firmware has timing or level requirements you can measure against a spec: pulse widths on a WS2812/NeoPixel strip, DShot ESC, servo, I2C, NEC IR, DS18B20/DHT sensor, A4988/DRV8825 stepper, PWM fan, HC-SR04, when a board prints a boot log over serial (ESP32, ESP8266, Zephyr, MicroPython, u-boot, Raspberry Pi…

not rated 0 15d ago A 149 tokens original MIT

beforeusersdo-qa

839

bhuman-ai/QAbro

Skill Codex

Route and run BeforeUsersDo QA through its MCP tools. Use when a user asks to test, QA, review, or get feedback on a website, app, feature, preview, or user flow; asks for AI QA, a personal self-review, a real human tester, QA status, video, transcript, findings, or a completed BeforeUsersDo report.

not rated 0 17d ago A 79 tokens

Caravaca-Labs/puzzletide-cli

Skill Claude CodeCodex

Use this skill when the user wants verifiable reasoning tasks to benchmark or test an LLM or agent — reproducible puzzle task sets (sudoku, word search) with objective, by-construction grading. No answer key to trust: answers are verified against the rules and the grid.

not rated 0 1mo ago A 65 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: