Skill Claude CodeCodex
Generate business-level test scenarios from a feature branch and push them to Kiwi TCMS. Use when asked to "generate tests for TICKET-XXXX", "push tests to Kiwi", or "update tests after PRD change".
11,704 tagged Testing, measured the same way as everything else here.
Browse within: LLM 188agentic-ai 140agents 134cli 107ai-coding 94agent 81skills 65javascript 57openai 52agentic-workflow 41agent-orchestration 40agent-browser 36claude-code-plugin 36software-architecture 35
Skill Claude CodeCodex
Generate business-level test scenarios from a feature branch and push them to Kiwi TCMS. Use when asked to "generate tests for TICKET-XXXX", "push tests to Kiwi", or "update tests after PRD change".
Skill Claude Code needs its repo
A coding workflow for implementing an existing feature or bug fix from a written plan, then adding unit tests and running integration tests.
rayven122/mcp-ie-migration-vrt
Skill Codex
Compare a legacy page rendered in Microsoft Edge IE mode with its migrated Chromium Edge page, diagnose visual differences, fix the migrated implementation, and repeat VRT. Use when working on IE-to-Edge migrations with the mcp-ie-migration-vrt MCP server, especially when both browser sessions must be operated into…
Skill Codex
Maintain, debug, test, document, package, or release the Nolane Habitat repository. Trigger when work changes Habitat workspace lifecycle, storage, semantic providers, mutation safety, MCP, Observatory, tests, Codex integration, or release artifacts.
Canvas-LMS-MCP/canvas-teacher-mcp
Skill Claude CodeCodex
GLOBAL, language-agnostic L1 abstract for building + testing a GitHub coding-assignment repo through the org-hub autograder. Owns the WHAT (audit → instruction-fidelity harness → prove 100/100 → ship the Starter → gates) with NO language mechanics. The HOW-to-compile/run/assert is delegated to an L2 language skill…
phenomenoner/adaptive-agent-harness
Skill Codex
Operate the Adaptive Agent Runtime through its public MCP tools. For tasks that need tool use or analysis, consider AAR MCP early when bounded stateful computation, brokered jobs, operations, contracts, or artifacts can materially help; software planning, development, testing, and troubleshooting belong to this class…
Skill Claude CodeCodex
Standard operating procedure for an auditable, headless HEC-RAS workflow via the hec-ras MCP server — how to plan, copy, edit boundaries, run, QA-gate, read results, compare, plot, audit, and safely stop. Use FIRST whenever an agent is handed a HEC-RAS project (.prj folder) or a results file (.p##.hdf) plus a…
Skill Claude CodeCodex
A skill for testing the BongoCat MCP connection by checking its response, reading status, changing an expression by name, and sending a speech bubble.
Skill Codex
Create and use disposable, token-scoped email addresses through a configured Agent Mailbox MCP server. Use for testing signups, email verification, magic links, password resets, one-time codes, inbound messages, attachments, and transactional email flows.
Skill Claude CodeCodex
Turn a captured Codex tool failure into reviewed .tape evidence and an offline regression test. Use when the user asks to inspect a failed run, fork captured evidence, save a regression, or prepare AgentTape evidence for CI.
ilkuru/mcp-server-antigravity-hyperv
Skill Claude CodeCodex
Standard operating procedure for safely deploying, testing, validating, and debugging PowerShell and CMD scripts inside isolated Hyper-V virtual machines using MCP tools.
Skill Claude CodeCodex
Use the Hronaut desktop Browser MCP for visible, persistent web workflows in isolated agent workspaces. Use for authenticated browser QA with human handoff, localhost or responsive testing, accessibility and performance diagnosis, or tasks that should survive one agent session while the user watches or takes over.
Hyeonu-Cha/dotnet-coverage-mcp
Skill Claude CodeCodex
Iteratively write .NET unit tests to improve code coverage toward 80% line and branch targets. USE FOR: autonomously running tests, finding gaps, writing targeted tests, and repeating until 80% coverage or plateau. Calls all 7 dotnet-coverage-mcp MCP tools: GetSourceFiles, RunTestsWithCoverage, GetCoverageSummary…
Skill Claude CodeCodex
Before finalizing structured or extracted output in a domain with declared safety invariants, verify the output against your invariant suite and refuse to ship when a locked field was model-auto-filled. Invoke when the user or task involves high-stakes structured extraction, compliance fields, or pass/fail results…
almasumdev/awesome-mobile-testing-agent-skills
Skill Claude CodeCodex
Expert guidance on running mobile tests on Firebase Test Lab, AWS Device Farm, BrowserStack App Automate, and Sauce Labs. Use when asked to set up device-farm coverage, design a device matrix, or compare vendors.
Skill Claude CodeCodex
Professional Android automation agent for device interaction, UI testing, and app management. Use for: (1) interacting with Android apps, (2) filling inputs, (3) structural UI inspection via XML/axtree, (4) complex multi-step mobile tasks. Prefer XML/axtree over screenshots — images are expensive tokens.
Skill Claude CodeCodex
Rules for AI-assisted infrastructure testing via mcp-tester-helper MCP server. Activate when the user asks to make HTTP requests to services (REST API), search logs (OpenSearch), or run SQL queries against databases (PostgreSQL) across different stands (environments).
Skill Claude CodeCodex
Adversarial hardening loop for security-critical code, especially hooks and guards that must never fail open. Two models from different labs attack the same code for concrete, reproducible failures; an orchestrator adjudicates; a separate fixer closes the enumerable class; repeat until both models return zero…
Skill Claude Code
How to run the behavioural skills-eval — drive headless claude -p with prompts and assert the right Shipeasy skill fires, the right MCP tools get called, and the resulting server state is correct. Covers the self-contained eval:fresh run (wipes + seeds its own isolated DB per run), the manual-backend eval, the…
Skill Claude Code
Wrap a scheduled OpenClaw task in a RunVouch run. Reports start and end, attaches the output file as evidence, and alerts on Telegram, Slack or e-mail when the task is missed, fails, or ends without the file it was supposed to write.
Dailyaiagents/daily-ai-agent-toolkit
Skill Claude CodeCodex
Verify that declared deliverable files exist, are non-empty, remain inside an allowed root, and avoid declared placeholder terms.
Skill Claude CodeCodex
Use when firmware has timing or level requirements you can measure against a spec: pulse widths on a WS2812/NeoPixel strip, DShot ESC, servo, I2C, NEC IR, DS18B20/DHT sensor, A4988/DRV8825 stepper, PWM fan, HC-SR04, when a board prints a boot log over serial (ESP32, ESP8266, Zephyr, MicroPython, u-boot, Raspberry Pi…
Skill Codex
Route and run BeforeUsersDo QA through its MCP tools. Use when a user asks to test, QA, review, or get feedback on a website, app, feature, preview, or user flow; asks for AI QA, a personal self-review, a real human tester, QA status, video, transcript, findings, or a completed BeforeUsersDo report.
Skill Claude CodeCodex
Use this skill when the user wants verifiable reasoning tasks to benchmark or test an LLM or agent — reproducible puzzle task sets (sudoku, word search) with objective, by-construction grading. No answer key to trust: answers are verified against the rules and the grid.
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: