Testing

18,324 mods in this category, of every kind an agent can take. Each one carries what it costs per session, what the scan found, and whether it is the original.

perf

625

QAInsights/perf-skills

Plugin Claude Code

Bundles 1 skill · 172 tokens together

Expert guidance for performance testing: planning, scripting, executing, and analyzing load/stress/soak/spike tests using JMeter, k6, Gatling, Locust, Artillery, NeoLoad, LoadRunner, OctoPerf, plus LLM inference benchmarking and SLO/capacity planning.

not rated 15 10d ago A tokens not measured original MIT

xops-agent

626

wentf9/xops-cli

Skill Claude CodeCodex

Automate server O&M, execute batch commands, test network connectivity, transfer files, and manage host assets using the XOps CLI toolkit. Triggers when managing server clusters, performing concurrent tasks across multiple machines, checking network ports, or synchronizing configuration files.

not rated 15 +5 yesterday C ✓ AI review 56 tokens original MIT

vetto-sandbox

627

shleder/vetto

Skill Codex

Enforce zero-daemon Landlock/Seatbelt security boundaries, network isolation, and subagent capability controls when executing untrusted commands or running subagents. Use when running terminal commands, testing untrusted scripts, isolating AI subagent workflows, or performing read-only session recovery for Codex and…

not rated 15 +8 yesterday A SkillSpector: pass 67 tokens original Apache-2.0

jest-testing-expert

628

duck4nh/antigravity-kit

Skill Claude CodeCodex

Expert in Jest testing framework, advanced mocking strategies, snapshot testing, async patterns, TypeScript integration, and performance optimization.

not rated 15 7mo ago A 28 tokens fork

playwright-electron

629

usrivastava92/gtv-desktop-remote

MCP server Claude CodeCodexCursor +2

Playwright Tools for MCP with Electron Support. Runs locally from the @robertn702/playwright-mcp-electron npm package.

not rated 15 +1 today A tokens not measured

oracle

631

qwerfunch/cladding

Skill Claude CodeCodex

Author an IMPL-BLIND spec-conformance oracle for an acceptance criterion the policy worklist (clad oracle --required) demands — an empty worklist means don't author unless the user explicitly asks. YOU spawn a blind sub-agent from a spec-only brief, then record it. Activate only when the connected project contains…

not rated 14 yesterday A SkillSpector: pass 82 tokens original MIT

review-work

633

daeryundf2-prog/LAZYANTIGRAVITY

Skill Claude CodeCodex

Post-implementation review orchestrator. Launches 5 parallel background sub-agents: Oracle (goal/constraint verification), Oracle (code quality), Oracle (security), unspecified-high (hands-on QA execution), unspecified-high (context mining from GitHub/git/Slack/Notion). All must pass for review to pass. MUST USE after…

not rated 14 changed 3d ago A 116 tokens

ttacart-pom-creator

634

PramodDutta/AdvancePlaywrightFramework1x

Skill Claude CodeCodex

Generate a Playwright Page Object (POM) class for the TTACart demo app (app.thetestingacademy.com/playwright/ttacart) by driving the live page and reading its real selectors. Use this whenever the user gives a TTACart page URL plus a flow to reach it and asks to "create a page object", "build the POM", "scaffold a…

not rated 14 2mo ago A 197 tokens

testing-helper

635

xun404/dify-agent-skill-plugin

Skill Claude CodeCodex

Generates unit tests, integration tests, and test strategies. Use for test creation, mocking, and coverage improvement.

not rated 14 26d ago A 26 tokens

simmer

637

2389-research/simmer

Plugin Claude Code

Bundles 6 skills · 456 tokens together

Iterative artifact refinement - hone any artifact or workspace over multiple rounds using criteria-driven judge feedback, runnable evaluators, and focused directional improvements.

not rated 14 2mo ago A tokens not measured original MIT

appium-python-expert

638

jmr85/e2e-agent-skills

Skill Claude CodeCodex

Specialist skill for mobile E2E testing with Appium 2.x + Python (pytest) for Android and iOS. Use this skill whenever the user asks about: setting up Appium with Python, writing mobile test cases, configuring Android/iOS drivers (UIAutomator2, XCUITest), implementing Page Object Model for mobile, running tests with…

not rated 14 +1 2mo ago A 158 tokens original MIT

screenhand

639

manushi4/Screenhand

Plugin Claude Code

Bundles 13 skills, 5 agents, 1 hook, 1 MCP server · 1,248 tokens together

Desktop and browser automation for Claude Code. Control any macOS/Windows app, automate social media, run QA tests, edit video, design in Figma/Canva, scrape the web, and orchestrate multi-agent workflows — 111 MCP tools.

not rated 14 +1 5mo ago A tokens not measured AGPL-3.0

anti-cheat

640

raydeStar/sir-thaddeus

Cursor rule Cursor

Test Harness Anti-Cheating Ruleset (No Hardcoding / No Answer Rigging) 0) Definition: What counts as “hardcoding” or “rigging”.

not rated 14 +1 8d ago A 914 tokens original Apache-2.0

minecraft-dev-loop

641

use-ai-for-mc/mcdev-mcp

Skill Claude CodeCodex

Test Minecraft mod changes in the running game - rebuild the mod, deploy the jar, restart the client via DebugBridge, and rejoin a server. Use when asked to "test this in game", restart Minecraft after a build, or run the build-deploy-relaunch-rejoin loop. Requires the mcdev-mcp MCP server and the DebugBridge mod with…

not rated 14 +1 8d ago A 80 tokens

apifox-test-scenario

642

apifox/apifox-cli-skills

Skill Claude CodeCodex

A guide for building and managing Apifox test scenarios, which are ordered workflows that check how several API calls and other actions work together.

not rated 14 +1 2mo ago A 111 tokens

hermes-bench

643

vcruz305/hermes-agentic-bench

Skill Claude CodeCodex

Run hermes-agentic-bench's agentic tool-use test battery against a model and report results.

not rated 14 +1 25d ago A 26 tokens original MIT

pos-verify

644

vincentmumme/personalos-boilerplate

Skill Claude CodeCodex

Use this immediately after files are created, edited, moved, deleted, or materially rewritten inside PersonalOS. Verifies that new truth was routed to the correct owner, written in the correct file shape, and still follows POS conventions. Do NOT use for whole-vault deep audits; use system-health-check.

not rated 14 +2 14d ago A SkillSpector: pass 65 tokens original MIT

ifBars/blender-agent-studio

Skill Codex

Benchmark Blender modeling agents, skills, prompts, scripts, or MCP tools with paired isolated runs. Use for baseline-versus-plugin comparisons, regression suites, skill forward-testing, MCP usefulness evaluation, score calibration, or claims that a Blender workflow improves mesh, visual, structural, animation, cost…

not rated 14 +5 changed today A 68 tokens original MIT

skill-creator

646

tae0y/python-project-template

Skill Claude Code

Create new skills, modify and improve existing skills, and measure skill performance. Use when users want to create a skill from scratch, edit, or optimize an existing skill, run evals to test a skill, benchmark skill performance with variance analysis, or optimize a skill's description for better triggering accuracy.

not rated 14 1mo ago A 64 tokens

android-testing-ui

647

krutikJain/android-agent-skills

Skill Codex

Validate Android UI behavior with Compose UI tests, Espresso-style checks, screenshot assertions, and accessibility verification.

not rated 14 5mo ago A 24 tokens original MIT

paranoid-qa

648

akovalion/paranoid-qa

Plugin Claude Code

Bundles 1 plugin

Turn Claude Code into a meticulous QA engineer: every Pass/Fail verdict must cite an observed artifact. Testing framework with 1000 checks, Playwright test review, test-case generation, Jira bug reports.

not rated 13 7d ago A tokens not measured original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: