Skill Claude CodeCodex
Use this skill whenever asked to join an autoevolve run, evolve code toward a measured target, work an evolution population, mutate a candidate under EVOLVE-BLOCK rules, or report measured autoevolve progress and artifacts.
11,690 tagged Testing, measured the same way as everything else here.
Browse within: LLM 179agentic-ai 140agents 130cli 103ai-coding 92agent 82skills 71javascript 57openai 52agent-browser 49agentic-workflow 41ai-testing 41agent-orchestration 40claude-code-plugin 37
Skill Claude CodeCodex
Use this skill whenever asked to join an autoevolve run, evolve code toward a measured target, work an evolution population, mutate a candidate under EVOLVE-BLOCK rules, or report measured autoevolve progress and artifacts.
Skill Claude CodeCodex
Apply Traceknot's ISTQB-aligned, evidence-bound QA process to repository changes across OMP, Codex, GajaeCode, Claude Code, and OpenCode, including session-scoped QA Board publication. Use for implementation verification, bug fixes, release checks, repository audits, defect confirmation, and residual-risk decisions…
rabee-elkholy/android-harness-kit
Skill Claude CodeCodex needs its repo
Use when developing business logic, UseCases, Repositories, ViewModels, or reproducing and fixing bugs using strict Red-Green-Refactor cycles. Requires writing and proving a failing test before writing implementation code.
Skill Claude CodeCodex
Deeply analyze an API endpoint and generate complete, structured test coverage. Use this skill whenever the user provides an API endpoint, a Swagger/OpenAPI spec, or describes an API to test. Triggers on: "analyze this API", "generate test cases for this endpoint", "what should I test on this API", "test coverage for…
Skill Claude CodeCodex
Simulate real keyboard/mouse input (Windows SendInput) and capture screenshots or short GIFs of any window — built for playtesting and validating desktop apps/games (e.g. a raylib/GLFW or similar native game window) from the CLI. Use this whenever a task requires actually driving a running Windows app with input and…
Skill Claude CodeCodex
Verify Affitor tracking pipeline end-to-end with CLI test commands. Triggers on "test tracking", "verify integration", "test affitor".
julianoczkowski/create-trimble-app
Skill Cursor
Scaffold form components with proper Modus input integration, event handling, validation, and checkbox bug handling.
Skill Claude CodeCodex
Iterative agent testing with session recovery. Execute, analyze, fix, resume from checkpoints. Use when testing an agent, debugging test failures, or verifying fixes without re-running from scratch.
Skill Claude CodeCodex
Use when reviewing PRs or before opening a PR – API design, errors, memory/threading, backward compatibility, dependencies, security, XCTest quality.
Skill Claude CodeCodex
Unleash a mob of AI virtual users (impatient shoppers, confused seniors, keyboard-only users, chaos monkeys) on a website to stress-test its UX and hunt bugs. Use when the user asks to "crowd test", "mob test", "send virtual users", or wants persona-based UX/QA feedback on a URL before launch.
Skill Claude Code
Verify the Next.js web surface from a deploy-like scratch copy when the repo's root Python app/ directory masks src/app locally.
Skill Codex
Use when designing, writing, or running Nido-backed VM tests, disposable production-like QA environments, template-accelerated test labs, or isolated multi-agent AI sandboxes. Provides workflows for Nido CLI spawn/provision/upload/test/delete automation, port forwarding, templates, cleanup hygiene, and VM isolation…
filearchitect/blankfiles-website
Skill Claude CodeCodex
Use blankfiles.com as a binary test-file gateway: discover formats, filter by type/category, and return direct download URLs from the public API.
Skill Claude CodeCodex
CI for ink interactive-fiction stories. Use it whenever a .ink file changes hands or changes state: "does my story compile?", "can any path crash it?", "are all my endings reachable?", "is there content nobody can ever see?", pre-commit checks, reviewing a story PR, or validating a generated/edited ink file before…
Skill Codex
Operate a hierarchical software-development team of up to ten agents to understand a codebase, turn product requirements and bug reports into safe production-ready changes, and verify the result end to end. Use for multi-file feature implementation, bug fixing, refactoring, migrations, integrations, performance…
Skill Claude CodeCodex
Use when confirmed planning artifacts and final execution plan are available and page verification, field comparison, evidence capture, or priority execution is requested.
whaojie797-design/Novera-AI-skills
Skill Claude CodeCodex
A test-generation tool for Python that creates runnable pytest tests. Pytest is a Python tool for writing and running automated tests.
Skill Claude CodeCodex
Validate a synthetic Agent Skill when testing SkillWitness locally or in CI.
Skill Claude Code
Statistically test whether a prompt or SKILL.md change is actually better than the old version. Use when the user asks to compare two prompts, A/B test a prompt change, check if a recent edit really improved things, find rule conflicts in a long prompt, or identify which sections of a prompt are pulling weight.…
Skill Cursor
Run end-to-end web quality audits with Playwright and Lighthouse using existing project scripts first. Use when the user says /tester, e2e test, smoke test, Playwright audit, Lighthouse audit, performance audit, accessibility audit, dark mode audit, mobile audit, or asks for 100/100 Lighthouse improvement guidance.
halflength-ampleness75/claude-code-recipes
Skill Claude CodeCodex
Follow the testing pyramid — more unit tests, fewer integration tests, even fewer e2e tests.
mcp-com-ai/mcp-server-evaluations-skills
Skill Claude Code
Test MCP servers for quality and reliability. Verify tool functionality, test error handling, generate tests, and assess response quality with no dependencies other than curl. Use this when validating MCP server implementations, testing OpenAPI-to-MCP conversions, or assessing API tool quality.
Skill Claude CodeCodex
Evolve Skill: measurement-first skill optimizer. Evaluates SKILL.md files against an anchored 9-dimension rubric, validates that the rubric itself is stable (test-retest), optimizes with a hill-climbing loop that only accepts improvements larger than measurement noise, protects against overfitting with train/holdout…
Skill Claude Code
Comprehensive testing and validation of Claude skills. Use this skill whenever the user wants to test, validate, audit, or quality-check a skill — whether they say "test my skill", "check this skill works", "validate my skill", "run skill-check", or anything similar. Also trigger when the user asks things like "does…
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: