verify
553Skill Claude CodeCodex
Verify a code change does what it should by running the app.
11,777 tagged Testing, measured the same way as everything else here.
Browse within: LLM 180agents 133agentic-ai 129cli 103ai-coding 99agent 80skills 72javascript 57agent-browser 54openai 51ai-testing 46agentic-workflow 41agent-orchestration 40claude-code-plugin 40
Skill Claude CodeCodex
Verify a code change does what it should by running the app.
Skill Claude CodeCodex
LTP Test Analyzer - evaluate test quality, robustness, and coverage.
Skill Claude Code
Set up Playwright E2E testing framework for any project — auto-detects stack, auth, routes, and selectors.
Skill Claude CodeCodex
A Russian-language style guide and editor for Allure test annotations. Allure is a tool that turns test metadata, titles, descriptions, and steps into readable test reports.
Skill Claude CodeCodex
Audits a repo's test suite against a short rubric and scaffolds a comprehensive one if it's thin or missing. Detects language and test framework, measures coverage where possible, identifies gaps (no unit tests, no integration tests, no CI hook, slow or flaky suite, uncovered critical paths), and presents a scored…
Skill Claude CodeCodex
A testing skill that checks whether a third-party coding-agent skill does what its documentation claims. It breaks claims into individual items and tests them in a temporary workspace.
aliceisjustplaying/claude-resources-monorepo
Skill Claude CodeCodex
Control iOS Simulators via accessibility APIs. Use this skill when the user wants to automate iOS simulator interactions, tap buttons by accessibility label, type text, swipe, take screenshots, describe the UI accessibility tree, or test iOS apps programmatically.
Skill Claude Code
Cortex XSOAR content pack development lifecycle - create packs, integrations, scripts, playbooks, run demisto-sdk lint/validate/pre-commit, build zip packs, manage versions and release notes, run unit tests, deploy to XSOAR instances, manage git branches/tags, handle marketplace vs local pack workflows. Use when the…
gzhanlei/claude-mob-programming-skill
Skill Claude CodeCodex
A team-based coding workflow that assigns different roles to two or three agents. TDD, or test-driven development, means writing tests before the code that makes them pass.
Skill Claude CodeCodex
A collaboration paradigm-driven development framework.
Skill Claude CodeCodex
Use when the user asks to build a feature using test-driven development.
Skill Claude CodeCodex
Plugin-shipped skill that wraps the workspace test runner (pytest / jest / go test) so the tester role can issue a single command and parse a single exit code. Defines the CLI surface and exit-code contract.
Skill Claude CodeCodex
Hunting skill for auth bypass vulnerabilities. Built from 12 public bug bounty reports across SAML XSW / parser-differential (GitHub Enterprise CVE-2025-25291/25292), SAML signature stripping (Uber, Rocket.Chat, samlify CVE-2025-47949), SAML domain enforcement bypass via control characters (HackerOne 2024)…
Skill Claude CodeCodex
End-to-end testing specialist using Playwright. Use PROACTIVELY for generating, maintaining, and running E2E tests. Manages test journeys, quarantines flaky tests, uploads artifacts (screenshots, videos, traces), and ensures critical user flows work.
Skill Claude Code
Markdown-first, AI-native parent skill for one repo-level Loop9 repo-complete loop. Use when you want the parent layer to hold repo-level owner, repo truth relocation, full finding-queue completion, repo-level closure, and final deliverable-oriented verdict while dispatching only thin handoffs to env-bootstrap /…
Skill Claude CodeCodex needs its repo
Design and build test automation frameworks from scratch. Load when asked about automation architecture, framework design, test pyramid strategy, choosing between POM vs Screenplay pattern, data-driven or keyword-driven frameworks, setting up CI/CD pipelines for test automation, parallel execution, folder structure…
wan-aishengtang/testing-master-skill
Skill Codex
Handle software testing in a test-owner style, including early test strategy, complete testing framework design, requirements analysis, test planning, manual functional testing, defect report writing, API testing, API automation regression planning, Web UI automation planning, performance testing, system acceptance…
chenxihuang1028-a11y/verification-before-completion
Skill Claude CodeCodex
Use when about to claim work is complete, fixed, or passing, before committing or creating PRs - requires running verification commands and confirming output before making any success claims; evidence before assertions always.
Skill Claude CodeCodex
Production patterns for Hono v4+ backends with Bun, Drizzle ORM, Better Auth, Stoker, hono-openapi + Zod + Scalar, and TDD with Bun test runner. Use when building features, writing tests, reviewing code, or setting up auth/RBAC in modular monolith Hono projects. Apply whenever you see Hono route definitions…
Skill Claude CodeCodex
Use when you need to run MintMaker controller code locally against a real Kubernetes cluster (minikube or kind) with repo hosted on GitHub, create test Component/DependencyUpdateCheck resources, and iterate quickly across macOS and Linux.
Skill Claude CodeCodex
Guidance for writing runevals.py scripts that use LangSmith evaluate() with a dataset ID.
Skill Claude CodeCodex
Use Stet to measure whether an AI coding change is safe to ship. Trigger on model comparisons, AGENTS.md or CLAUDE.md effectiveness, shared instruction, policy, skill, harness, tool-policy, reasoning, or runtime rollouts, repo eval setup, dataset building, regression detection, benchmarking, promote/rollback…
alexhvastovich/playwright-ai-pom-starter
Skill Codex
Assign or review stable Playwright test case IDs, spec headings, file names, describe blocks, and test titles in this repository. Use when adding, renaming, generating, or reviewing a scenario, or when a request mentions test IDs, naming conventions, traceability, file naming, or duplicate test identity.
Skill Claude Code
Analyzes test files (.test.ts) to generate code-slice.json specification files. These capture behavioral contracts in given/when/then form, enabling drift detection between tests and slice.json design documentation.
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: