Testing commands

2,999 tagged Testing, measured the same way as everything else here.

Browse within: agentic-workflow 44claude-plugin 33ai-development 32agentic-coding 31code-quality 31spec-driven-development 27agent-orchestration 26agent-framework 25ai-assistant 25agentic 21ai-workflow 21Multi-Agent 20claude-code-skills 20documentation 20

kaku-runtime-test

457

cklxx/elephant.ai

Command Claude Code

An end-to-end test command for Kaku, a tool that manages programming-agent panes, together with Feishu messages and runtime monitoring. It starts from a simulated Feishu message and checks the full agent-team workflow.

11 5mo ago A 0 tokens original MIT

run-action

458

Lykhoyda/rn-dev-agent

Command

Part of rn-dev-agent

Execute a learned Maestro flow ("action") by name with optional -e KEY=VALUE parameters. Looks the flow up via packages/rn-dev-agent-core/dist/learned-actions.js (same inventory as /rn-dev-agent:list-learned-actions), then replays it via cdprunaction — auto-repair-aware orchestration with structured RunRecords (GH.

11 3d ago A 72 tokens original MIT

e2e-test

459

J0hnG4lt/metabase-flightsql-driver

Command Claude Code

Run end-to-end tests for the Metabase Arrow Flight SQL driver. Starts all services from scratch, runs setup, and validates the dashboard works.

11 1mo ago A 29 tokens original Apache-2.0

np:eval

460

kayac/nepp-chan

Command Claude Code

An interactive evaluation command for scoring and displaying the answer quality of a knowledge agent. An evaluation checks whether an agent gives useful and correct answers.

11 4d ago A 30 tokens AGPL-3.0

negatives

461

Ecro/embedeval

Command Claude Code

Sequentially author negatives.py for every case that lacks one, with interactive judgment gates where LLM cannot reliably decide.

11 27d ago A 0 tokens original Apache-2.0

smoke-test

462

trainual/tiptap-collaboration-mcp

Command Claude Code

Smoke test the tiptap-collaboration MCP tools against a live server. Use when user says "smoke test", "test tiptap", or "test mcp tools".

11 6d ago A 36 tokens original MIT

build-unit-tests

463

armoin2018/ai-ley

Command Claude Code

Generate comprehensive unit testing infrastructure including test configuration, health checks, regression tests, synthetic transactions, coverage dashboards, and CI/CD integration.

11 7mo ago A 0 tokens original CC0-1.0

polygen

464

cognitive-fab/polygraph

Command

Part of polygraph

Run polygen — draft a contract from a feature description, author a verifiable SAM v2 strict-profile module against it, self-repair against reachable invariant violations, and synthesize a demo/regression trace corpus.

11 6d ago A 43 tokens original Apache-2.0

polyvers

465

cognitive-fab/polygraph

Command

Part of polygraph

Run polyvers — classify a state-machine version change into compatibility lanes, run the gates those lanes require against fleet snapshots (shape round-trip, vocabulary, in-flight stimuli, migration validation, seeded model check), scaffold migrations, and check parent×child version matrices. No API key.

11 6d ago A 56 tokens original Apache-2.0

verify

466

cognitive-fab/polygraph

Command

Part of polygraph

Run the Polygraph verification loop — generate N transition-function specs from a source file and replay real traces against them, reporting spec-errors vs code-findings.

11 6d ago A 31 tokens original Apache-2.0

tailtest-hunt

467

avansaber/tailtest

Command

Part of tailtest

Run an adversarial pass on $ARGUMENTS -- explicitly try to break the source code.

11 +1 2mo ago A 0 tokens original MIT

calibrate

469

let-sunny/canicode

Command Claude Code

Part of canicode

Run the calibration pipeline for a single fixture, all active fixtures, or resume a failed run.

10 2mo ago A 0 tokens original MIT

review-run

470

let-sunny/canicode

Command Claude Code

Part of canicode

Review a completed /develop pipeline run and provide a structured QA assessment.

10 2mo ago A 0 tokens original MIT

test

471

Caspian-Sun/claude-code-workflow

Command Claude Code

You are now acting as a Test Engineer. Generate comprehensive test cases for the specified components/functions.

10 6d ago A 0 tokens original MIT

run-tests

472

zircote-plugins/sdlc-quality

Command Claude Code

Part of sdlc

Run automated functional tests using the hook-driven test framework. Execute the test suite to validate all project functionality.

10 1mo ago A 24 tokens original MIT

bugfix

474

rube-de/cc-skills

Command

TDD-driven bugfix workflow: tester writes failing test (RED) → developer fixes (GREEN) → developer refactors (REFACTOR) → reviewer validates. Accepts issue number, description, or both. Auto-creates PR unless --no-pr flag is passed.

10 6d ago A 55 tokens original MIT

create-feature-plan

476

paulbreuler/limps

Command Claude Code

Generate a TDD plan with verbose planning docs and minimal agent execution files using MCP planning tools.

10 6mo ago A 0 tokens original MIT

create-python-tests

477

codeready-toolchain/tarsy

Command Cursor

Command "create-python-tests" from codeready-toolchain/tarsy, covering writing tests for python llm service, running tests, from project root, from llm-service/ directory and critical rules.

10 3d ago A 0 tokens original Apache-2.0

harness

480

xiongchenyu6/dotfiles

Command Claude Code

A command that uses browser automation to run an experience check on the current application or game.

10 4d ago A 0 tokens

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: