18,481 mods in this category, of every kind an
agent can take. Each one carries what it costs per session, what the
scan found, and whether it is the original.
Screenshot every page of the running localhost dev server and report which pages changed since the last snapshot. Use after editing anything users can see (pages, components, CSS/Tailwind, layouts, templates) to visually verify before saying you're done, and when the user says "check the UI", "does it look right"…
Suit up. Build anything. A full engineering organization for Claude Code — deep idea interrogation, TDD sprints, enforced quality gates. 43 skills, 8 agents, 13 hooks, path-scoped rules.
★not rated 4 23d agoA
tokens not measured
originalMIT
Turns an objective into a delivered, VERIFIED OUTCOME: frames a hash-pinned sealed acceptance suite, designs the right loop (the loop-design wizard, formerly LoopPrint), war-games it forward, executes with dynamic domain-expert sub-agents (maker != checker), and gates completion on a separate verifier re-running the…
★not rated 4 4d agoA
tokens not measured
originalMIT
Engineering discipline for AI agent systems — four always-on principles, test-driven development, verification-before-completion, system-architecture playbooks, a scope-disciplined product manager, and docs/-native project management.
★not rated 4 yesterdayA
tokens not measured
originalMIT
Run an isolated persona-based UX test through the real desktop or browser UI. Creates a precise non-developer user persona, gives the tester only an approved product introduction, prevents source-code and design-document leakage, and produces an evidence-based Chinese evaluation report. Use when the user asks for…
Your agent said the tests pass. This checks that a test actually ran. Installing turns on two hooks: one strips the tail or grep that eats test results before a test command runs, one audits the session against git the moment it ends and warns only when something was caught. It also adds /red-handed:audit and…
★not rated 4 1mo agoA
tokens not measured
originalMIT
Claude Code skills for AI-context, testing, and MCP development. IANA-registered format (application/vnd.faf+yaml). Create .faf project DNA, score AI-readiness, sync with CLAUDE.md, build MCP servers, generate test suites.
★not rated 4 24d agoA
tokens not measured
originalMIT
Spec-Build-Test loop — the user defines a spec, then three agents iterate (Builder implements, Tester validates, Supervisor monitors for freezes) until the result matches. Works for any digital function — UI, APIs, CLI tools, conversational AI, data pipelines, and more.
Capture screenshots, short recordings, and milestone evidence during long-running agent tasks. Use when an agent is asked to perform multi-step browser, desktop, QA, troubleshooting, deployment, or admin workflows where the user wants checkpoint artifacts, progress evidence, error snapshots, or a final timeline of…
Plan focused tests with CodeMeridian by finding relevant test shields, coverage gaps, impacted behavior, and the smallest useful test set before implementation.
Domain skill - run screenshot-light PIE playtest episodes with structured entity observations, bounded semantic actions, transition polling, and in-memory traces for QA and external policy or RL runners.
Operator guide for MCPLab config authoring, Test Case Assistant workflows, execution, and result analysis. Use when users need to create or refine test cases from runs/traces, suggest deterministic checks or value capture, write or debug MCPLab eval YAML, run or queue evaluations, troubleshoot failures, or compare…
Use when a Looper-managed GitHub repo needs scheduled pre-merge QA — a PR carries the looper:qa label, the spec stage reaches looper:spec-ready, or the Looper reviewer loop requests an independent second pass. Runs the full QA cycle (Looper state probe → PR checkout → ffs code review → language-specific test suite →…
Implement and verify CoCo features end-to-end (Telegram commands, callbacks, app-server transport, queueing, watchdogs, approvals, and tests). Use when changing this repository's bot behavior and needing repo-specific file targets, workflows, and validation commands. NOT for generic Python tasks outside CoCo.
Automated testing execution using OpenTester DSL. Use when the user wants to create tests, run tests, validate test syntax, or manage test projects. Supports CLI testing with a YAML-based DSL.
Before diagnosing, explaining, or fixing a CI, build, test, or deployment failure, read .ai-context/failure.md when it exists. It contains the latest captured CI incident and relevant repository context.
A project rule requiring checks before code is committed and pushed. For this Deno project, the required checks are formatting, linting, and the full test suite; new functions also need documentation and tests.
★not rated 4 2mo agoA2,210 tokens
originalMIT
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: