18,481 mods in this category, of every kind an
agent can take. Each one carries what it costs per session, what the
scan found, and whether it is the original.
Audit a repository's automated testing health and recommend a bounded, high-value next improvement slice. Use when assessing an existing test suite, introducing tests into an untested or legacy project, deciding what to test next, evaluating testing strategy or coverage, choosing between unit, integration, and…
Engineering & Claude-coding craft skills (diagnose, tdd, prototype, zoom-out, grill-with-docs, improve-codebase-architecture, to-issues, to-prd) adapted from mattpocock/skills (MIT) to the bootstrap conventions and the AFK/HITL model, plus homegrown skills (checkpoint).
★not rated 3 1mo agoA
tokens not measured
originalMIT
The senior dev, unbundled. A cast of opinionated skills that make an AI agent push back on bad ideas, verify before it claims done, attack its own code, and finish like a senior engineer.
Claude Code plugin for Python development with Pyright LSP integration, 15 automated hooks for type checking, linting (ruff), formatting (black), testing (pytest), and security scanning.
★not rated 3 1mo agoA
tokens not measured
originalMIT
Mechanically validate the sonu plugin repo before a PR — manifest sync, YAML frontmatter, shell-fence syntax, named-source and AI-attribution scans, cross-reference integrity, skill reachability. Only meaningful inside the claude-plugins repo; in any other repo, say so and stop.
Design evaluation contracts and test plans for agentic systems. Create deterministic tests, trajectory evals, quality dimensions, gold-set criteria, and CI gates before or after implementation. Use when asked for tests first, an eval plan, success criteria, non-deterministic testing, LLM-as-judge setup, or…
Cursor rule "rspec-let-do-blocks" from panozzaj/cursor-rules, covering multi-line array, multi-line method call with keyword args, multi-line hash definition, complex single line expression and using { ... } for multi-line blocks.
Entry point for the disciplined reasoning pipeline. Invoke to make a cheap model plan, implement, self-critique, and verify with hard no-premature-done gates.
Use when writing an implementation plan in a project — wraps superpowers:writing-plans with a fresh-subagent test-rigor review loop (≤3 rounds) after the plan is written. Each round asks "what realistic failure mode would survive these tests?" and either extends the plan or marks gaps out-of-scope in a Deferred Risks…
End-to-end test a Claude Code plugin by launching an interactive Claude session in a tmux pane with the plugin loaded, sending prompts, and verifying behavior.
Advanced Model Context Protocol (MCP) server for Vitest testing with intelligent resources, coverage analysis, and AI-assisted development workflows. Runs locally from the @madrus/vitest-mcp-server npm package.
★not rated 3 1y agoA
tokens not measured
originalMIT
MCP server "cypress-mcp-server" as configured in glh230/cypress-mcp-server. Runs locally from the cypress-mcp-server npm package.
★not rated 3 10mo agoA
tokens not measured
originalMIT
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: