18,394 mods in this category, of every kind an
agent can take. Each one carries what it costs per session, what the
scan found, and whether it is the original.
Use when creating or reviewing a shipd.ai Olympus quest submission — picking a candidate GitHub repo, designing a challenge task hard enough for the ≤50% pass-rate bar, writing the test patch, solution patch, test.sh, or Dockerfile, or when a platform check fails (naming collisions, Dockerfile warnings, description…
Specification-Driven Development — a structured workflow for Claude Code that enforces specify → clarify → plan → tasks → implement → validate before writing code.
★not rated 6 6mo agoA
tokens not measured
originalMIT
Adds a solid-RED opsx:tdd gate to OpenSpec (between opsx:propose and opsx:apply). Converts a change's tests.md into REAL failing tests (real render, real data-testids, real assertions) that fail only because the implementation is missing. No expect.fail placeholders.
★not rated 6 2mo agoA
tokens not measured
originalMIT
Run vLLM Buildkite CI-equivalent tests locally on NVIDIA GPUs using the current shell environment. Use when the user asks to run CI tests locally, reproduce CI failures, run a specific test file or test area, or match Buildkite test behavior.
Use this skill to aggregate and visualize test RESULTS into dashboards/reports — a self-contained HTML dashboard from JUnit XML (embedded zero-dep Python generator), Allure (rich reports with history/trends), k6 → Grafana for performance, and CI-native summaries (GitHub Actions job summary, test-reporter actions…
A skill for creating test cases for cloud-product features and interfaces. It can cover normal and error situations, with optional interface-level cases and automation code.
Review test code for resource cleanup, mock hygiene, and best practices in Bun/Node.js projects. Use after writing tests or when debugging flaky tests.
Loop Engineering portable: bucle autónomo que revisa un proyecto por flujos de ejecución (GitNexus), escribe tests (test-first), arregla lo seguro en worktrees aislados, lo verifica con un evaluador adversarial independiente y abre PRs — nunca auto-mergea. Incluye la skill /forja y 4 subagentes…
★not rated 6 2mo agoA
tokens not measured
originalMIT
Complete feature development workflow with code mapping, TDD, task-based implementation, multi-agent review, UI verification, PR management, and state persistence.
★not rated 6 7mo agoA
tokens not measured
originalMIT
Fix Codecov patch coverage gaps reported on a pull request. Use when Codecov bot flags missing or partial lines in a PR comment, when patch coverage is below the project threshold (≥87.55%), or when coverage regresses after new code is merged. Covers reading the Codecov report, identifying uncovered lines per file…
Multi-language MCP server for API testing with TypeScript/Playwright, JavaScript/Jest, Python/pytest support. Runs locally from the @kirti676/api-tester-mcp npm package.
★not rated 6 5mo agoA
tokens not measured
originalMIT
Agent-native MCP server for driving, inspecting, and asserting on real Electron desktop apps. Runs locally from the @electron-stagewright/core npm package.
★not rated 6▲
+1 1mo agoA
tokens not measured
originalMIT
Mine for hidden bugs that pattern-based auditors miss — logic errors, broken assumptions, state machine gaps, and semantic fragility.
★not rated 6▲
+1 13d agoA
tokens not measured
originalApache-2.0
At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: