BenchBox-dev

19 mods across 1 repository, 12 stars between them.

pr

01

BenchBox-dev/BenchBox

Command Claude Code

BenchBox PR workflow - path-aware preflight, push, open PR vs develop; does not enable auto-merge unless READY=1.

not rated 12 today A 27 tokens original MIT

BenchBox

02

BenchBox-dev/BenchBox

Settings file Claude Code

Agent settings configuring enabledMcpjsonServers, extraKnownMarketplaces, enabledPlugins.

not rated 12 today A tokens not measured copy · 83% MIT

benchbox

03

BenchBox-dev/BenchBox

Skill Claude Code

Use when the user asks to "test TPC-H", "check compliance", "review architecture", "run quality checks", "check binaries", "test dialect translation", "compare implementations", "run live platform tests", "cut a release", "finalize a release", or "plan and execute" a benchmark feature.

not rated 12 today A 68 tokens original MIT

bossmode

04

BenchBox-dev/BenchBox

Skill Claude Code

Organize and execute complex multi-step work through an executive, persistent named managers, focused workers, and independent review. Use when work divides across parallel workstreams or requires an independent review gate.

not rated 12 today A 41 tokens original MIT

code

05

BenchBox-dev/BenchBox

Skill Claude Code

Use for "implement code", "build a feature", "refactor code", "commit code", "review code", "adversarially review code", "review a code change", "review all code work in this session", "address PR review follow-ups", "run a PR review follow-up sweep", "clear PR backlog", "process PR backlog", "fix lint/type error"…

not rated 12 today A 129 tokens original MIT

docs

06

BenchBox-dev/BenchBox

Skill Claude Code

Use when the user asks to "create documentation", "build docs", "review docs", "compare documents", "compress docs", "adversarial review docs", or "commit docs".

not rated 12 today A 40 tokens original MIT

agent-execution

07

BenchBox-dev/BenchBox

Skill Claude Code

Select model tiers, map reasoning effort, and dispatch delegated work through native or external agent harnesses. Use when a workflow must choose an agent model or effort level, or launch a manager, worker, or independent reviewer; do not use for direct, undelegated tool calls.

not rated 12 today A 60 tokens original MIT

change-framework

08

BenchBox-dev/BenchBox

Skill Claude Code

Unified source-code selection and change-execution workflow: reuse ladder, vertical slicing, post-edit verification, named branches, commits, and authorized-write PRs.

not rated 12 today A 34 tokens original MIT

BenchBox-dev/BenchBox

Skill Claude Code

Unified investigation workflow: comparing artifacts, pre-edit research, context trust/authority handling, root-cause debugging, and validation-driven compression.

not rated 12 today A 31 tokens original MIT

review-protocol

10

BenchBox-dev/BenchBox

Skill Claude Code

Shared protocol for review-shaped actions, authorization scope, defect routing, solution-fit assessment, L1/L2/L3 planning-depth layers, local-only capture, and plan prior-decision reconciliation.

not rated 12 today A 42 tokens original MIT

skill-sync

11

BenchBox-dev/BenchBox

Skill Claude Code

Use when the user asks to sync, set up, inspect, validate, verify, pin, prune, promote, or configure skills managed by skill-sync.

not rated 12 today A 34 tokens original MIT

test

12

BenchBox-dev/BenchBox

Skill Claude Code

Use when the user asks to "run tests", "create tests", "fix failing test", "add test coverage", "fix slow tests", or "commit test changes".

not rated 12 today A 37 tokens original MIT

tidy-perms

13

BenchBox-dev/BenchBox

Skill Claude CodeCodex

Consolidate accumulated permission grants across Claude Code, Codex, and Gemini: move trusted commands into project settings, clean garbage entries, verify cross-agent consistency, commit project-level configs.

not rated 12 today B 42 tokens original MIT

todo-db

14

BenchBox-dev/BenchBox

MCP server Claude CodeCodexCursor +2

MCP server "todo-db" as configured in BenchBox-dev/BenchBox. Runs locally from the _project/scripts Python package. Needs 1 environment variable to run.

not rated 12 changed today A tokens not measured original MIT

BenchBox CLAUDE.md

15

BenchBox-dev/BenchBox

Instructions file Claude Code

Claude Code instructions for BenchBox-dev/BenchBox: Read and follow AGENTS.md; it is the active BenchBox authority. Load relevant generated skills from .claude/skills/.

not rated 12 today A 33 tokens original MIT

BenchBox GEMINI.md

16

BenchBox-dev/BenchBox

Instructions file Gemini CLI

Gemini CLI instructions for BenchBox-dev/BenchBox: Read and follow AGENTS.md; it is the active BenchBox authority. Load relevant generated skills from .agents/skills/. Keep the shared mirror at that path; .gemini/skills/ is no longer populated.

not rated 12 today A 52 tokens original MIT

todo-db

17

BenchBox-dev/BenchBox

Skill Claude Code

Use when working in a project tracked by todo-db — "what should I work on", "what's ready", "claim a TODO", "start work on an item", "record progress", "finish an item", "create a TODO", "add a work item", "check scope", "release my claim", "why did finish fail", "tracker stats", "prioritize TODOs", "batch…

not rated 12 today A 124 tokens original MIT

At most 3 mods per repository are shown here, and a mod shipped inside a plugin is left to that plugin's page — the rest are on their repository pages: