scorecard

scorecard is a skill for Claude Code, Codex from karanb192/claude-code-hooks. It costs 20 tokens per session (237 once invoked), scanned B, original, MIT.

A report showing how often Claude follows the rules in CLAUDE.md, a file that gives coding instructions for a project. It highlights rules that are often ignored and may be better enforced automatically.

In plain words
What is it for?
Use it to review recent rule compliance, identify the worst offenders, and decide whether to replace repeated reminders with hooks or checks.
Why use it?
It shows which written instructions are not working reliably in practice.

Skill for Claude CodeCodex

Installs and runs on its own, but its text points at files inside its plugin — anything it tells you to read at a ${CLAUDE_PLUGIN_ROOT} path is only there once the plugin is installed. Installing the plugin gets both.

Part of the dead-rules-audit plugin — 1 skill shipped together

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/karanb192/claude-code-hooks/scorecard
Any agent
npx skills add karanb192/claude-code-hooks --skill scorecard
Clone the repo
git clone --depth 1 https://github.com/karanb192/claude-code-hooks

Made for: Claude Code, Codex.

Or install dead-rules-audit, the plugin that ships this one along with the rest of its 1 skill.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for scorecard

README.md
[![agentmods](https://agentmods.dev/badge/skills/karanb192/claude-code-hooks/scorecard.svg)](https://agentmods.dev/skills/karanb192/claude-code-hooks/scorecard)
Your own site
<a href="https://agentmods.dev/skills/karanb192/claude-code-hooks/scorecard"><img src="https://agentmods.dev/badge/skills/karanb192/claude-code-hooks/scorecard.svg" alt="Measured on agentmods" height="20"></a>
Per session 20 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 237 The whole file, excluding the scripts and references it only reads on demand.
Security scan B 1 finding. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00020 $0.00237
Opus 5 $0.00010 $0.00118
Sonnet 5 $0.00004 $0.00047
Haiku 4.5 $0.00002 $0.00024

Measured 4d ago against content hash f06d6af3064b, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade B, and why

scorecard scanned grade B with 1 finding against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 4d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Unrestricted tool accessmediumExcessive agency

A wildcard tool grant or "run any command" leaves no least-privilege boundary at all.

re-run any commands.
plugins/dead-rules-audit/skills/scorecard/SKILL.md · 21 lines

What it actually says

CLAUDE.md compliance scorecard

!node "${CLAUDE_SKILL_DIR}/../../dead-rules-audit.js" --render 2>/dev/null || node "${CLAUDE_PLUGIN_ROOT}/dead-rules-audit.js" --render

The card above is a worst-first audit of your CLAUDE.md rules, aggregated across your recent sessions: how often each rule was relevant to an edit, how often it was violated, its heuristic compliance %, and a ⚠ promote→hook flag for rules Claude chronically ignores. It is produced by the dead-rules-audit hook, which records as you work.

Briefly call out the 1-3 worst offenders and, for any rule flagged promote→hook, suggest making it deterministic (a PreToolUse/PostToolUse hook or a lint rule) instead of relying on Claude to remember it. Keep it short; do not re-run any commands.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 4d ago First seen · 21 lines · 0 tokens per session scan B f06d6af3064b

Subscribe to this mod's changes

scorecard is a skill published in the GitHub repository karanb192/claude-code-hooks (498 stars, last pushed 11d ago), licensed MIT. It adds 20 tokens to every session and 237 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it B with 1 finding (unrestricted tool access). No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other skills, from other repositories

system-audit

Periodic evidence-based self-audit of the whole agent system — the orchestrator, its docs/SSOTs, rules, memory, subagents, tools, infra and the work it claims to have delivered — followed by immediate cheap-safe fixes and a ranked backlog. Use when the user says 'аудит системы', 'проверь себя', 'самопроверка', 'system…

awrshift/claude-memory-kit · 225 tokens

session-review

End-of-session adversarial review loop. Assemble the session's work into a role-assigned, self-contained brief, then run independent reviewers in parallel — an isolated code-reader (the idea-validator agent) that reads the ACTUAL files and web-checks technology currency, plus an external-family model if you have one …

awrshift/claude-memory-kit · 149 tokens

close-session

End-of-session ritual — audit today's patterns against accumulated memory, propose promotions, refresh MEMORY.md, and write the session handoff. Use when the user says "/memory-kit:close-session", "закрой сессию", "закрываем", "we're done for today", "wrap up".

awrshift/claude-memory-kit · 64 tokens

memory-audit

Audit MEMORY.md against the memory discipline — oversized sections, settled multi-session patterns that belong in knowledge/concepts/, stacked chronicle blocks, stale entries. Produces a move plan as a table for approval, then executes the approved moves atomically. Use when the SessionStart hook reports a tripped…

awrshift/claude-memory-kit · 119 tokens

second-opinion

Cross-check the agent's own answer with independent reviewers before bringing it to the user. Use when the user says 'second opinion', 'sanity check', 'cross-check', 'am I missing something', 'stress-test', 'devil's advocate', 'run a full review', 'this is important', 'high-stakes', 'help me choose between', 'critique…

awrshift/claude-memory-kit · 130 tokens

qa-sweep

Run a multi-lens agent QA sweep of the RUNNING product: spawn qa agents (one per lens — user-flow · edge-state · honesty · contract · ux-critique), collect their structured findings, integrator-verify the load-bearing ones, and land verified findings as backlog tickets + a run record in the project's qa/ folder. Use…

awrshift/claude-memory-kit · 199 tokens