skill-optimizer

skill-optimizer is a skill for Claude Code from s977043/river-review. It costs 85 tokens per session (965 once invoked), scanned A, original, MIT.

A skill for reviewing and improving an existing Claude Code skill through defined success criteria, small changes, and evaluation checks.

In plain words
What is it for?
Auditing skill instructions, tightening trigger rules and workflows, improving examples, and adding or refining evaluations.
Why use it?
It helps identify whether a skill fails because it is triggered incorrectly, carried out poorly, or validated inadequately before changing it.

Skill for Claude Code

Written for Claude Code: allowed-tools in frontmatter. Also seen: mentions Claude Code.

Part of the river-review plugin — 138 skills, 15 commands, 5 agents, 2 hooks shipped together

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/s977043/river-review/skill-optimizer
Any agent
npx skills add s977043/river-review --skill skill-optimizer
Clone the repo
git clone --depth 1 https://github.com/s977043/river-review

Made for: Claude Code.

Or install river-review, the plugin that ships this one along with the rest of its 138 skills, 15 commands, 5 agents, 2 hooks.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for skill-optimizer

README.md
[![agentmods](https://agentmods.dev/badge/skills/s977043/river-review/skill-optimizer.svg)](https://agentmods.dev/skills/s977043/river-review/skill-optimizer)
Your own site
<a href="https://agentmods.dev/skills/s977043/river-review/skill-optimizer"><img src="https://agentmods.dev/badge/skills/s977043/river-review/skill-optimizer.svg" alt="Measured on agentmods" height="20"></a>
Per session 85 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 965 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00085 $0.00965
Opus 5 $0.00043 $0.00483
Sonnet 5 $0.00017 $0.00193
Haiku 4.5 $0.00009 $0.00097

Measured 6d ago against content hash 99a4535dc6c4, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-06, from the pricing page.

Security

Grade A, and why

skill-optimizer scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 6d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

.claude/skills/skill-optimizer/SKILL.md · 198 lines

How it starts

The opening of the file, as written. The whole thing — 198 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Pattern declaration

Primary pattern: Reviewer Secondary patterns: Pipeline Why: optimization starts with diagnosis against criteria, then applies small changes in a controlled sequence with eval checkpoints.

Purpose

Improve an existing skill without breaking what already works.

Your job is to:

  1. inspect the current skill package
  2. define or refine success criteria
  3. identify failure modes
  4. propose one small change at a time
  5. attach each change to an eval hypothesis
  6. reject changes that cannot be evaluated

Core optimization rule

Never apply a large rewrite first.

Optimize in small units:

  • description
  • gate logic
  • workflow order
  • examples
  • prohibited behaviors
  • output contract
  • review checklist
  • supporting file structure

Change only one major unit per proposal.

Phase 0: Baseline audit

Inspect:

  • current SKILL.md
  • current supporting files
  • invocation settings
  • current examples
  • current failure reports or user complaints
  • existing eval cases if any

Then summarize:

  • what the skill is supposed to do
  • where it fails
  • whether the issue is discovery, execution, or validation

Phase 1: Success criteria

Define 3 to 6 evaluation criteria.

Each criterion must be:

  • specific
  • observable
  • pass/fail or narrowly scored

Separate:

  • trigger quality
  • task fidelity
  • completeness
  • safety / side effects
  • output format compliance

Phase 2: Failure mapping

Classify failures into:

  • over-triggering
  • under-triggering
  • missing context collection
  • vague output
  • hallucinated assumptions
  • skipped verification
  • unnecessary tool use
  • high token or step cost

Phase 2.5: Pattern mismatch diagnosis

Before proposing wording or structure fixes, you must check whether the failure is caused by the wrong pattern or a missing secondary pattern. Do not skip this phase.

Check:

  • acts too early on ambiguous input → missing Inversion
  • output structure is inconsistent across runs → missing Generator
  • returns unvalidated or unchecked output → missing Reviewer
  • skips required steps or loses sequence control → missing Pipeline
  • lacks domain-specific accuracy → missing Tool Wrapper

Read the full file on GitHub · 198 lines

Files

What ships with it

3 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 6d ago First seen · 198 lines · 85 tokens per session scan A 99a4535dc6c4

Subscribe to this mod's changes

skill-optimizer is a skill published in the GitHub repository s977043/river-review (3 stars, last pushed yesterday), licensed MIT. It adds 85 tokens to every session and 965 once invoked, about $0.0004 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other skills, from other repositories

review-all

Multi-agent code review for diffs (project-agnostic). Covers standards, bugs, security, DRY, smells, perf, tests, API contracts, a11y/i18n. Verifies each finding to eliminate false positives. Use for /review-all, pre-PR/pre-commit review, or auditing uncommitted/staged changes.

ncoevoet/claude-review-all · 74 tokens

logic-health

Sweep a directory, module, or full codebase for logic correctness and produce a scored health dashboard with systemic patterns. Trigger when the user requests a health view — "audit the whole codebase", "health check", "health overview", "logic health overview", "audit src/", "audit auth and payments modules", "where…

hyhmrright/logic-lens · 180 tokens

logic-diff

Compare two code versions for semantic equivalence via semi-formal tracing of both versions side-by-side. Trigger when the user shares a refactor, rewrite, migration, or A/B implementation and wants to confirm behavior is unchanged — "did I break anything", "is this equivalent", "are these equivalent", "semantically…

hyhmrright/logic-lens · 192 tokens

omnicheck-gitlab

Use when checking if MR review findings have been applied — verifies both OmniForge-generated and human reviewer comments against the current diff, posts nudge replies on unaddressed threads.

nexiouscaliver/OmniForge · 41 tokens

omnicheck-github

Use when checking if PR review findings have been applied — verifies both OmniForge-generated and human reviewer comments against the current diff, posts nudge replies on unaddressed threads.

nexiouscaliver/OmniForge · 40 tokens

omnicreate-gitlab

Use when creating a GitLab merge request (OmniForge). Auto-populates title and description from commits, supports draft MRs, labels, assignees, reviewers, and issue linking.

nexiouscaliver/OmniForge · 45 tokens