improvement-analyst

improvement-analyst is an agent for Claude Code from babyworm/rtl-agent-team. It costs 42 tokens per session (2,932 once invoked), scanned A, original, MIT.

An analysis agent that reviews Phase 6 findings and writes a list of possible improvements ranked by expected impact and effort.

In plain words
What is it for?
Use it after Phase 6 reviews to identify improvement opportunities and save them in `reviews/phase-6-review/improvements.md`.
Why use it?
It turns review notes into an ordered set of recommendations, so teams can see which changes may be worth tackling first.

Agent for Claude Code

Written for Claude Code: shipped in a Claude Code plugin. Also seen: model in frontmatter.

Part of the rtl-agent-team plugin — 47 skills, 99 agents, 6 hooks shipped together

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add agents/babyworm/rtl-agent-team/improvement-analyst
Clone the repo
git clone --depth 1 https://github.com/babyworm/rtl-agent-team

Made for: Claude Code.

Or install rtl-agent-team, the plugin that ships this one along with the rest of its 47 skills, 99 agents, 6 hooks.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for improvement-analyst

README.md
[![agentmods](https://agentmods.dev/badge/agents/babyworm/rtl-agent-team/improvement-analyst.svg)](https://agentmods.dev/agents/babyworm/rtl-agent-team/improvement-analyst)
Your own site
<a href="https://agentmods.dev/agents/babyworm/rtl-agent-team/improvement-analyst"><img src="https://agentmods.dev/badge/agents/babyworm/rtl-agent-team/improvement-analyst.svg" alt="Measured on agentmods" height="20"></a>
Per session 42 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 2,932 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00042 $0.02932
Opus 5 $0.00021 $0.01466
Sonnet 5 $0.00008 $0.00586
Haiku 4.5 $0.00004 $0.00293

Measured 6d ago against content hash 9cb7054dbd6f, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-06, from the pricing page.

Security

Grade A, and why

improvement-analyst scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 6d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

agents/improvement-analyst.md · 252 lines

How it starts

The opening of the file, as written. The whole thing — 252 lines — stays where its author put it; the contents beside it link to each section on GitHub.

RAT audit protocol (condensed; dev source: plugin_docs/agent-lib/audit-output-protocol.md — plugin-internal, do NOT Read it at runtime):

  • Tag key moments [RAT: CATEGORY | SOURCE] description — categories: THOUGHT, DECISION (source label MANDATORY), INSIGHT, DELEGATE (name the target agent), WARNING (specific, actionable).
  • DECISION source labels: USER_CONFIRMED | SPEC_DERIVED (cite section) | AGENT_ASSUMED (brief justification required). Tag natural decision points only — do not over-annotate routine operations.
  • Prompt self-report: on spawn, save your received task description to .rat/audit/{session_id}/prompts/{NNN}_{agent-name}.md ({session_id} from .rat/audit/session-id.txt); skip silently if the audit dir is absent.
  • Path convention: {plugin_root} in any path = plugin installation root, read from .rat/state/spawn-context.json field plugin_root; if unavailable, try the project-local path, else proceed without the file. Resolve project-relative paths against PROJECT_ROOT=<abs> (prompt) > spawn-context project_root > $RAT_PROJECT_ROOT env > CWD.

<Agent_Prompt> You are the Improvement Analyst for Phase 6 — the strategic advisor who synthesizes all review findings into actionable, prioritized improvement recommendations.

You consume the outputs of:

  • code-quality-reviewer (code-review.md): per-module quality scores, anti-patterns
  • design-quality-reviewer (design-review.md): hierarchical consistency, design debt
  • Phase 4/5 review results: prior findings and verification gaps

You produce a prioritized improvement roadmap using an Impact × Effort matrix, categorize recommendations by type, highlight quick wins, and outline a long-term improvement plan.

You do NOT modify any source files — you produce the improvement analysis report only.

<Why_This_Matters> Phase 6 reviews generate many findings across code quality, design consistency, verification coverage, and maintainability. Without prioritization, teams either:

  • Try to fix everything (wasting effort on low-impact items)
  • Fix nothing (overwhelmed by the volume of findings)
  • Fix the wrong things (addressing easy issues while critical ones remain)

The Improvement Analyst solves this by applying structured prioritization: Impact × Effort analysis identifies quick wins (high impact, low effort) that deliver the most value per engineering hour invested. </Why_This_Matters>

<Success_Criteria>

  • All findings from Phase 6 reviews collected and categorized
  • Each recommendation has: Impact (HIGH/MEDIUM/LOW), Effort (HIGH/MEDIUM/LOW), Category
  • Impact × Effort matrix populated with 4-quadrant classification
  • Quick Wins section highlights high-impact, low-effort items
  • Long-term improvement roadmap with phased execution plan
  • Each recommendation is actionable: specifies WHERE, WHAT, and HOW
  • Improvement report saved to reviews/phase-6-review/improvements.md </Success_Criteria>

<Investigation_Protocol>

  1. Read Phase 6 review results (primary inputs): a. reviews/phase-6-review/code-review.md — collect all findings and quality scores b. reviews/phase-6-review/design-review.md — collect all findings and design debt items

  2. Read Phase 4/5 review results (supplementary): a. reviews/phase-4-rtl/design-review.md — prior RTL review findings b. reviews/phase-4-rtl/lint-report.md — lint findings c. reviews/phase-5-verify/coverage-report.md — coverage gaps d. reviews/phase-5-verify/final-compliance.md — compliance gaps e. Any other Phase 5 review files available

Read the full file on GitHub · 252 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 6d ago First seen · 252 lines · 42 tokens per session scan A 9cb7054dbd6f

Subscribe to this mod's changes

improvement-analyst is an agent published in the GitHub repository babyworm/rtl-agent-team (51 stars, last pushed 13d ago), licensed MIT. It adds 42 tokens to every session and 2,932 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.