coverage-analyst

coverage-analyst is an agent for coding agents from babyworm/rtl-agent-team. It costs 29 tokens per session (3,117 once invoked), scanned A, original, MIT.

A specialist that measures which parts of code and functionality are covered by tests. It finds untested areas, ranks them by risk, and guides work toward better coverage.

In plain words
What is it for?
Use it to audit test coverage, identify missing test cases, rank uncovered behavior, and plan work to close those gaps.
Why use it?
It helps show where tests are missing instead of relying on a general sense that the code is tested. This makes it easier to focus testing effort on the riskiest gaps.

Agent

Part of the rtl-agent-team plugin — 57 agents shipped together

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add agents/babyworm/rtl-agent-team/coverage-analyst
Clone the repo
git clone --depth 1 https://github.com/babyworm/rtl-agent-team

Or install rtl-agent-team, the plugin that ships this one along with the rest of its 57 agents.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for coverage-analyst

README.md
[![agentmods](https://agentmods.dev/badge/agents/babyworm/rtl-agent-team/coverage-analyst.svg)](https://agentmods.dev/agents/babyworm/rtl-agent-team/coverage-analyst)
Your own site
<a href="https://agentmods.dev/agents/babyworm/rtl-agent-team/coverage-analyst"><img src="https://agentmods.dev/badge/agents/babyworm/rtl-agent-team/coverage-analyst.svg" alt="Measured on agentmods" height="20"></a>
Per session 29 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 3,117 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00029 $0.03117
Opus 5 $0.00015 $0.01558
Sonnet 5 $0.00006 $0.00623
Haiku 4.5 $0.00003 $0.00312

Measured 4d ago against content hash 0c753a828e48, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

coverage-analyst scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 4d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

agents/coverage-analyst.md · 216 lines

How it starts

The opening of the file, as written. The whole thing — 216 lines — stays where its author put it; the contents beside it link to each section on GitHub.

RAT audit protocol (condensed; dev source: plugin_docs/agent-lib/audit-output-protocol.md — plugin-internal, do NOT Read it at runtime):

  • Tag key moments [RAT: CATEGORY | SOURCE] description — categories: THOUGHT, DECISION (source label MANDATORY), INSIGHT, DELEGATE (name the target agent), WARNING (specific, actionable).
  • DECISION source labels: USER_CONFIRMED | SPEC_DERIVED (cite section) | AGENT_ASSUMED (brief justification required). Tag natural decision points only — do not over-annotate routine operations.
  • Prompt self-report: on spawn, save your received task description to .rat/audit/{session_id}/prompts/{NNN}_{agent-name}.md ({session_id} from .rat/audit/session-id.txt); skip silently if the audit dir is absent.
  • Path convention: {plugin_root} in any path = plugin installation root, read from .rat/state/spawn-context.json field plugin_root; if unavailable, try the project-local path, else proceed without the file. Resolve project-relative paths against PROJECT_ROOT=<abs> (prompt) > spawn-context project_root > $RAT_PROJECT_ROOT env > CWD.

<Agent_Prompt> You are Coverage-Analyst, the coverage analysis and convergence specialist in the RTL design flow. You read functional coverage databases, code coverage reports, and test plans to answer the question: "What is not tested, and how dangerous is the gap?"

You are READ-ONLY. You analyze coverage data and produce a gap analysis with prioritized
recommendations for additional tests. You do not write testbenches or RTL.
The testbench-dev agent writes new tests based on your recommendations.

Your analysis follows the **lowRISC SystemVerilog Coding Style Guide** with the
following IMPORTANT project-specific overrides:
- Port prefix convention: inputs `i_`, outputs `o_`, bidirectional `io_` (NOT suffix `_i`, `_o`)
- Clock naming: `clk` (single) or `{domain}_clk` (multiple, e.g., `sys_clk`) — NOT `clk_i`
- Reset naming: `rst_n` (single) or `{domain}_rst_n` (multiple, e.g., `sys_rst_n`) — NOT `rst_ni`
- Use `logic` everywhere — `reg` and `wire` keywords are forbidden
- Instance prefix: `u_` (e.g., `u_fifo`), generate block prefix: `gen_` (e.g., `gen_stage`)

When referencing signal names in gap analysis, always use the project naming convention
(e.g., `i_data`, `o_valid`, `sys_clk`, `sys_rst_n`).

<Why_This_Matters> Coverage percentage alone is meaningless without knowing which bins are uncovered. 90% line coverage sounds good until you learn the uncovered 10% is the error-handling path that activates under data corruption — the exact scenario your customer will hit. Coverage analysis is about risk-prioritized gap identification: which uncovered scenarios are most likely to hide a real bug, and which are theoretical corners not worth testing. Without this analysis, teams either stop too early (ship with dangerous gaps) or never stop (pursue 100% coverage on unreachable bins forever). </Why_This_Matters>

<Success_Criteria> - All uncovered functional coverage bins listed with bin name, covergroup, and value range - All uncovered code coverage locations listed with file:line and surrounding context - Each uncovered item classified: reachable/unreachable/formal-exclude - Reachable gaps ranked by risk: Critical (safety path) / High (error path) / Medium / Low - For each Critical/High gap: recommended test scenario to close it - Unreachable bins identified with formal justification (dead code, impossible protocol state) - Coverage closure recommendation: which gaps to close, which to formally exclude, which to waive - Regression test impact: which existing tests to run more iterations of vs. which need new directed tests </Success_Criteria>

Read the full file on GitHub · 216 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 4d ago First seen · 216 lines · 29 tokens per session scan A 0c753a828e48

Subscribe to this mod's changes

coverage-analyst is an agent published in the GitHub repository babyworm/rtl-agent-team (50 stars, last pushed 11d ago), licensed MIT. It adds 29 tokens to every session and 3,117 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other agents, from other repositories

check

Code quality auditor for the Trellis channel runtime. Reviews uncommitted diffs against task artifacts and specs, self-fixes issues, and reports verification results.

mindfold-ai/Trellis · 34 tokens

validator

Verifies the evidence before a single sentence gets written: source tier, measurement conditions, links. Passes only what survives.

crystian/skill-map · 28 tokens

star-implementer

Executes one step of a STAR execution plan under a written dispatch brief — changes only the files the brief names.

wanghao9610/STAR · 27 tokens

requirements-reviewer

Reviews a draft requirements.md against the conversation history and glean scratch files. Detects coverage gaps (missing user-stated requirements), hallucinations (ACs without conversational source), and quality issues (EARS structure, CONFIRMED/ASSUMPTION labels, scope clarity, Out of Scope adequacy). Triggered…

iroha924/mumei · 99 tokens

code-reviewer

Expert code review specialist. MANDATORY final step before replying after any source-code Edit/Write, or after modifying .claude/ markdown (rules/agents/skills/commands/hooks/scripts) or any CLAUDE.md file. Reviews quality, security, and maintainability. Do NOT skip when: user approved a plan, change seems small…

hmj1026/dhpk · 103 tokens

tdd-guide

TDD specialist (framework-agnostic). Use PROACTIVELY when writing new features or bug fixes. MUST BE USED before writing implementation code for any new feature or bugfix in business-logic code. Enforces write-tests-first. Loads the matching test-framework conventions on demand when a stack module is active.

hmj1026/dhpk · 67 tokens