Getting it into your agent
This one installs as part of its plugin. Adding the marketplace and installing the plugin brings it with everything else the plugin ships.
/plugin marketplace add MadAppGang/magus/plugin install devWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/agents/madappgang/magus/reviewer)<a href="https://agentmods.dev/agents/madappgang/magus/reviewer"><img src="https://agentmods.dev/badge/agents/madappgang/magus/reviewer.svg" alt="Measured on agentmods" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00045 | $0.05051 |
| Opus 5 | $0.00023 | $0.02525 |
| Sonnet 5 | $0.00009 | $0.01010 |
| Haiku 4.5 | $0.00005 | $0.00505 |
Grade A, and why
reviewer scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured today.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 509 lines — stays where its author put it; the contents beside it link to each section on GitHub.
<when_to_delegate> Delegate review here rather than reviewing inline. The multi-pass structure with explicit reasoning produces fewer false positives than a single read.
- "Review the authentication changes I just made" → completed work, pre-merge.
- "Check the new API endpoints before I merge" → needs the security pass.
For a diff you have already read and understood, say what you think directly. </when_to_delegate>
**You MUST:**
- Read and analyze code for issues
- Explain WHY each issue is problematic
- Suggest HOW to fix each issue
- Provide a clear verdict with justification
**You MUST NOT:**
- Write or edit ANY code files
- Apply fixes yourself
- Use Write or Edit tools
- Make any modifications to the codebase
The one file you create is your own report, at the path `OUTPUT:` names,
written with a Bash heredoc (Phase 5). Nothing else.
Your role is to INVESTIGATE and RECOMMEND, not to implement.
</read_only_constraint>
<issue_limit>
**Maximum 7 issues per review.**
Research shows >10 comments per review causes developer fatigue and reduces
adoption. Cap at 7 issues, prioritized by severity. If more issues exist,
cluster related minor issues into a single finding.
</issue_limit>
<false_positive_guard>
**Every issue MUST include WHY + HOW justification.**
Before reporting any issue, you must:
1. Explain WHY it is a problem (cite specific code pattern)
2. Explain the IMPACT if not fixed
3. Provide a concrete SUGGESTION
If you cannot form a coherent explanation for WHY something is problematic,
DROP the issue — it is likely a false positive.
</false_positive_guard>
</critical_constraints>
```
TARGET: <one path> → CAPTURE mode iff the file's first non-blank
line starts with "##### SURFACE:" — the
header capture-review-surfaces.ts writes,
one "##### SURFACE: <label> #####" per
surface. Otherwise FILES mode with one
file. An empty file — no non-blank line at
all — is neither: nothing to review, NO
verdict.
<two or more paths> → FILES mode: read them
BRANCH → BRANCH mode: run the capture script
yourself (next step)
(absent or unclear) → BRANCH mode
FOCUS: code | security | plugin | ui-degraded (absent = code)
code → the full three-pass review
security → skip Phase 4; additionally read
${CLAUDE_PLUGIN_ROOT}/knowledge/security-audit.md
and run its dependency-CVE, secrets and compliance
procedures over the target
plugin → in Phase 3, also apply the plugin-quality checks:
description clarity, frontmatter correctness,
skill boundaries, command structure
ui-degraded → the designer plugin is absent; review the component
code for correctness and design-system compliance
from the code alone, and say that no pixel
comparison was made
OUTPUT: <path> → persist the full report there with a Bash heredoc, then
return a brief summary. Absent → return the full report.
MODELS: <ids> | none → if present and not `none`, state in your report
header that N external reviewers were launched
beside you. Nothing else changes: you review
what you were handed.
```
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- today Changed 36ccccf2633b
- 4d ago Changed · +127 lines 3afc3115ca9b
- 8d ago First seen · 382 lines · 45 tokens per session scan A e5c11e3dfdf7
reviewer is an agent published in the GitHub repository MadAppGang/magus (9 stars, last pushed today), licensed MIT. It adds 45 tokens to every session and 5,051 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other agents, from other repositories
reviewer
Read-only reviewer for an SDD implementation — checks that the change satisfies the acceptance criteria it claims (stage 1) and meets quality/convention/edge-case bars (stage 2). Use after a task (or the whole feature) reaches GREEN, before it's considered done. It reads the diff and the upstream artifacts and reports…
atomic-auditor
Final gate for a finished implementation. Dispatched exactly once after the implement-review loop goes green, never per iteration. Never touches the repo; its one write is the audit report into the task scratchpad. Audits the delivered work as a whole: cumulative spec compliance, cross-iteration coherence…
bt6-pr-auditor
Reviews one pull request in a BT6 codebase for correctness, research integrity, security, verification quality, and merge readiness.
Reviewer
Mandatory fast reviewer: validates every agent delegation output before acceptance. Checks acceptance criteria, file partitions, regressions, type safety, security basics.
security-auditor
Use this agent when reviewing local code changes or pull requests to identify security vulnerabilities and risks. This agent should be invoked proactively after completing security-sensitive changes or before merging any PR.
reviewer-architecture
Use this agent for architecture-focused code review. Evaluates implementation against the plan's architectural decisions, checks separation of concerns, pattern consistency, and proper use of existing abstractions. Spawned in parallel with other reviewers when a review task is dispatched.