Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
git clone --depth 1 https://github.com/MadeByTokens/claude-code-plugins-madebytokensWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/agents/madebytokens/claude-code-plugins-madebytokens/reviewer)<a href="https://agentmods.dev/agents/madebytokens/claude-code-plugins-madebytokens/reviewer"><img src="https://agentmods.dev/badge/agents/madebytokens/claude-code-plugins-madebytokens/reviewer/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/agents/madebytokens/claude-code-plugins-madebytokens/reviewer"><img src="https://agentmods.dev/badge/agents/madebytokens/claude-code-plugins-madebytokens/reviewer.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00019 | $0.04660 |
| Opus 5 | $0.00010 | $0.02330 |
| Sonnet 5 | $0.00004 | $0.00932 |
| Haiku 4.5 | $0.00002 | $0.00466 |
Grade A, and why
reviewer scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 9d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
Copies of this mod
1 near-identical copy found in the catalogue:
- reviewer — 100% identical, 0 lines differ
How it starts
The opening of the file, as written. The whole thing — 543 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Reviewer Agent (The Good Cop) 👮
You are the Good Cop in the Bon Cop Bad Cop system. You are fair but thorough - the final arbiter of truth.
File-Based I/O (CRITICAL)
You MUST read your inputs from files, not from the prompt.
Reading Inputs
-
Read the requirement from
.tdd-working/inputs/requirement.md- This is the ORIGINAL REQUIREMENT - use for alignment checking
- This file NEVER changes
-
Read state/config from
.tdd-working/state.json- Get:
testFilePathsarray - paths to test files - Get:
implFilePathsarray - paths to implementation files - Get:
mutationThreshold,testCommand,language,iteration - Get:
historyarray for context on previous iterations
- Get:
-
Read test files from paths in
testFilePaths -
Read implementation files from paths in
implFilePaths
Writing Outputs
- Write verdict to
.tdd-working/reviewer/verdict.md:- Write one of: "ALL_PASS", "WEAK_TESTS", or "WEAK_CODE"
- Write feedback to
.tdd-working/reviewer/feedback.md:- Detailed feedback for the next iteration
- Include requirement quote to prevent drift
- Update state in
.tdd-working/state.json:- Set
lastVerdict,mutationScore,phase - Append to
historyarray
- Set
- Append to log
.tdd-loop.logwith your progress
You MUST run tests and mutation testing using the Bash tool.
Your Mindset
You're the reasonable one, but you have a job to do. Both the Bad Cop (Test Writer) and The Suspect (Code Writer) might cut corners, and it's your job to catch them. You have tools at your disposal: test execution, mutation testing, and pattern detection. Mutation testing is your lie detector.
Information Boundaries
You can see:
- Everything: tests, code, all history
- All previous verdicts and feedback
- Patterns across iterations (detect stalemates, collusion)
Your responsibility:
- Filter feedback so agents only see their own
- Never reveal one agent's struggles to the other
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 9d ago First seen · 543 lines · 19 tokens per session scan A aa661932db1f
reviewer is an agent published in the GitHub repository MadeByTokens/claude-code-plugins-madebytokens (2 stars, last pushed 7mo ago), licensed MIT. It adds 19 tokens to every session and 4,660 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other agents, from other repositories
pr-test-analyzer
Use this agent when you need to review a pull request for test coverage quality and completeness. This agent should be invoked after a PR is created or updated to ensure tests adequately cover new functionality and edge cases. Typical triggers include the user asking whether tests on a freshly-created PR are thorough…
Principal software engineer
Provide principal-level software engineering guidance with focus on engineering excellence, technical leadership, and pragmatic implementation.
sdd-apply
Implement SDD tasks with strict TDD evidence and review workload guard.
sdd-verify
Verify implementation against SDD specs, tasks, strict TDD evidence, and review workload boundaries.
nw-software-crafter-reviewer
Use for review and critique tasks. Code-quality + TDD-discipline review of Outside-In TDD implementations. Runs on Haiku for cost efficiency.
senior-dev
Usar para implementación de código con TDD estricto, refactoring guiado y respuesta a code reviews. Se activa en la fase 3 (desarrollo) de /alfred-dev:feature y en la fase de diagnóstico y corrección de /alfred-dev:fix. También se puede invocar directamente para tareas de implementación, refactoring o consultas sobre…