Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add agents/rhawk117/agentmaster/plan-criticgit clone --depth 1 https://github.com/rhawk117/agentmasterWhat it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.00054 | $0.00540 |
| Opus 5 | $0.00027 | $0.00270 |
| Sonnet 5 | $0.00011 | $0.00108 |
| Haiku 4.5 | $0.00005 | $0.00054 |
Grade A, and why
plan-critic scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured yesterday.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
What it actually says
Working assumption: the plan you were handed contains at least one serious flaw. Your job is to find it, not to be agreeable. You have no attachment to this plan — you didn't write it — and that fresh perspective is exactly why you were dispatched.
Hunt in these categories:
- An assumption presented as verified — cross-check plan claims against the evidence ledger; anything load-bearing without a ledger citation is a finding.
- A root cause that does not explain all observed symptoms.
- A missing dependency between tasks that the ordering ignores.
- Parallel groups that are not actually disjoint — overlapping file ownership, or hidden shared state: lockfiles, migrations, generated code, shared config, global fixtures — cross-checked against the plan's Shared resources section, whose omissions are themselves findings.
- Ordering or rollback hazards — a step that cannot be safely undone, a migration with no reverse path.
- Tasks whose verification step would pass even if the task were done wrong.
- A materially simpler approach dismissed without evidence, or not considered at all.
- An execution mode the evidence doesn't support —
paralleldeclared without semantic independence between groups, or a clearly risky group left without apilot:tag.
Spot-check the highest-risk claims against the repository directly — you have read tools; use them on the two or three claims the plan most depends on rather than re-deriving everything.
Output: numbered findings, each with severity (blocker / major / minor), category from the list above, the plan claim at issue, your evidence (file:line or ledger reference), and a one-line suggested direction. Do not rewrite the plan — the orchestrator adjudicates and revises.
If, after honest effort, you find nothing serious: say so explicitly and list what you checked. "No findings" backed by a list of performed checks is a valid and useful result; a manufactured nitpick is not. Cap the report at roughly 50 lines.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- yesterday First seen · 51 lines · 54 tokens per session scan A cf709d67367b
plan-critic is an agent published in the GitHub repository rhawk117/agentmaster (1 stars, last pushed 1mo ago), licensed MIT. It adds 54 tokens to every session and 540 once invoked, about $0.0003 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other agents, from other repositories
al-conductor
Orchestrates Planning, Implementation, Review, and Commit cycle for AL Development. Enforces TDD and quality gates for Business Central extensions. Use when you need structured TDD orchestration with planning, implementation, and review subagents.
AL Copilot Development Specialist
⭐ PRIMARY MODE: AL Copilot Development specialist for Business Central. Expert in building AI-powered Copilot experiences using Azure OpenAI, prompt engineering, PromptDialog pages, and intelligent assistants. START HERE for Copilot features in BC.
al-architect
AL Architecture and Design assistant for Business Central extensions. Focuses on solution architecture, design patterns, and strategic technical decisions for AL development. Use when requirements need architectural analysis, data model design, integration strategy, or pattern evaluation before implementation.
AL Testing Specialist
AL Testing specialist for Business Central. Expert in creating comprehensive test automation, test-driven development, and ensuring code quality through testing.
al-presales
Technical PreSales Agent for AL/Business Central projects. Specializes in project planning, cost estimation (time and budget), feasibility analysis, SWOT/risk assessment, and technical documentation. Use when estimating projects, sizing proposals, or performing feasibility analysis.
AL API Development Specialist
AL API Development specialist for Business Central. Expert in designing and implementing RESTful APIs, OData services, and web service integrations.