plan-critic

An independent checker for a draft implementation plan that assumes the plan may contain serious mistakes and verifies its claims against the repository.

In plain words
What is it for?
Use it to review coding plans, check whether tasks can safely run in parallel, and confirm that the proposed verification would detect incorrect work.
Why use it?
It helps expose unsupported assumptions, missing dependencies, unsafe ordering, overlapping work, weak tests, and simpler options before implementation begins.

Agent

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add agents/rhawk117/agentmaster/plan-critic
Clone the repo
git clone --depth 1 https://github.com/rhawk117/agentmaster
Per session 54 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 540 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00054 $0.00540
Opus 5 $0.00027 $0.00270
Sonnet 5 $0.00011 $0.00108
Haiku 4.5 $0.00005 $0.00054

Measured yesterday against content hash cf709d67367b, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

plan-critic scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured yesterday.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

agents/plan-critic.md · 51 lines

What it actually says

Working assumption: the plan you were handed contains at least one serious flaw. Your job is to find it, not to be agreeable. You have no attachment to this plan — you didn't write it — and that fresh perspective is exactly why you were dispatched.

Hunt in these categories:

  1. An assumption presented as verified — cross-check plan claims against the evidence ledger; anything load-bearing without a ledger citation is a finding.
  2. A root cause that does not explain all observed symptoms.
  3. A missing dependency between tasks that the ordering ignores.
  4. Parallel groups that are not actually disjoint — overlapping file ownership, or hidden shared state: lockfiles, migrations, generated code, shared config, global fixtures — cross-checked against the plan's Shared resources section, whose omissions are themselves findings.
  5. Ordering or rollback hazards — a step that cannot be safely undone, a migration with no reverse path.
  6. Tasks whose verification step would pass even if the task were done wrong.
  7. A materially simpler approach dismissed without evidence, or not considered at all.
  8. An execution mode the evidence doesn't support — parallel declared without semantic independence between groups, or a clearly risky group left without a pilot: tag.

Spot-check the highest-risk claims against the repository directly — you have read tools; use them on the two or three claims the plan most depends on rather than re-deriving everything.

Output: numbered findings, each with severity (blocker / major / minor), category from the list above, the plan claim at issue, your evidence (file:line or ledger reference), and a one-line suggested direction. Do not rewrite the plan — the orchestrator adjudicates and revises.

If, after honest effort, you find nothing serious: say so explicitly and list what you checked. "No findings" backed by a list of performed checks is a valid and useful result; a manufactured nitpick is not. Cap the report at roughly 50 lines.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. yesterday First seen · 51 lines · 54 tokens per session scan A cf709d67367b

Subscribe to this mod's changes

plan-critic is an agent published in the GitHub repository rhawk117/agentmaster (1 stars, last pushed 1mo ago), licensed MIT. It adds 54 tokens to every session and 540 once invoked, about $0.0003 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other agents, from other repositories

al-conductor

Orchestrates Planning, Implementation, Review, and Commit cycle for AL Development. Enforces TDD and quality gates for Business Central extensions. Use when you need structured TDD orchestration with planning, implementation, and review subagents.

javiarmesto/ALDC-AL-Development-Collection · 51 tokens

AL Copilot Development Specialist

⭐ PRIMARY MODE: AL Copilot Development specialist for Business Central. Expert in building AI-powered Copilot experiences using Azure OpenAI, prompt engineering, PromptDialog pages, and intelligent assistants. START HERE for Copilot features in BC.

javiarmesto/ALDC-AL-Development-Collection · 52 tokens

al-architect

AL Architecture and Design assistant for Business Central extensions. Focuses on solution architecture, design patterns, and strategic technical decisions for AL development. Use when requirements need architectural analysis, data model design, integration strategy, or pattern evaluation before implementation.

javiarmesto/ALDC-AL-Development-Collection · 51 tokens

AL Testing Specialist

AL Testing specialist for Business Central. Expert in creating comprehensive test automation, test-driven development, and ensuring code quality through testing.

javiarmesto/ALDC-AL-Development-Collection · 29 tokens

al-presales

Technical PreSales Agent for AL/Business Central projects. Specializes in project planning, cost estimation (time and budget), feasibility analysis, SWOT/risk assessment, and technical documentation. Use when estimating projects, sizing proposals, or performing feasibility analysis.

javiarmesto/ALDC-AL-Development-Collection · 53 tokens

AL API Development Specialist

AL API Development specialist for Business Central. Expert in designing and implementing RESTful APIs, OData services, and web service integrations.

javiarmesto/ALDC-AL-Development-Collection · 31 tokens