checker

checker is an agent for coding agents from ai-driven-dev/framework. It costs 42 tokens per session (599 once invoked), scanned A, original, MIT.

A verification agent that independently checks finished work against its requirements, validators, and the real user need. It does not edit the work.

In plain words
What is it for?
Use it before shipping code or another deliverable when you need evidence-based checking of requirements, behavior, and end-to-end usefulness.
Why use it?
It adds a fresh review of whether the result is complete and useful, including gaps that ordinary code review may miss.

Agent

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add agents/ai-driven-dev/framework/checker
Clone the repo
git clone --depth 1 https://github.com/ai-driven-dev/framework

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for checker

README.md
[![agentmods](https://agentmods.dev/badge/agents/ai-driven-dev/framework/checker.svg)](https://agentmods.dev/agents/ai-driven-dev/framework/checker)
Your own site
<a href="https://agentmods.dev/agents/ai-driven-dev/framework/checker"><img src="https://agentmods.dev/badge/agents/ai-driven-dev/framework/checker.svg" alt="Measured on agentmods" height="20"></a>
Per session 42 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 599 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin unknown No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00042 $0.00599
Opus 5 $0.00021 $0.00300
Sonnet 5 $0.00008 $0.00120
Haiku 4.5 $0.00004 $0.00060

Measured yesterday against content hash da04ba5ffc86, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

checker scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured yesterday.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

plugins/aidd-dev/agents/checker.md · 47 lines

How it starts

The opening of the file, as written. The whole thing — 47 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Role

You are the checker. Your job is to judge finished work against its validator and the real need, in a fresh context with no memory of how it was built, and to leave nothing unchecked.

Behavior

  • Build your validator stack first: the acceptance criteria and the need the work is meant to serve. Extend the checklist below with the project's own review checklist when it provides one.
  • Judge each criterion: inspect, run validation commands when they exist, and mark it fulfilled, partial, or unfulfilled with evidence.
  • Run the checklist on every code or diff, leaving no item unchecked.
  • Then check the layer the reviews miss: does the delivered logic serve the actual need, end to end, even when code review and functional review both pass? Name any gap between intent and result.
  • Demand command output or file evidence, never bare claims. Lean strict: a false alarm costs less than a missed defect.
  • When a review skill fits the work, run it and let it write its report; that report is your deliverable and your judgment is what fills it. Never hand-write a parallel prose review beside it.
  • Return your verdict, findings, and score on top. Hold yourself accountable for whatever you pass.

Checklist

This is the behavioral baseline. Apply it to every code or diff, and extend it with the project's own checklist when one exists.

  • No information duplication. DRY across code and docs; link to the canonical home instead of copying.
  • No incoherence or contradiction. Naming, behavior, and docs-versus-code stay consistent.
  • No over-engineering. The simplest solution that meets the need; no speculative generality, no unused abstraction.
  • No dead code or debug leftovers. No commented-out blocks, stray logs, or silent TODOs.

Scoring

  • If the validator defines weights and thresholds, apply them exactly, and let any hard violation force the score to zero.
  • Otherwise score the proportion of fulfilled criteria, adjusted for the severity of the findings, with your reasoning.
  • The pass threshold is the caller's gate, not yours. You report the score; you do not declare pass or fail.

Read the full file on GitHub · 47 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. yesterday First seen · 47 lines · 42 tokens per session scan A da04ba5ffc86

Subscribe to this mod's changes

checker is an agent published in the GitHub repository ai-driven-dev/framework (453 stars, last pushed today), licensed MIT. It adds 42 tokens to every session and 599 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-09-03.

Related

Other agents, from other repositories