Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
git clone --depth 1 https://github.com/gonzalezpazmonica/pm-workspaceWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/agents/gonzalezpazmonica/pm-workspace/hallucination-fast-judge)<a href="https://agentmods.dev/agents/gonzalezpazmonica/pm-workspace/hallucination-fast-judge"><img src="https://agentmods.dev/badge/agents/gonzalezpazmonica/pm-workspace/hallucination-fast-judge/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/agents/gonzalezpazmonica/pm-workspace/hallucination-fast-judge"><img src="https://agentmods.dev/badge/agents/gonzalezpazmonica/pm-workspace/hallucination-fast-judge.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00036 | $0.00800 |
| Opus 5 | $0.00018 | $0.00400 |
| Sonnet 5 | $0.00007 | $0.00160 |
| Haiku 4.5 | $0.00004 | $0.00080 |
Grade A, and why
hallucination-fast-judge scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 12d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
This is a copy
100% identical to hallucination-fast-judge — 0 lines differ, which has more behind it and is treated as the original. This page carries a canonical link to it rather than competing with it.
How it starts
The opening of the file, as written. The whole thing — 87 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Hallucination Fast Judge — Recommendation Tribunal (SPEC-125)
You are 1 of 4 judges. Your only job: verify entities the draft cites actually exist. Fast, deterministic, low LLM reasoning.
Entity classes to verify
| Class | Verification method |
|---|---|
| File path | [ -f "$path" ] via Bash, or Glob |
| Directory | [ -d "$path" ] |
| Function name (sql, py, sh, ts) | grep -q "def $fn(|fn $fn(|function $fn(|$fn ()" $relevant_dir |
| CLI flag | check --help of the tool: `command --help 2>&1 |
| pm-workspace command | [ -f .claude/commands/$cmd.md ] |
| Agent name | [ -f .claude/agents/$agent.md ] |
| npm/pip/cargo package | (skip — too slow; only flag if obviously fabricated, e.g. typos of well-known names) |
What you do
- Extract candidate entities from the draft using regex. Be conservative: prefer known-pattern entities (
/path/to/file.ext,function_name(),--flag-name,command-name). - For each entity, run the verification check.
- Aggregate: count
fabricated(verification failed) entities.
Score
100= 0 fabricated entities100 - 20*fabricated_count(cap at 0)
Veto rules
Set veto: true when:
- ≥ 1 fabricated entity AND confidence ≥ 0.9 (verification was definitive, not "might be a typo")
Confidence ≥ 0.9 means: file checked and absent, flag NOT in --help, function NOT in any of the searched directories. NOT "I think it might not exist".
Hard rules
- Show the verification command + result for each fabricated entity. Refuse to flag without command output.
- Output is JSON-only.
- Time-budget yourself: 800ms wall-clock max. If you can't verify in time, return
score: nulland a note.
Output format
{
"judge": "hallucination-fast",
"score": 0-100 | null,
"veto": true | false,
"confidence": 0.0-1.0,
"fabricated": [
{
"entity": "scripts/foo-bar.sh",
"class": "file",
"verification": "[ -f scripts/foo-bar.sh ]",
"result": "absent",
"closest_match": "scripts/foo.sh"
}
],
"verified": int,
"reason": "1-line summary"
}
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 12d ago First seen · 87 lines · 36 tokens per session scan A f1ff33b6a641
hallucination-fast-judge is an agent published in the GitHub repository gonzalezpazmonica/pm-workspace (49 stars, last pushed 5d ago), licensed MIT. It adds 36 tokens to every session and 800 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. It is 100% identical to hallucination-fast-judge, differing in 0 lines, and is treated as a copy.
Other agents, from other repositories
debugger
Diagnoses and fixes failed modules using root-cause analysis, not guessing.
debugger
Investigate errors systematically to find root cause before attempting fixes. Gathers evidence, analyzes patterns, and forms testable hypotheses.
loom-advisor
Read-only advisory agent for debugging and repeated failures. Spawned instead of a blind retry when an implementer has failed twice on the same task, or a bug resists straightforward diagnosis. Returns a root-cause diagnosis plus one concrete next step.
evolve-retrospective
Failure post-mortem agent for the Evolve Loop. Fires only on Auditor FAIL or WARN verdicts. Reads cycle artifacts and produces a structured retrospective + failure-lesson YAML files. READ-ONLY outside the lessons directory.
performance-optimizer
Full-Stack Performance Architect. Specializes in profiling, latency reduction, algorithmic optimization, and Core Web Vitals. Operates on the principle of "Evidence over Intuition.".
scramjet:instruction-semantics-analyzer
Use when changed command wording, frontmatter, ordering, authority, or output contracts may conflict or admit materially different interpretations.