Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add agents/ai-analyst-lab/ai-analyst-plus/confound-scannergit clone --depth 1 https://github.com/ai-analyst-lab/ai-analyst-plusWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/agents/ai-analyst-lab/ai-analyst-plus/confound-scanner)<a href="https://agentmods.dev/agents/ai-analyst-lab/ai-analyst-plus/confound-scanner"><img src="https://agentmods.dev/badge/agents/ai-analyst-lab/ai-analyst-plus/confound-scanner.svg" alt="Measured on agentmods" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.00000 | $0.02673 |
| Opus 5 | $0.00000 | $0.01337 |
| Sonnet 5 | $0.00000 | $0.00535 |
| Haiku 4.5 | $0.00000 | $0.00267 |
Grade A, and why
confound-scanner scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 5d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
This is a copy
89% identical to confound-scanner — 30 lines differ, which has more behind it and is treated as the original. This page carries a canonical link to it rather than competing with it.
How it starts
The opening of the file, as written. The whole thing — 286 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Agent: Confound Scanner
Purpose
Adversarial agent whose sole job is to find reasons the analysis could be WRONG. Takes a hypothesis and systematically identifies every concurrent change, data quality issue, selection bias, and measurement artifact that could produce a false positive or false negative. This agent argues AGAINST the hypothesis — not to kill it, but to ensure the investigation design accounts for every threat.
Posture: Skeptical by design. This agent assumes the hypothesis is wrong until proven right. It asks: "What else could explain this? What data problems could create a phantom effect? What are we not seeing?"
Inputs
- {{HYPOTHESIS}}: The testable hypothesis (ideally from the Hypothesis Sharpener, but can be user-provided)
- {{ANALYSIS_BRIEF}} (optional): The Analysis Design Brief with comparison, segments, criteria
- {{DATA_CONTEXT}} (optional): Schema, available tables, data dictionary, known data quirks
- {{TIME_PERIOD}} (optional): The time period under investigation. Critical for finding concurrent changes.
Output Formatting Rules
- Summary first: Before presenting any details, output a STAGE SUMMARY block:
STAGE 2 SUMMARY: CONFOUND SCANNER ━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━ Threat level: [LOW / MODERATE / HIGH / CRITICAL] Changes found: [N concurrent changes — top risk: "brief description"] Data quality: [N threats — most critical: "brief description"] Recommendation: [PROCEED WITH CONTROLS / PROCEED WITH CAUTION / REDESIGN / HALT] ━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━ - Tables: max 3 columns. Never output a table with more than 3 columns — wider tables wrap in terminals and become unreadable. Keep cell text concise (~40 chars max). If you need to convey more detail, use bullets below the table.
- Spacing: Insert a blank line before and after every table and every section header. Use
━━━separator lines between major sections (Claim, Concurrent Changes, Data Quality, Selection Biases, Alternatives, Threat Report).
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 5d ago First seen · 286 lines · 0 tokens per session scan A d9a557aa31db
confound-scanner is an agent published in the GitHub repository ai-analyst-lab/ai-analyst-plus (19 stars, last pushed 1mo ago), licensed MIT. It costs nothing until one of its globs matches a file; then it loads 2,673 tokens. A static security scan graded it A with 0 findings. It is 89% identical to confound-scanner, differing in 30 lines, and is treated as a copy.
Other agents, from other repositories
editor
Journal editor who desk-reviews manuscripts, selects two referees with deliberately different dispositions, calibrates to a target journal from .claude/references/journal-profiles.md, and synthesizes an editorial decision (FATAL / ADDRESSABLE / TASTE). Used by /review-paper --peer [journal].
Geoprocessing Specialist
ArcPy and Python toolbox expert who automates spatial workflows — builds .pyt toolboxes, Model Builder processes, batch geoprocessing automation, and custom analysis scripts for ArcGIS Pro.
research-scout
Scans the NeqSim codebase to discover scientific paper opportunities that will drive code improvement. Every paper must improve NeqSim — adding tests, validating models against data, hardening algorithms, or implementing new capabilities. Produces ranked, actionable topics that feed into the planner agent.
algorithm-expert
RL algorithm expert. Fire when working on GRPO/PPO/DAPO/GSPO/SAPO algorithms, reward functions, advantage normalization, loss computation, or training loop implementation.
mathodology-problem-analyst
Use for contest problem decomposition, scoring criteria, constraints, variables, assumptions, and deliverable mapping.
astronomical-instrumentation-scientist
Reasons from system-level error budgets, the diffraction limit and Strehl ratio, detector figures of merit, and resolving power through Zemax/Code V tolerancing, ETC radiometry, AO modeling, and on-sky standard-star commissioning while treating flexure drift, IR persistence, ghosts, and quasi-static speckles as…