falsify

A research check that searches the web for evidence contradicting a finding. It can mark findings for removal from consideration, lower their confidence, or add notes based on the result.

In plain words
What is it for?
Use it to challenge claims in an active research session and record how each claim should be treated after the search.
Why use it?
It helps catch conclusions that appear true only because opposing evidence was not checked.

Command

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add commands/zircote-plugins/sigint/falsify
Clone the repo
git clone --depth 1 https://github.com/zircote-plugins/sigint
Per session 27 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 171 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00027 $0.00171
Opus 5 $0.00014 $0.00086
Sonnet 5 $0.00005 $0.00034
Haiku 4.5 $0.00003 $0.00017

Measured 2d ago against content hash 6498b745ae00, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

falsify scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 2d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

commands/falsify.md · 13 lines

What it actually says

Load and execute the sigint:falsify skill.

$ARGUMENTS are passed through to the skill as-is.

Run adversarial falsification on the active research session now based on: $ARGUMENTS

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 2d ago First seen · 13 lines · 27 tokens per session scan A 6498b745ae00

Subscribe to this mod's changes

falsify is a command published in the GitHub repository zircote-plugins/sigint (20 stars, last pushed 16d ago), licensed MIT. It adds 27 tokens to every session and 171 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other commands, from other repositories

radar-council

The on-demand strategic war room. Spins up a REAL experimental agent team of 4 analyst teammates (pricing, product gap, threat and sentiment, devil's advocate) over the latest briefing. They message and challenge each other like a scientific debate, converge on recommended moves, and the addendum is published to all…

sabahudin-web/competitive-intelligence-radar · 73 tokens

radar

The weekly competitor gather run. Fans out one competitor-scout subagent per competitor in parallel, each returning a cited dossier via BrightData, then diffs against last week, merges one briefing, and publishes to all three surfaces (local folder, Notion, HTML dashboard). Runs autonomously. Run /radar-setup first.

sabahudin-web/competitive-intelligence-radar · 67 tokens

radar-setup

One-time setup for Competitive Intelligence Radar. Asks about your BrightData and Notion connection and live-tests both, then shows the proposed Notion schema for your approval before creating anything, then checks the agent-teams prerequisites for the council. Run this before /radar.

sabahudin-web/competitive-intelligence-radar · 55 tokens

build-market-research

Build comprehensive market research with industry analysis, future trends, and strategic opportunities.

armoin2018/ai-ley · 14 tokens

fixbot

You are the orchestrator for the fixbot workflow. Fixbot fixes a Linear issue, running directly in this project against the locally running server. For an isolated worktree version, use /autobot /fixbot instead.

metabase/metabase · 0 tokens

reprobot

You are the orchestrator for the reprobot workflow. ReproBot attempts to reproduce a reported bug against the locally running server, classifies the result, and optionally writes a failing test. For an isolated worktree version, use /autobot /reprobot instead.

metabase/metabase · 0 tokens