Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
/plugin marketplace add tonyblu331/research-proofnpx agentmods add plugins/tonyblu331/research-proof/research-proof-plugingit clone --depth 1 https://github.com/tonyblu331/research-proofGrade A, and why
research-proof-plugin scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 2d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
What it actually says
{
"name": "research-proof-plugin",
"description": "Falsifiable research planning for AI agents: pressure-test claims, freeze evaluators, use AI-lab and medical evidence patterns, build proof ladders, and maintain proof ledgers.",
"version": "1.3.0",
"author": {
"name": "Tony Blanco"
},
"homepage": "https://github.com/tonyblu331/research-proof",
"repository": "https://github.com/tonyblu331/research-proof",
"license": "MIT",
"keywords": [
"agent-skills",
"claude-code",
"research",
"research-agent",
"ai-research",
"evaluation",
"evals",
"benchmarking",
"proof",
"proof-ledger",
"falsification",
"scientific-method",
"medical-research",
"evidence-certainty",
"prompt-engineering",
"autonomous-agents"
]
}
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 2d ago First seen · 30 lines scan A 0566be1f807a
research-proof-plugin is a plugin published in the GitHub repository tonyblu331/research-proof (45 stars, last pushed 3mo ago), licensed MIT. Its token cost is not measured: this kind of file is read by the harness, not the model. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other plugins, from other repositories
airas
Automated research with AIRAS: paper search, hypothesis generation, experiment execution on GitHub Actions/AIXS, figure rendering, and paper writing — bundles the AIRAS MCP server (uvx airas) and research-workflow skills.
airas
Automated ML research tools: the AIRAS MCP server plus research-workflow skills.
agent-lifecycle-kit
Plugin marketplace listing 1 plugin: agent-lifecycle-kit.
vukkt-plugins
Plugins by Vuk Topalovic — currently token-warden, a measurement and selection layer that makes Claude Code subagents measurably cheaper over time.
warranted
Make AI research agents accountable — give every conclusion a traceable argument graph.
tmcp
Standalone TMCP skill-packet workflows, first-run checks, skill harvesting, adaptive workflow-pack recommendations, packet substance checks, and expert rubric remediation for Claude Code.