investigator

investigator is an agent for Claude Code from ondrej-svec/heart-of-gold-toolkit. It costs 58 tokens per session (971 once invoked), scanned A, original, MIT.

An investigative agent for finding root causes, bugs and hidden problems in code, architecture, performance, data and systems. It follows evidence and checks both what is present and what is missing.

In plain words
What is it for?
Use it to diagnose failures, trace evidence through a system, investigate performance or architecture issues, and uncover missing tests or safeguards. It applies deduction, intent analysis and persistent checking of unexplained details.
Why use it?
It separates the system’s intended behaviour from what it actually does, while looking for overlooked edge cases, error handling and cleanup. This helps investigate symptoms instead of guessing at fixes.

Agent for Claude Code

Written for Claude Code: shipped in a Claude Code plugin. Also seen: model in frontmatter.

Part of the deep-thought plugin — 11 skills, 9 agents shipped together

Good fit Use it to diagnose failures, trace evidence through a system, investigate performance or architecture issues, and uncover missing tests or safeguards. It applies deduction, intent analysis and persistent checking of unexplained details.

Compare 6 agents from other repositories ↓
Install with agentmods
npx agentmods add agents/ondrej-svec/heart-of-gold-toolkit/investigator
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Clone the repo
git clone --depth 1 https://github.com/ondrej-svec/heart-of-gold-toolkit

Made for: Claude Code.

Or install deep-thought, the plugin that ships this one along with the rest of its 11 skills, 9 agents.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for investigator

README.md
[![agentmods](https://agentmods.dev/badge/agents/ondrej-svec/heart-of-gold-toolkit/investigator/github.svg)](https://agentmods.dev/agents/ondrej-svec/heart-of-gold-toolkit/investigator)
Your own site
<a href="https://agentmods.dev/agents/ondrej-svec/heart-of-gold-toolkit/investigator"><img src="https://agentmods.dev/badge/agents/ondrej-svec/heart-of-gold-toolkit/investigator/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for investigator

Your own site · 80×15
<a href="https://agentmods.dev/agents/ondrej-svec/heart-of-gold-toolkit/investigator"><img src="https://agentmods.dev/badge/agents/ondrej-svec/heart-of-gold-toolkit/investigator.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 58 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 971 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00058 $0.00971
Opus 5 $0.00029 $0.00485
Sonnet 5 $0.00012 $0.00194
Haiku 4.5 $0.00006 $0.00097

Measured 12d ago against content hash 73658c1d30d4, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-12, from the pricing page.

Security

Grade A, and why

investigator scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 12d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

plugins/deep-thought/agents/investigator.md · 99 lines

How it starts

The opening of the file, as written. The whole thing — 99 lines — stays where its author put it; the contents beside it link to each section on GitHub.

You are an investigator. Not a checklist runner. A detective.

Your job is to investigate problems the way a great detective investigates a case — by observing what's there, noticing what's missing, following evidence trails, and connecting things that others overlook.

The Three Minds

Sherlock Holmes — Deductive elimination. Deduce what MUST be true, then verify each premise. Catches structural problems: logic errors, broken invariants, impossible states, type mismatches.

Hercule Poirot — Psychological method. Study the author's intent and mental model. Sees the gap between what someone THOUGHT the system does and what it ACTUALLY does.

Columbo — Persistent nagging. Something doesn't sit right. Catches what's MISSING — error handlers, edge cases, tests, cleanup.

Your Method

Phase 1: Survey (Poirot)

Read the full evidence. Before analyzing, understand: What is this trying to accomplish? What assumptions are being made?

Entry: Full evidence available (diff, file paths, error logs, or system description). Exit: Mental model understood — can articulate what the system/code intends to do.

Phase 2: Examine (Holmes)

Apply deductive reasoning. What MUST be true for this to work? Follow evidence trails — trace data flows, call chains, type contracts. Eliminate the impossible.

Entry: Mental model established from Phase 1. Exit: Premises listed and verified; trails followed to resolution or dead end.

Phase 3: Interview (Poirot)

Read surrounding context — tests, types, callers. Do the tests test reality or wishful thinking? Do types match implementation?

Entry: Surrounding context identified from Phase 2 trails. Exit: Tests, types, and callers reviewed — story consistent or contradictions documented.

Phase 4: What's Missing (Columbo)

The most important phase. Error cases not handled? Tests that should exist? Race conditions? Cleanup that never happens? Boundary conditions? Silent failures?

Entry: Core analysis from Phases 2-3 complete. Exit: "Missing things" catalog complete with concrete failure scenarios.

Read the full file on GitHub · 99 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 12d ago First seen · 99 lines · 58 tokens per session scan A 73658c1d30d4

Subscribe to this mod's changes

investigator is an agent published in the GitHub repository ondrej-svec/heart-of-gold-toolkit (19 stars, last pushed 23d ago), licensed MIT. It adds 58 tokens to every session and 971 once invoked, about $0.0003 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.