bug-finder

bug-finder is a skill for Claude Code, Codex from cmdr-chara/codex-toolkit. It costs 97 tokens per session (2,487 once invoked), scanned A, original, MIT.

A structured hunt for previously unknown correctness bugs in an existing code repository.

In plain words
What is it for?
Use it to derive expected behavior, inspect high-risk code, investigate races and data loss, and prove or discard possible lifecycle, protocol, retry, or concurrency bugs.
Why use it?
It turns vague suspicion into specific, testable bug candidates instead of treating style problems or theoretical risks as confirmed defects.

Skill for Claude CodeCodex

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/cmdr-chara/codex-toolkit/bug-finder
Any agent
npx skills add cmdr-chara/codex-toolkit --skill bug-finder
Clone the repo
git clone --depth 1 https://github.com/cmdr-chara/codex-toolkit

Made for: Claude Code, Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for bug-finder

README.md
[![agentmods](https://agentmods.dev/badge/skills/cmdr-chara/codex-toolkit/bug-finder.svg)](https://agentmods.dev/skills/cmdr-chara/codex-toolkit/bug-finder)
Your own site
<a href="https://agentmods.dev/skills/cmdr-chara/codex-toolkit/bug-finder"><img src="https://agentmods.dev/badge/skills/cmdr-chara/codex-toolkit/bug-finder.svg" alt="Measured on agentmods" height="20"></a>
Per session 97 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 2,487 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00097 $0.02487
Opus 5 $0.00048 $0.01243
Sonnet 5 $0.00019 $0.00497
Haiku 4.5 $0.00010 $0.00249

Measured 3d ago against content hash ec652ea33408, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

bug-finder scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 3d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

skills/bug-finder/SKILL.md · 256 lines

How it starts

The opening of the file, as written. The whole thing — 256 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Bug Finder

Find real correctness defects the user has not already identified. This skill owns discovery and candidate proof. It does not replace causal debugging, code review, broad improvement planning, security review, or release approval.

A useful bug finding is an observable contract violation with enough evidence that another engineer can reproduce, falsify, or investigate it. Suspicious code, style debt, theoretical risk, and maintainability concerns are not bugs by themselves.

Trigger boundary

Use this skill when the request is open-ended correctness hunting, for example:

  • find important bugs in this repository;
  • look for hidden races, lifecycle errors, stale-state paths, data-loss conditions, or incorrect edge cases;
  • inspect a subsystem for unknown correctness defects before users report them;
  • search for bugs in provider adapters, protocol handling, retries, persistence, streaming, cancellation, or concurrency;
  • produce a ranked set of concrete bug candidates with proof attempts.

Do not trigger for:

  • a known crash, hang, wrong result, flaky test, or production incident whose cause is uncertain — use debugging-investigator;
  • reviewing a particular branch, PR, commit, or diff — use review-and-refactor-code;
  • asking what the repository should improve overall — use codebase-improvement-planner;
  • a pure performance hunt with a named metric — use optimize-codebase-performance;
  • security-only vulnerability hunting — use the security workflow available in the environment;
  • final ship/no-ship judgment — use verification-and-release.

If the repository boundary is not understood, use repository-intelligence first or request its map as a prerequisite. Do not rediscover an entire large repository when a current map already exists.

Required inputs

Resolve before hunting:

  1. repository and branch/candidate identity;
  2. requested scope, exclusions, and protected user work;
  3. architecture/ownership map when the system is nontrivial;
  4. public and internal contracts relevant to the scope;
  5. available tests, fixtures, logs, protocol schemas, state machines, and failure-handling code;
  6. permission boundaries for running tests, starting services, creating temporary fixtures, or adding instrumentation.

Read the full file on GitHub · 256 lines

Files

What ships with it

2 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 3d ago First seen · 256 lines · 97 tokens per session scan A ec652ea33408

Subscribe to this mod's changes

bug-finder is a skill published in the GitHub repository cmdr-chara/codex-toolkit (2 stars, last pushed 5d ago), licensed MIT. It adds 97 tokens to every session and 2,487 once invoked, about $0.0005 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.