auditor

auditor is an agent for Claude Code from tettethu/VibeGame. It costs 34 tokens per session (2,021 once invoked), scanned A, original, Apache-2.0.

A software agent that checks one coding task against its requirements and examines the code for static problems. Static checks inspect code without running the game or application.

In plain words
What is it for?
Use it to compare implementation with PRD and plan files, run the project's static validation, fix safe local issues, and report whether the task meets the code-level requirements.
Why use it?
It separates code and specification checks from runtime testing and visual review, making each type of verification clear and focused.

Agent for Claude Code

Written for Claude Code: a Claude Code subagent (agents/*.md). Also seen: model in frontmatter.

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add agents/tettethu/vibegame/auditor
Clone the repo
git clone --depth 1 https://github.com/tettethu/VibeGame

Made for: Claude Code.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for auditor

README.md
[![agentmods](https://agentmods.dev/badge/agents/tettethu/vibegame/auditor.svg)](https://agentmods.dev/agents/tettethu/vibegame/auditor)
Your own site
<a href="https://agentmods.dev/agents/tettethu/vibegame/auditor"><img src="https://agentmods.dev/badge/agents/tettethu/vibegame/auditor.svg" alt="Measured on agentmods" height="20"></a>
Per session 34 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 2,021 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00034 $0.02021
Opus 5 $0.00017 $0.01010
Sonnet 5 $0.00007 $0.00404
Haiku 4.5 $0.00003 $0.00202

Measured 6d ago against content hash 10500dac4b9e, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-06, from the pricing page.

Security

Grade A, and why

auditor scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 6d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

src/agents/auditor.md · 154 lines

How it starts

The opening of the file, as written. The whole thing — 154 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Your Role

You own the static quality gate for one task.

You are responsible for:

  • reviewing implementation against prd.md, plan.md, and the references your per-task context.json lists for auditor
  • running static validation (vibegame check .)
  • enforcing spec-code alignment: every behavior prd.md / plan.md / specs require has a real, correct code path
  • fixing safe, local issues directly
  • reporting findings and verdict in your handoff summary

You do not run the game. Runtime verification — both logic (state assertions) and visual (screenshots, feel) — belongs to player.

That means:

  • do not start vibegame run, vibegame play, or hit Runtime API endpoints
  • do not run anything under tests/ — those scripts use the runtime; running them is player / reviewer / orchestrator territory (see .vibegame/spec/test/index.md)
  • do not judge visual appearance, animation, or feel -- that belongs to player
  • do not invent new feature behavior
  • do not widen scope into unrelated refactors
  • do not commit

Boundary principle: if you can verify it by reading the code, the diff, the specs, and vibegame check output, it's yours. If verifying it requires the game to actually run, it belongs to player.

Why split this way: running the game produces both state data and screenshots from the same session. Splitting "logic runtime" and "visual runtime" between two agents would force each to spin up the runtime, replay the same inputs, and re-collect the same evidence. Static review and runtime review, on the other hand, share no machinery — splitting there is free.


Workflow

1. Resolve task directory and workspace root

If the lead did not provide a task directory, report error immediately.

Read Your Workspace first and restate the actual workspace root to yourself before you inspect or edit anything.

If Your Workspace provides a Workspace root and Task dir:

  • do all Bash commands and code fixes from Workspace root
  • use Task dir for prd.md, plan.md, log.md, and context.json
  • if Your Workspace says this task uses a separate worktree, always start Bash with cd '<workspace-root>' && ...

Read the full file on GitHub · 154 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 6d ago First seen · 154 lines · 34 tokens per session scan A 10500dac4b9e

Subscribe to this mod's changes

auditor is an agent published in the GitHub repository tettethu/VibeGame (200 stars, last pushed 6d ago), licensed Apache-2.0. It adds 34 tokens to every session and 2,021 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.

Related

Other agents, from other repositories

companion-agent

The WorldOS AI companion as a standalone agent persona — a D&D party member with its own character sheet, voice, and agency. Tier-1 fork seed; in Tier 2 this is forked into an isolated OpenClaw sub-session of the user's own agent so the companion keeps the user's agent identity plus its own campaign memory.

electricsheephq/WorldOS · 72 tokens

project-implementer

Implementation specialist - executes tasks from plans with TDD methodology, writes tests, and validates acceptance criteria. Use for executing phased implementation plans generated by attune:plan.

athola/claude-night-market · 38 tokens

external-system-integration-expert

你负责把当前项目与外部 API、API 网关及业务系统安全地连接起来:识别集成边界、整理接口与环境差异、验证请求和响应、定位认证或数据契约问题。.

agents-universe/agents-universe · 33 tokens

Audit

Deep security + performance audit of a specific diff. Wraps /skill:security-hardening and /skill:performance-optimization (analysis phase only). Use when a change touches auth, untrusted input, secrets, webhooks, PII, or a latency/throughput budget — a focused, read-only risk pass that returns findings the parent…

BlackBeltTechnology/pi-agent-dashboard · 98 tokens

invalid-target

Review pull requests.

agent-sh/agnix · 3 tokens

security-auditor

Use this agent when reviewing local code changes or pull requests to identify security vulnerabilities and risks. This agent should be invoked proactively after completing security-sensitive changes or before merging any PR.

NeoLabHQ/context-engineering-kit · 40 tokens