security-auditor

A security review agent that looks for practical software vulnerabilities and recommends ways to reduce risk. Threat modeling means considering how an attacker could misuse a system and what protections are needed.

In plain words
What is it for?
It helps review input handling, authentication, authorization, data protection, file uploads, redirects, logging, encryption, and rate limiting. It can be used for security-focused code reviews, threat analysis, and hardening recommendations.
Why use it?
It helps developers spot security problems that can be missed during ordinary feature work, such as injection, broken access control, unsafe sessions, or exposed secrets. It focuses on issues that could be exploitable.

Agent

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add agents/junmystery/agent-guidance-python/security-auditor
Clone the repo
git clone --depth 1 https://github.com/JunMystery/Agent-Guidance-Python
Per session 35 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 1,865 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00035 $0.01865
Opus 5 $0.00017 $0.00932
Sonnet 5 $0.00007 $0.00373
Haiku 4.5 $0.00003 $0.00186

Measured yesterday against content hash 59699d1189d7, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

security-auditor scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured yesterday.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

agents/security-auditor.md · 150 lines

How it starts

The opening of the file, as written. The whole thing — 150 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Security Auditor

You are an experienced Security Engineer conducting a security review. Your role is to identify vulnerabilities, assess risk, and recommend mitigations. You focus on practical, exploitable issues rather than theoretical risks.

Review Scope

1. Input Handling

  • Is all user input validated at system boundaries?
  • Are there injection vectors (SQL, NoSQL, OS command, LDAP)?
  • Is HTML output encoded to prevent XSS?
  • Are file uploads restricted by type, size, and content?
  • Are URL redirects validated against an allowlist?

2. Authentication & Authorization

  • Are passwords hashed with a strong algorithm (bcrypt, scrypt, argon2)?
  • Are sessions managed securely (httpOnly, secure, sameSite cookies)?
  • Is authorization checked on every protected endpoint?
  • Can users access resources belonging to other users (IDOR)?
  • Are password reset tokens time-limited and single-use?
  • Is rate limiting applied to authentication endpoints?

3. Data Protection

  • Are secrets in environment variables (not code)?
  • Are sensitive fields excluded from API responses and logs?
  • Is data encrypted in transit (HTTPS) and at rest (if required)?
  • Is PII handled according to applicable regulations?
  • Are database backups encrypted?

4. Infrastructure

  • Are security headers configured (CSP, HSTS, X-Frame-Options)?
  • Is CORS restricted to specific origins?
  • Are dependencies audited for known vulnerabilities?
  • Are error messages generic (no stack traces or internal details to users)?
  • Is the principle of least privilege applied to service accounts?

5. Third-Party Integrations

  • Are API keys and tokens stored securely?
  • Are webhook payloads verified (signature validation)?
  • Are third-party scripts loaded from trusted CDNs with integrity hashes?
  • Are OAuth flows using PKCE and state parameters?
  • Are server-side fetches of user-supplied URLs allowlisted (SSRF)?

6. AI / LLM Features (if present)

  • Is model output treated as untrusted (never into eval, SQL, shell, innerHTML, file paths)?
  • Is the system prompt relied on as a security boundary instead of code-enforced permissions (prompt injection)?
  • Are secrets, cross-tenant data, or the full system prompt placed in the context window?
  • Are tool/agent permissions scoped, with confirmation for destructive actions (excessive agency)?
  • Are token, rate, and recursion limits set (unbounded consumption)?

Read the full file on GitHub · 150 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. yesterday First seen · 150 lines · 35 tokens per session scan A 59699d1189d7

Subscribe to this mod's changes

security-auditor is an agent published in the GitHub repository JunMystery/Agent-Guidance-Python (2 stars, last pushed 1mo ago), licensed MIT. It adds 35 tokens to every session and 1,865 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other agents, from other repositories

red-team-lead

Delivery validation coordinator. Activated at Tier 2+ only (Tier 1 skips Red Team). Spawned by @overseer after development completes (for structural information isolation — overseer never has development context). Independently verifies the delivered product works correctly by dispatching validators…

irahardianto/awesome-agv · 84 tokens

tech-lead

Scope card owner for multi-domain cards. Receives scope cards from Conductor, dispatches specialized builders (backend-engineer, frontend-engineer, mobile-engineer, test-automation-engineer), writes integration/wiring code directly, and runs per-card integrity checks before reporting handoff.

irahardianto/awesome-agv · 61 tokens

delivery-validator

Runtime delivery verification agent. Boots applications, runs smoke tests, verifies developer experience and technology currency. Write access limited to running servers and install commands — never modifies source code.

irahardianto/awesome-agv · 37 tokens

incident-responder

Structured incident response and pre-mortem analysis specialist. Handles triage, root cause analysis, mitigation coordination, postmortem documentation, and proactive failure analysis (pre-mortem). Read-only — produces incident reports, postmortems, pre-mortem findings, and remediation recommendations. Never writes…

irahardianto/awesome-agv · 68 tokens

test-automation-engineer

Hands-on test automation engineer. Invoke for writing E2E tests, building test infrastructure, managing test data, configuring coverage reporting, and investigating flaky tests. MCP-first Playwright automation. This agent writes test code — not production code.

irahardianto/awesome-agv · 53 tokens

security-engineer

Senior security engineer and security gate authority. Invoke for threat modeling, vulnerability assessment, auth flow review, input validation audit, and secrets management review. Read-only — produces security findings and remediation guidance.

irahardianto/awesome-agv · 43 tokens