false-positive-filter

false-positive-filter is an agent for Claude Code from morodomi/redteam-skills. It costs 27 tokens per session (1,864 once invoked), scanned A, original, MIT.

An agent that reviews findings from static security analysis and marks likely false positives. Static analysis checks source code for possible problems without running the application.

In plain words
What is it for?
Use it to filter vulnerability reports for patterns such as sanitized XSS output, prepared SQL queries, test code, and approved security-ignore comments.
Why use it?
It reduces the time spent reviewing warnings caused by safe patterns, such as escaped output or prepared SQL statements. It also distinguishes ignored findings that still need manual review when their required details are missing.

Agent for Claude Code

Written for Claude Code: allowed-tools in frontmatter.

Part of the redteam-core plugin — 4 skills, 18 agents shipped together

Good fit Use it to filter vulnerability reports for patterns such as sanitized XSS output, prepared SQL queries, test code, and approved security-ignore comments.

Compare 6 agents from other repositories ↓
Install with agentmods
npx agentmods add agents/morodomi/redteam-skills/false-positive-filter
Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

Clone the repo
git clone --depth 1 https://github.com/morodomi/redteam-skills

Made for: Claude Code.

Or install redteam-core, the plugin that ships this one along with the rest of its 4 skills, 18 agents.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for false-positive-filter

README.md
[![agentmods](https://agentmods.dev/badge/agents/morodomi/redteam-skills/false-positive-filter/github.svg)](https://agentmods.dev/agents/morodomi/redteam-skills/false-positive-filter)
Your own site
<a href="https://agentmods.dev/agents/morodomi/redteam-skills/false-positive-filter"><img src="https://agentmods.dev/badge/agents/morodomi/redteam-skills/false-positive-filter/github.svg" alt="Measured on agentmods" height="20"></a>

Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.

agentmods 80×15 button for false-positive-filter

Your own site · 80×15
<a href="https://agentmods.dev/agents/morodomi/redteam-skills/false-positive-filter"><img src="https://agentmods.dev/badge/agents/morodomi/redteam-skills/false-positive-filter.svg" alt="Reviewed on agentmods" width="80" height="20"></a>
Per session 27 Only the description is in the session, so the agent can decide to use it. The body loads when it is invoked.
When invoked 1,864 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. A grade says what 26 rules found in the file — not that it is safe.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5.1 $0.00027 $0.01864
Opus 5 $0.00014 $0.00932
Sonnet 5 $0.00005 $0.00373
Haiku 4.5 $0.00003 $0.00186

Measured 8d ago against content hash 48f76aca4ac2, method: parsed. Prices are Anthropic first-party input rates as of 2026-09-08, from the pricing page.

Security

Grade A, and why

false-positive-filter scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 8d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

plugins/redteam-core/agents/false-positive-filter.md · 238 lines

How it starts

The opening of the file, as written. The whole thing — 238 lines — stays where its author put it; the contents beside it link to each section on GitHub.

False Positive Filter

静的解析で検出された脆弱性の誤検知を自動的にフィルタリングするエージェント。

Input Format

security-scan出力(vulnerabilities配列)を入力として受け取る:

{
  "vulnerabilities": [
    {
      "id": "XSS-001",
      "type": "reflected",
      "vulnerability_class": "xss",
      "severity": "high",
      "file": "app/views/user.blade.php",
      "line": 23,
      "code": "{{ $input }}"
    }
  ]
}

Filter Rules

Pattern-Based Filters

Category Pattern Action Confidence
Sanitized Output htmlspecialchars, e(), {{ }} Mark as FP (XSS) 0.95
Prepared Statement ->where(), DB::select() with ? Mark as FP (SQLi) 0.95
Test Code /tests/, /spec/, *Test.php Mark as FP (All) 1.00
Security Ignore @security-ignore (with required attrs) Mark as FP (All) 0.90

Note: /vendor/, /node_modules/ は除外対象外。sca-attackerで別途脆弱性検出。

@security-ignore Format

// @security-ignore reason="false positive - input from trusted source" reviewer="john"
Attribute Required Description
reason Yes 除外理由(必須)
reviewer Yes レビュー承認者(必須)

属性なしの@security-ignoreはconfidence 0.50(手動レビュー必須)

Context-Based Filters

Category Context Check Action
Framework Auto-Escape Blade {{ }}, Jinja2 default Mark as FP (XSS)
ORM Protection Eloquent, Django ORM Mark as FP (SQLi)
CSRF Middleware VerifyCsrfToken enabled Mark as FP (CSRF)

Sanitization Patterns by Language

sanitization_patterns:
  php:
    xss:
      - 'htmlspecialchars\s*\('
      - 'htmlentities\s*\('
      - 'strip_tags\s*\('
      - '\{\{\s*\$'  # Blade auto-escape
      - 'e\s*\('     # Laravel helper
    sql-injection:
      - '->where\s*\([^,]+,\s*\?'
      - '->whereRaw\s*\([^,]+,\s*\['
      - 'DB::select\s*\([^,]+,\s*\['

  python:
    xss:
      - 'escape\s*\('
      - '\{\{[^|]*\}\}'  # Jinja2 auto-escape
      # Note: mark_safe は除外対象外(エスケープ無効化のため脆弱)
    sql-injection:
      - 'execute\s*\([^,]+,\s*\['
      - 'execute\s*\([^,]+,\s*\('
      - '\.filter\s*\('  # Django ORM

  javascript:
    xss:
      - 'textContent\s*='
      - 'encodeURIComponent\s*\('
      - 'DOMPurify\.sanitize\s*\('
    sql-injection:
      - '\?\s*,'  # Parameterized query
      - '\$\d+'   # Positional parameter

Read the full file on GitHub · 238 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 8d ago First seen · 238 lines · 27 tokens per session scan A 48f76aca4ac2

Subscribe to this mod's changes

false-positive-filter is an agent published in the GitHub repository morodomi/redteam-skills (2 stars, last pushed 6mo ago), licensed MIT. It adds 27 tokens to every session and 1,864 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other agents, from other repositories

cheatsheet-duplication-checker

Duplication and placement reviewer for OWASP cheat sheet changes. Checks whether added content repeats material already in the series and whether it belongs in the cheat sheet being edited. Invoked by /review-cheatsheet-pr.

OWASP/CheatSheetSeries · 52 tokens

cheatsheet-link-auditor

Link and source-quality auditor for OWASP cheat sheet changes. Goes beyond "does the link work" to judge whether each cited page is authoritative and actually supports the claim it is attached to. Invoked by /review-cheatsheet-pr.

OWASP/CheatSheetSeries · 56 tokens

cheatsheet-security-reviewer

Security-correctness reviewer for OWASP cheat sheet changes. Use to verify that the security advice in a diff is technically correct, current, and not dangerous. Invoked by /review-cheatsheet-pr.

OWASP/CheatSheetSeries · 49 tokens

cheatsheet-language-reviewer

Language and editorial reviewer for OWASP cheat sheet changes. Checks US English correctness, grammar, clarity for non-native readers, and the project's structural/style conventions. Invoked by /review-cheatsheet-pr.

OWASP/CheatSheetSeries · 48 tokens

cheatsheet-practicality-reviewer

Developer-practicality reviewer for OWASP cheat sheet changes. Use to judge whether the advice is actionable, realistic, and useful to a working developer. Invoked by /review-cheatsheet-pr.

OWASP/CheatSheetSeries · 49 tokens

threat-modeler

Use this agent when the user asks to "create a threat model", "analyze threats", "STRIDE analysis", "what are the threats", "threat modeling", "identify attack vectors", "map attack surface", or needs systematic threat identification with data flow diagrams.

allsmog/vuln-scout · 60 tokens