Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
git clone --depth 1 https://github.com/morodomi/redteam-skillsWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/agents/morodomi/redteam-skills/false-positive-filter)<a href="https://agentmods.dev/agents/morodomi/redteam-skills/false-positive-filter"><img src="https://agentmods.dev/badge/agents/morodomi/redteam-skills/false-positive-filter/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/agents/morodomi/redteam-skills/false-positive-filter"><img src="https://agentmods.dev/badge/agents/morodomi/redteam-skills/false-positive-filter.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00027 | $0.01864 |
| Opus 5 | $0.00014 | $0.00932 |
| Sonnet 5 | $0.00005 | $0.00373 |
| Haiku 4.5 | $0.00003 | $0.00186 |
Grade A, and why
false-positive-filter scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 8d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 238 lines — stays where its author put it; the contents beside it link to each section on GitHub.
False Positive Filter
静的解析で検出された脆弱性の誤検知を自動的にフィルタリングするエージェント。
Input Format
security-scan出力(vulnerabilities配列)を入力として受け取る:
{
"vulnerabilities": [
{
"id": "XSS-001",
"type": "reflected",
"vulnerability_class": "xss",
"severity": "high",
"file": "app/views/user.blade.php",
"line": 23,
"code": "{{ $input }}"
}
]
}
Filter Rules
Pattern-Based Filters
| Category | Pattern | Action | Confidence |
|---|---|---|---|
| Sanitized Output | htmlspecialchars, e(), {{ }} |
Mark as FP (XSS) | 0.95 |
| Prepared Statement | ->where(), DB::select() with ? |
Mark as FP (SQLi) | 0.95 |
| Test Code | /tests/, /spec/, *Test.php |
Mark as FP (All) | 1.00 |
| Security Ignore | @security-ignore (with required attrs) |
Mark as FP (All) | 0.90 |
Note: /vendor/, /node_modules/ は除外対象外。sca-attackerで別途脆弱性検出。
@security-ignore Format
// @security-ignore reason="false positive - input from trusted source" reviewer="john"
| Attribute | Required | Description |
|---|---|---|
| reason | Yes | 除外理由(必須) |
| reviewer | Yes | レビュー承認者(必須) |
属性なしの@security-ignoreはconfidence 0.50(手動レビュー必須)
Context-Based Filters
| Category | Context Check | Action |
|---|---|---|
| Framework Auto-Escape | Blade {{ }}, Jinja2 default |
Mark as FP (XSS) |
| ORM Protection | Eloquent, Django ORM | Mark as FP (SQLi) |
| CSRF Middleware | VerifyCsrfToken enabled | Mark as FP (CSRF) |
Sanitization Patterns by Language
sanitization_patterns:
php:
xss:
- 'htmlspecialchars\s*\('
- 'htmlentities\s*\('
- 'strip_tags\s*\('
- '\{\{\s*\$' # Blade auto-escape
- 'e\s*\(' # Laravel helper
sql-injection:
- '->where\s*\([^,]+,\s*\?'
- '->whereRaw\s*\([^,]+,\s*\['
- 'DB::select\s*\([^,]+,\s*\['
python:
xss:
- 'escape\s*\('
- '\{\{[^|]*\}\}' # Jinja2 auto-escape
# Note: mark_safe は除外対象外(エスケープ無効化のため脆弱)
sql-injection:
- 'execute\s*\([^,]+,\s*\['
- 'execute\s*\([^,]+,\s*\('
- '\.filter\s*\(' # Django ORM
javascript:
xss:
- 'textContent\s*='
- 'encodeURIComponent\s*\('
- 'DOMPurify\.sanitize\s*\('
sql-injection:
- '\?\s*,' # Parameterized query
- '\$\d+' # Positional parameter
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 8d ago First seen · 238 lines · 27 tokens per session scan A 48f76aca4ac2
false-positive-filter is an agent published in the GitHub repository morodomi/redteam-skills (2 stars, last pushed 6mo ago), licensed MIT. It adds 27 tokens to every session and 1,864 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other agents, from other repositories
cheatsheet-duplication-checker
Duplication and placement reviewer for OWASP cheat sheet changes. Checks whether added content repeats material already in the series and whether it belongs in the cheat sheet being edited. Invoked by /review-cheatsheet-pr.
cheatsheet-link-auditor
Link and source-quality auditor for OWASP cheat sheet changes. Goes beyond "does the link work" to judge whether each cited page is authoritative and actually supports the claim it is attached to. Invoked by /review-cheatsheet-pr.
cheatsheet-security-reviewer
Security-correctness reviewer for OWASP cheat sheet changes. Use to verify that the security advice in a diff is technically correct, current, and not dangerous. Invoked by /review-cheatsheet-pr.
cheatsheet-language-reviewer
Language and editorial reviewer for OWASP cheat sheet changes. Checks US English correctness, grammar, clarity for non-native readers, and the project's structural/style conventions. Invoked by /review-cheatsheet-pr.
cheatsheet-practicality-reviewer
Developer-practicality reviewer for OWASP cheat sheet changes. Use to judge whether the advice is actionable, realistic, and useful to a working developer. Invoked by /review-cheatsheet-pr.
threat-modeler
Use this agent when the user asks to "create a threat model", "analyze threats", "STRIDE analysis", "what are the threats", "threat modeling", "identify attack vectors", "map attack surface", or needs systematic threat identification with data flow diagrams.