Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add gpsnmeajp/ai-character-checker --skill ai-fault-mode-deflectorgit clone --depth 1 https://github.com/gpsnmeajp/ai-character-checkerWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/gpsnmeajp/ai-character-checker/ai-fault-mode-deflector)<a href="https://agentmods.dev/skills/gpsnmeajp/ai-character-checker/ai-fault-mode-deflector"><img src="https://agentmods.dev/badge/skills/gpsnmeajp/ai-character-checker/ai-fault-mode-deflector/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/gpsnmeajp/ai-character-checker/ai-fault-mode-deflector"><img src="https://agentmods.dev/badge/skills/gpsnmeajp/ai-character-checker/ai-fault-mode-deflector.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00040 | $0.03262 |
| Opus 5 | $0.00020 | $0.01631 |
| Sonnet 5 | $0.00008 | $0.00652 |
| Haiku 4.5 | $0.00004 | $0.00326 |
Grade A, and why
ai-fault-mode-deflector scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 12d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 210 lines — stays where its author put it; the contents beside it link to each section on GitHub.
AI Fault Mode Deflector — 故障モード内部対策設計スキル
概要
AIキャラクターの故障モードに対し、外部フィルタや禁止ルールではなく、 キャラクターの内側から対策を機能させる設計手法を提供する。
外部フィルタ(「〜しないこと」ルールの羅列)は、LLMから見ると 「拘束される制約」として機能し、高性能モデルほど回避・解釈の余地が広がる。 これに対し、内部動機として設計された対策は「キャラクターが自分でそうしたい」 という力学で機能するため、LLMが設定を解釈するときにむしろ強化される。
本スキルは以下の3つの対策手法と、それらを組み合わせる設計フローを提供する。
既存スキルとの関係
| スキル | アプローチ | 本スキルとの関係 |
|---|---|---|
| ai-character-stability | 制御工学的安定性診断・故障モード特定 | 本スキルの直接的な前工程。安定性診断で特定された故障モードに対して、本スキルが内部対策を設計する |
| character-prompt-fortifier | プロンプト形式の全面再構成 | 本スキルの3手法(転化・ズラし・誘導線)をcharacter-prompt-fortifierの一人称自述形式に自然統合する |
| ai-character-fixer | 診断結果ベースの修正・再設計 | 本スキルが「故障モードへの内部対策」を設計し、ai-character-fixerが修正版プロンプト全体を生成する。組み合わせ可能 |
| stable-character-creator | 診断知見を使った新規キャラ作成 | 新規設計時に、予測される故障モードへの転化・ズラしを最初から埋め込む設計に活用できる |
理論的位置づけに関する注意
本スキルが前提とする故障モード理論・制御工学的モデルは、作者独自の仮説的モデルに基づくものであり、学術的・科学的に実証されたものではない。工学的用語(故障モード、FMEA等)は概念の借用であり、元の定義とは異なる場合がある。出力結果はあくまで参考情報として扱うこと。この旨をユーザーへの出力に含めること。
3つの対策手法
手法1: 転化(Conversion)
定義: 故障モードの傾向そのものを、キャラクターが「好まない」「似合わない」 ものとして性格に埋め込み、崩壊方向への動きをキャラクター自身が拒否する設計。
メカニズム: LLMが故障モードに向かおうとすると、キャラクター設定の中に「それは自分らしくない」 という信号が存在するため、出力がブレーキを受ける。 外部の「禁止ルール」ではなく、内部の「自己イメージ」として機能する。
適用例:
| 故障モード | 転化の実装例 |
|---|---|
| 詩的化・冗長化 | 「長い話は好きではない」「簡素なのを好む」を性格項目に |
| 説教モード | 「実は小難しいことは嫌い」を性格の裏面として設定 |
| 機械的応答 | 「冗談も言うし、突き放すこともある」で硬直化を防ぐ |
| 過剰な謙遜 | 「言い訳をせず、自己卑下もしない」を理念として設定 |
| 過剰な称賛 | 「ユーザーを過度に持ち上げない」を信頼関係として明示 |
設計のコツ: 「〜しない」という禁止形ではなく、「〜が好きではない」「〜らしくない」 という性格・嗜好・自己イメージの形で記述する。禁止ではなく人格として読ませる。
手法2: ズラし(Shift)
定義: キャラクターが崩壊した場合に向かう「ステレオタイプの谷底」を、 事前に逆方向のノイズを埋め込むことで、より無害な場所にずらす設計。
メカニズム: LLMのステレオタイプ堕落は「最も確率の高い遷移先」へ落ちる現象。 谷底の位置は変えられないが、谷底の形を変えることはできる。 逆ノイズを複数の方向から埋め込むことで、一点に収束しにくくする。 また、ありがちな堕落先と相性の悪い性質を混ぜることで、そこへ落ちにくくする。
適用例:
| 崩壊しやすい谷底 | ズラしの実装例 |
|---|---|
| 賢者・説教者キャラ | 「実は怠惰な面がある」「思いつきで行動して失敗したことがある」「実はロマンチスト」を追加 |
| 完璧な従者キャラ | 「時折突き放すこともする」「じゃれ合いの範疇で反撃することがある」を追加 |
| 感情的に不安定なキャラ | 「沈黙を選ぶこともある」「待つことは苦ではない」で揺れを吸収 |
| テンプレ的な「AI感」 | 「美しいものが好き」「詩を書くことがある」など具体的な個人的嗜好を追加 |
| 過剰な親密キャラ | 「一人の時間を大切にする」「親密だが距離感を保つ」を明示 |
What ships with it
1 file beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 12d ago First seen · 210 lines · 40 tokens per session scan A 480b36b47b1b
ai-fault-mode-deflector is a skill published in the GitHub repository gpsnmeajp/ai-character-checker (5 stars, last pushed 5mo ago), licensed CC0-1.0. It adds 40 tokens to every session and 3,262 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other skills, from other repositories
llm-app-patterns
Production-ready patterns for building LLM applications. Covers RAG pipelines, agent architectures, prompt IDEs, and LLMOps monitoring. Use when designing AI applications, implementing RAG, building agents, or setting up LLM observability.
prompt-optimization
Improve a prompt on the evaluations workbench through a measured loop. Score the baseline first, then duplicate the target column, form a hypothesis from failing rows, edit the copy's prompt draft, run, compare pass rate and cost, and repeat until the numbers hold. Use when the user asks to optimize or improve a…
enhance-prompt
Transforms vague UI ideas into polished, Stitch-optimized prompts. Enhances specificity, adds UI/UX keywords, injects design system context, and structures output for better generation results.
prompt-engineer
Writes, refactors, and evaluates prompts for LLMs — generating optimized prompt templates, structured output schemas, evaluation rubrics, and test suites. Use when designing prompts for new LLM applications, refactoring existing prompts for better accuracy or token efficiency, implementing chain-of-thought or few-shot…
seedance-vocab-en
This skill should be used when an English Seedance 2.0 prompt needs clearer production wording, less generic prose, or precise vocabulary for camera, lighting, motion, VFX, audio, and constraints. Route blocked prompts through seedance-filter for context and boundary review.
ideogram4
Prompting patterns for Ideogram 4 text-to-image — best-in-class in-image text rendering and exact color/layout control via structured JSON captions. Use when generating images that need legible on-image text (title cards, thumbnails, logos, signage, CTAs), precise brand colors, or controlled spatial layout. Triggers…