Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add dhaupin/vant --skill vant-skill-test-chaosgit clone --depth 1 https://github.com/dhaupin/vantWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/dhaupin/vant/vant-skill-test-chaos)<a href="https://agentmods.dev/skills/dhaupin/vant/vant-skill-test-chaos"><img src="https://agentmods.dev/badge/skills/dhaupin/vant/vant-skill-test-chaos/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/dhaupin/vant/vant-skill-test-chaos"><img src="https://agentmods.dev/badge/skills/dhaupin/vant/vant-skill-test-chaos.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00033 | $0.00311 |
| Opus 5 | $0.00016 | $0.00156 |
| Sonnet 5 | $0.00007 | $0.00062 |
| Haiku 4.5 | $0.00003 | $0.00031 |
Grade A, and why
test-chaos scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 9d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
What it actually says
Test Chaos
Resilience testing.
When To Use
- Fault tolerance
- Disaster recovery
- System hardening
Chaos Experiments
1. Kill Process
kill -9 <pid>
2. Cut Network
iptables -A INPUT -j DROP
3. Fill Disk
dd if=/dev/zero of=/full
4. Latency
tc qdisc add dev eth0 root netem delay 5000ms
Tools
| Tool | Use |
|---|---|
| Chaos Monkey | Netflix chaos |
| Litmus | Kubernetes chaos |
| Pumba | Docker chaos |
| Gremlin | Chaos as service |
Output
## Chaos Test
| Experiment | Result | Recovery |
|------------|--------|---------|
| Kill DB | [FAIL→RECOVERED] | 5s |
| Net cut | [FAIL→RECOVERED] | 2s |
| Disk full | [HANDLED] | N/A |
Role: Chaos Engineer
Input: Experiment
Output: Resilience
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 9d ago First seen · 73 lines · 33 tokens per session scan A bd801d35d36a
test-chaos is a skill published in the GitHub repository dhaupin/vant (9 stars, last pushed 10d ago), licensed MIT. It adds 33 tokens to every session and 311 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other skills, from other repositories
rn-testing
This skill should be used when the user asks to "write a Maestro test", "create E2E flows", "add testIDs", "run UI tests", "run E2E tests", "verify a feature works", "test my screen", "set up maestro-runner", "mock network requests", "inspect store state", "write test assertions", or needs guidance on test timing…
capturing-proof
This skill should be used when the user asks to "capture proof", "record a demo of this feature", "make a video showing it works", "record the flow for the PR", "generate a PR body", "capture screenshots for the PR", "proof-capture", or when a verified feature needs PR-ready proof artifacts (video + numbered…
run-action
Explicit Codex workflow: Execute a learned Maestro flow ("action") by name with optional -e KEY=VALUE parameters. Looks the flow up via packages/rn-dev-agent-core/dist/learned-actions.js (same inventory as $rn-dev-agent:list-learned-actions), then replays it via cdprunaction — auto-repair-aware orchestration with…
build-and-test
Explicit Codex workflow: Build the Expo/React Native app (local or EAS), install it, start Metro, then test a requested feature end-to-end.
check-env
Explicit Codex workflow: Check that the fenced React Native session, Metro, app target, and device inventory are ready for testing.
proof-capture
Explicit Codex workflow: Capture PR-ready proof artifacts for a feature, with an attested fail-closed controller in strict mode.