Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add MythicAgents/sage --skill sage-focused-capability-testsgit clone --depth 1 https://github.com/MythicAgents/sageWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/mythicagents/sage/sage-focused-capability-tests)<a href="https://agentmods.dev/skills/mythicagents/sage/sage-focused-capability-tests"><img src="https://agentmods.dev/badge/skills/mythicagents/sage/sage-focused-capability-tests/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/mythicagents/sage/sage-focused-capability-tests"><img src="https://agentmods.dev/badge/skills/mythicagents/sage/sage-focused-capability-tests.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00083 | $0.00825 |
| Opus 5 | $0.00042 | $0.00413 |
| Sonnet 5 | $0.00017 | $0.00165 |
| Haiku 4.5 | $0.00008 | $0.00082 |
Grade A, and why
sage-focused-capability-tests scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 10d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
The source is not reproduced here
Licensed GPL-3.0
The repository is licensed GPL-3.0, which this catalogue does not treat as permission to reproduce the file. Read it at the source.
What ships with it
17 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.
- agents/openai.yaml 228 B
- scripts/build_capability_smoke.py 3.5 KB runs code
- scripts/probe_slack_findings_webhook.py 4.0 KB runs code
- scripts/run_focused_account_context.py 8.9 KB runs code
- scripts/run_focused_adcs_ca_export.py 15 KB runs code
- scripts/run_focused_adcs_certificate_auth.py 5.9 KB runs code
- scripts/run_focused_dcsync_account.py 7.6 KB runs code
- scripts/run_focused_endpoint_protection.py 12 KB runs code
- scripts/run_focused_local_admin_access.py 9.9 KB runs code
- scripts/run_focused_local_admin_remote_exec.py 14 KB runs code
- scripts/run_focused_managed_secret_read.py 9.2 KB runs code
- scripts/run_focused_parameter_group_reference.py 7.8 KB runs code
- scripts/run_focused_parent_dcsync.py 5.9 KB runs code
- scripts/run_focused_sid_history.py 6.6 KB runs code
- scripts/run_focused_ticket_context_proof.py 8.6 KB runs code
- scripts/run_offline_suite.py 3.9 KB runs code
- tests/test_probe_slack_findings_webhook.py 4.5 KB runs code
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 10d ago First seen · 66 lines · 83 tokens per session scan A b1481dbce4eb
sage-focused-capability-tests is a skill published in the GitHub repository MythicAgents/sage (24 stars, last pushed 14d ago), licensed GPL-3.0. It adds 83 tokens to every session and 825 once invoked, about $0.0004 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other skills, from other repositories
systematic-debugging
4-phase root cause debugging: understand bugs before fixing.
mem0-test-integration
Verify a Mem0 integration produced by /mem0-integrate. Runs in the same workspace on the same branch (loose coupling) — installs dependencies, runs the repo's native test suite, then exercises a real end-to-end smoke flow against the user's API key. Produces a scorecard. TRIGGER when: user has just run /mem0-integrate…
darwinian-evolver
Evolve prompts/regex/SQL/code with Imbue's evolution loop.
evaluating-with-leakage-gates
Evaluate an OpenMed de-identification or clinical NER model against the leakage-first release gates G1a through G8, which gate releases on residual PHI leakage rather than on F1. Use when the user wants to run the OpenMed eval harness on a synthetic golden set, decide whether a de-id model is RELEASABLE or…
bat-story-eval
Compare MCP tool behavior between target and baseline versions using pre-built and custom stories with diff-based triage.
connect-agent
Connect the codebase's AI agent to LangWatch agent simulations, so test suites run against the real agent process. Adds a small connect function beside the service startup that calls the agent already in the codebase, which opens an outbound connection and registers the agent with its environment and its run…