Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add skills/mystilleef/spae-framework/testnpx skills add mystilleef/spae-framework --skill testgit clone --depth 1 https://github.com/mystilleef/spae-frameworkWhat it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.00014 | $0.01321 |
| Opus 5 | $0.00007 | $0.00660 |
| Sonnet 5 | $0.00003 | $0.00264 |
| Haiku 4.5 | $0.00001 | $0.00132 |
Grade C, and why
test scanned grade C with 1 finding against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 2d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Hidden instructionshighPrompt injection
Directives inside HTML comments, invisible characters or bidirectional overrides are read by the model and not by the person reviewing the file.
<!-- prettier-ignore-start --> How it starts
The opening of the file, as written. The whole thing — 161 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Test
When to use
- Files, or changed code, have gaps in coverage
Goal
- Write comprehensive and exhaustive tests that address gaps in code coverage.
Input
Determine scope from the first available source:
- Files or folders provided by the user.
- Current changes in the repository.
Abort if no scope exists.
Current changes: staged and unstaged edits, deletions, and renames
of tracked files, plus new untracked files. Requires a versioned
project. Abort with a clear message if none detected.
Testing
See references/testing-guide.md for test structure, isolation,
mocking, assertion, and performance standards.
Shell commands
See references/shell-command-guide.md for command safety, timeouts,
redirects, and non-interactive environment directives.
Cleanup
See references/cleanup-guide.md for the self-introduced-artifact
checklist and diff-only audit scope.
Behavioral surface
Target only methods and functions with business logic, state
transitions, or error handling. Exclude: trivial getters/setters,
POJOs, generated code, framework boilerplate, and any method where
every path delegates trivially, accesses a field, or returns a computed
value with no state change, resource interaction, or error-propagation
decision.
Workflow
- GATE—Confirm scope: user-provided files or current repository changes. Abort immediately if none.
- ORIENT—Goal: cover all behavioral gaps in scope. Production code unchanged.
- PLAN—Read
references/testing-guide.md,references/shell-command-guide.md, andreferences/cleanup-guide.md. Inspect production code, adjacent tests, and coverage commands. List every gap across all four categories per behavioral-surface target.- Short-circuit: zero gaps found → emit
Result: No Gapsand halt.
- Short-circuit: zero gaps found → emit
- ACT—Write tests for every enumerated gap. Cover all four categories per target before declaring it complete.
- VERIFY—Loop over every criterion declared in PLAN:
- Run targeted tests; run broader suite or coverage tool.
- Audit every new test file against CI Parity rules in
references/testing-guide.md; fix any violation before proceeding. - Audit the task's own
git diff/git statusagainst the Self-cleanup checklist inreferences/cleanup-guide.md; remove every self-introduced artifact before proceeding. - For each unmet criterion: return to
ACT, execute, then re-enterVERIFY. - Exit only when all pass and no regressions remain.
- Halt only for out-of-scope blockers.
- PERSIST—Confirm all test files written; no partial writes.
- REPORT—Emit the result following the result directives and using the result template.
What ships with it
3 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 2d ago First seen · 161 lines · 14 tokens per session scan C a1103a50af65
test is a skill published in the GitHub repository mystilleef/spae-framework (1 stars, last pushed 29d ago), licensed MIT. It adds 14 tokens to every session and 1,321 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it C with 1 finding (hidden instructions). No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other skills, from other repositories
exploit-xss
Cross-site scripting (XSS) vulnerability detection and exploitation. Supports reflected XSS, stored XSS, DOM-based XSS, and blind XSS testing. Use this skill when user mentions XSS, cross-site scripting, script injection, or needs to test JavaScript injection in parameters, forms, headers, or DOM sources.
results-storage
SQLite-based persistent storage and reporting system for penetration testing results. Use this skill when user needs to store scan results, query vulnerabilities, generate reports, or manage pentest data across sessions.
exploit-sqli
SQL injection detection and exploitation using sqlmap, manual techniques, and custom payloads. Use this skill when user needs to test for SQL injection vulnerabilities, extract database information, or exploit SQLi in parameters, headers, or cookies.
recon-dir-scan
Directory and file enumeration using ffuf, gobuster, dirsearch, and feroxbuster. Use this skill when user needs to discover hidden directories, enumerate files, find backup files, or map application structure through path fuzzing.
recon-fingerprint
Web fingerprinting and WAF detection using wafw00f, whatweb, nuclei, and httpx. Use this skill when user needs to identify web technologies, detect WAF/CDN, analyze server headers, or fingerprint web applications and frameworks.
recon-subdomain
Subdomain enumeration and DNS reconnaissance using subfinder, amass, dnsx, and other tools. Use this skill when user needs to discover subdomains, perform DNS enumeration, gather DNS records, or find hidden subdomains of a target domain.