test

A testing assistant that writes tests for selected files, folders, or current repository changes. Code coverage means how much of the code is exercised by tests.

In plain words
What is it for?
It is for adding thorough tests to existing code changes or areas with coverage gaps.
Why use it?
It helps find untested business logic, state changes, and error paths without spending time identifying every missing test by hand.

Skill for Claude CodeCodex

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/mystilleef/spae-framework/test
Any agent
npx skills add mystilleef/spae-framework --skill test
Clone the repo
git clone --depth 1 https://github.com/mystilleef/spae-framework

Made for: Claude Code, Codex.

Per session 14 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 1,321 The whole file, excluding the scripts and references it only reads on demand.
Security scan C 1 finding. Scan, not verified.
Origin original No closer match found in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00014 $0.01321
Opus 5 $0.00007 $0.00660
Sonnet 5 $0.00003 $0.00264
Haiku 4.5 $0.00001 $0.00132

Measured 2d ago against content hash a1103a50af65, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade C, and why

test scanned grade C with 1 finding against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 2d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Hidden instructionshighPrompt injection

Directives inside HTML comments, invisible characters or bidirectional overrides are read by the model and not by the person reviewing the file.

<!-- prettier-ignore-start -->
skills/test/SKILL.md · 161 lines

How it starts

The opening of the file, as written. The whole thing — 161 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Test

When to use

  • Files, or changed code, have gaps in coverage

Goal

  • Write comprehensive and exhaustive tests that address gaps in code coverage.

Input

Determine scope from the first available source:

  • Files or folders provided by the user.
  • Current changes in the repository.

Abort if no scope exists.

Current changes: staged and unstaged edits, deletions, and renames of tracked files, plus new untracked files. Requires a versioned project. Abort with a clear message if none detected.

Testing

See references/testing-guide.md for test structure, isolation, mocking, assertion, and performance standards.

Shell commands

See references/shell-command-guide.md for command safety, timeouts, redirects, and non-interactive environment directives.

Cleanup

See references/cleanup-guide.md for the self-introduced-artifact checklist and diff-only audit scope.

Behavioral surface

Target only methods and functions with business logic, state transitions, or error handling. Exclude: trivial getters/setters, POJOs, generated code, framework boilerplate, and any method where every path delegates trivially, accesses a field, or returns a computed value with no state change, resource interaction, or error-propagation decision.

Workflow

  1. GATE—Confirm scope: user-provided files or current repository changes. Abort immediately if none.
  2. ORIENT—Goal: cover all behavioral gaps in scope. Production code unchanged.
  3. PLAN—Read references/testing-guide.md, references/shell-command-guide.md, and references/cleanup-guide.md. Inspect production code, adjacent tests, and coverage commands. List every gap across all four categories per behavioral-surface target.
    • Short-circuit: zero gaps found → emit Result: No Gaps and halt.
  4. ACT—Write tests for every enumerated gap. Cover all four categories per target before declaring it complete.
  5. VERIFY—Loop over every criterion declared in PLAN:
    • Run targeted tests; run broader suite or coverage tool.
    • Audit every new test file against CI Parity rules in references/testing-guide.md; fix any violation before proceeding.
    • Audit the task's own git diff/git status against the Self-cleanup checklist in references/cleanup-guide.md; remove every self-introduced artifact before proceeding.
    • For each unmet criterion: return to ACT, execute, then re-enter VERIFY.
    • Exit only when all pass and no regressions remain.
    • Halt only for out-of-scope blockers.
  6. PERSIST—Confirm all test files written; no partial writes.
  7. REPORT—Emit the result following the result directives and using the result template.

Read the full file on GitHub · 161 lines

Files

What ships with it

3 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 2d ago First seen · 161 lines · 14 tokens per session scan C a1103a50af65

Subscribe to this mod's changes

test is a skill published in the GitHub repository mystilleef/spae-framework (1 stars, last pushed 29d ago), licensed MIT. It adds 14 tokens to every session and 1,321 once invoked, about $0.0001 per session on Opus 5. A static security scan graded it C with 1 finding (hidden instructions). No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.

Related

Other skills, from other repositories

exploit-xss

Cross-site scripting (XSS) vulnerability detection and exploitation. Supports reflected XSS, stored XSS, DOM-based XSS, and blind XSS testing. Use this skill when user mentions XSS, cross-site scripting, script injection, or needs to test JavaScript injection in parameters, forms, headers, or DOM sources.

crazyMarky/pentest-skills · 71 tokens

results-storage

SQLite-based persistent storage and reporting system for penetration testing results. Use this skill when user needs to store scan results, query vulnerabilities, generate reports, or manage pentest data across sessions.

crazyMarky/pentest-skills · 40 tokens

exploit-sqli

SQL injection detection and exploitation using sqlmap, manual techniques, and custom payloads. Use this skill when user needs to test for SQL injection vulnerabilities, extract database information, or exploit SQLi in parameters, headers, or cookies.

crazyMarky/pentest-skills · 51 tokens

recon-dir-scan

Directory and file enumeration using ffuf, gobuster, dirsearch, and feroxbuster. Use this skill when user needs to discover hidden directories, enumerate files, find backup files, or map application structure through path fuzzing.

crazyMarky/pentest-skills · 52 tokens

recon-fingerprint

Web fingerprinting and WAF detection using wafw00f, whatweb, nuclei, and httpx. Use this skill when user needs to identify web technologies, detect WAF/CDN, analyze server headers, or fingerprint web applications and frameworks.

crazyMarky/pentest-skills · 55 tokens

recon-subdomain

Subdomain enumeration and DNS reconnaissance using subfinder, amass, dnsx, and other tools. Use this skill when user needs to discover subdomains, perform DNS enumeration, gather DNS records, or find hidden subdomains of a target domain.

crazyMarky/pentest-skills · 54 tokens