cli-audit-test

A review framework for judging the quality and maturity of a test plan or test suite.

In plain words
What is it for?
Use it to assess test plans, test directories, test strategies, and test-suite structure against established testing practices.
Why use it?
It exposes missing coverage, weak negative testing, unbalanced test levels, and gaps in automation or continuous integration before they cause delivery risk.

Skill for Claude CodeCodex

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/destynova2/cli-code-skills/cli-audit-test
Any agent
npx skills add Destynova2/cli-code-skills --skill cli-audit-test
Clone the repo
git clone --depth 1 https://github.com/Destynova2/cli-code-skills

Made for: Claude Code, Codex.

Per session 94 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 2,624 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin 88% copy Near-identical to another mod in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00094 $0.02624
Opus 5 $0.00047 $0.01312
Sonnet 5 $0.00019 $0.00525
Haiku 4.5 $0.00009 $0.00262

Measured 2d ago against content hash ab52aec1a43d, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

cli-audit-test scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 2d ago.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

Origin

This is a copy

88% identical to api-audit — 693 lines differ, which has more behind it and is treated as the original. This page carries a canonical link to it rather than competing with it.

cli-audit-test/SKILL.md · 203 lines

How it starts

The opening of the file, as written. The whole thing — 203 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Optimization: This skill uses on-demand loading. Heavy content lives in references/ and is loaded only when needed.

Language rule: Skill instructions are written in English. When generating user-facing output, detect the project's primary language (from README, comments, docs, commit messages) and produce the report in that language. If the project is bilingual, ask the user which language to use before proceeding.

Audit Test — Test Plan Quality & Maturity Scorer

Evaluate a test plan or test suite against ISTQB standards, TMMi maturity model, and industry best practices.

"A test plan that only says 'we will test' without specifying which techniques and why is a major red flag." — ISTQB Foundation Level Syllabus v4.0

Core Principle

A test plan is a contract with quality. Every dimension left vague is a risk accepted silently. This skill measures how explicitly and completely that contract is defined, using a 12-dimension framework calibrated on ISTQB, TMMi, and proven industry patterns.

Input

$ARGUMENTS is the target to audit:

  • Test plan document (.md, .txt, .adoc, .pdf): audit the plan as written
  • Tests directory (tests/, tests/e2e/, etc.): infer the plan from the test structure, files, and config
  • Empty: search for test plans in the project root, then fall back to tests/ directory analysis

Input Discovery

  1. Glob for test plan documents: **/test-plan*, **/test-strategy*, **/TESTING*
  2. Glob for test directories: **/tests/**, **/test/**, **/e2e/**, **/spec/**
  3. Glob for CI config: .github/workflows/*, .gitlab-ci.yml, Jenkinsfile, justfile
  4. Glob for test config: **/pict/**, **/fixtures/**, **/mocks/**, **/toxiproxy*
  5. Read project manifest (Cargo.toml, package.json, etc.) for test dependencies

13-Dimension Framework

Score each dimension 0-4, then compute a weighted final score. Read references/dimensions.md for detailed scoring criteria, evidence patterns, and per-dimension guidance.

Read the full file on GitHub · 203 lines

Files

What ships with it

4 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. 2d ago First seen · 203 lines · 94 tokens per session scan A ab52aec1a43d

Subscribe to this mod's changes

cli-audit-test is a skill published in the GitHub repository Destynova2/cli-code-skills (5 stars, last pushed 10d ago), licensed MIT. It adds 94 tokens to every session and 2,624 once invoked, about $0.0005 per session on Opus 5. A static security scan graded it A with 0 findings. It is 88% identical to api-audit, differing in 693 lines, and is treated as a copy.