skill-test

skill-test is a skill for Claude Code, Codex from striderZA/OpenCodeGameStudios. It costs 31 tokens per session (3,266 once invoked), scanned A, a copy of skill-test, MIT.

A validator for coding-agent skill files that checks their structure, behavior, category fit, and overall test coverage. It offers static checks, behavioral checks, category-specific checks, and an audit report.

In plain words
What is it for?
It helps lint one or all skills, verify a skill against its test specification, assess category requirements, and report which skills or agent specifications need testing.
Why use it?
A skill can look correctly written but still fail to follow required rules or behave incorrectly. These checks expose structural problems, failed expectations, and missing test coverage.

Skill for Claude CodeCodex

Install

Getting it into your agent

One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.

agentmods
npx agentmods add skills/striderza/opencodegamestudios/skill-test
Any agent
npx skills add striderZA/OpenCodeGameStudios --skill skill-test
Clone the repo
git clone --depth 1 https://github.com/striderZA/OpenCodeGameStudios

Made for: Claude Code, Codex.

Wrote this? Show the measurements

A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.

agentmods badge for skill-test

README.md
[![agentmods](https://agentmods.dev/badge/skills/striderza/opencodegamestudios/skill-test.svg)](https://agentmods.dev/skills/striderza/opencodegamestudios/skill-test)
Your own site
<a href="https://agentmods.dev/skills/striderza/opencodegamestudios/skill-test"><img src="https://agentmods.dev/badge/skills/striderza/opencodegamestudios/skill-test.svg" alt="Measured on agentmods" height="20"></a>
Per session 31 Skills are progressive disclosure: only the name and description are preloaded; the body loads when the skill is used.
When invoked 3,266 The whole file, excluding the scripts and references it only reads on demand.
Security scan A 0 findings. Scan, not verified.
Origin 92% copy Near-identical to another mod in the catalogue.
Token cost

What it costs to keep this loaded

Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.

ModelPer sessionOnce invoked
Fable 5 $0.00031 $0.03266
Opus 5 $0.00015 $0.01633
Sonnet 5 $0.00006 $0.00653
Haiku 4.5 $0.00003 $0.00327

Measured yesterday against content hash ef9cf919b22d, method: parsed. Prices are Anthropic first-party input rates as of 2026-08-30, from the pricing page.

Security

Grade A, and why

skill-test scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured yesterday.

A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.

Nothing flagged

None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.

Origin

This is a copy

92% identical to skill-test — 13 lines differ, which has more behind it and is treated as the original. This page carries a canonical link to it rather than competing with it.

.agents/modules/qa/skills/skill-test/SKILL.md · 357 lines

How it starts

The opening of the file, as written. The whole thing — 357 lines — stays where its author put it; the contents beside it link to each section on GitHub.

Skill Test

Validates .agents/skills/*/SKILL.md files for structural compliance and behavioral correctness. No external dependencies — runs entirely within the existing skill/hook/template architecture.

Four modes:

Mode Command Purpose Token Cost
static /skill-test static [name|all] Structural linter — 7 compliance checks per skill Low (~1k/skill)
spec /skill-test spec [name] Behavioral verifier — evaluates assertions in test spec Medium (~5k/skill)
category /skill-test category [name|all] Category rubric — checks skill against its category-specific metrics Low (~2k/skill)
audit /skill-test audit Coverage report — skills, agent specs, last test dates Low (~3k total)

Phase 1: Parse Arguments

Determine mode from the first argument:

  • static [name] → run 7 structural checks on one skill
  • static all → run 7 structural checks on all skills (Glob .agents/skills/*/SKILL.md)
  • spec [name] → read skill + test spec, evaluate assertions
  • category [name] → run category-specific rubric from CCGS Skill Testing Framework/quality-rubric.md
  • category all → run category rubric for every skill that has a category: in catalog
  • audit (or no argument) → read catalog, list all skills and agents, show coverage

If argument is missing or unrecognized, output usage and stop.


Phase 2A: Static Mode — Structural Linter

For each skill being tested, read its SKILL.md fully and run all 7 checks:

Check 1 — Required Frontmatter Fields

The file must contain all of these in the YAML frontmatter block:

  • name:
  • description:
  • argument-hint:
  • user-invocable:
  • allowed-tools:

FAIL if any are absent.

Check 2 — Multiple Phases

The skill must have ≥2 numbered phase headings. Look for patterns like:

  • ## Phase N or ## Phase N:
  • ## N. (numbered top-level sections)
  • At least 2 distinct ## headings if phases aren't explicitly numbered

Read the full file on GitHub · 357 lines

Changes

What this file has done since we first saw it

Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.

  1. yesterday First seen · 357 lines · 31 tokens per session scan A ef9cf919b22d

Subscribe to this mod's changes

skill-test is a skill published in the GitHub repository striderZA/OpenCodeGameStudios (84 stars, last pushed 24d ago), licensed MIT. It adds 31 tokens to every session and 3,266 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it A with 0 findings. It is 92% identical to skill-test, differing in 13 lines, and is treated as a copy.