Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add skills/matrixfounder/agentic-development/artifact-formalizernpx skills add MatrixFounder/Agentic-development --skill artifact-formalizergit clone --depth 1 https://github.com/MatrixFounder/Agentic-developmentWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/matrixfounder/agentic-development/artifact-formalizer)<a href="https://agentmods.dev/skills/matrixfounder/agentic-development/artifact-formalizer"><img src="https://agentmods.dev/badge/skills/matrixfounder/agentic-development/artifact-formalizer.svg" alt="Measured on agentmods" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.00126 | $0.04429 |
| Opus 5 | $0.00063 | $0.02214 |
| Sonnet 5 | $0.00025 | $0.00886 |
| Haiku 4.5 | $0.00013 | $0.00443 |
Grade A, and why
artifact-formalizer scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured today.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 286 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Artifact formalizer (specification register)
Purpose
A reader of an essay-register artifact performs two passes: one to locate the requirement, one to decide whether a given sentence carried one. This skill removes the second pass. It has two modes.
| Mode | When | Instrument | Owns |
|---|---|---|---|
| A — Authoring | before and during writing | references/authoring-contract.md |
the defect is not written |
| B — Audit | on an existing document | scripts/scan_register.py + a reading pass |
what Mode A missed |
Mode A prevents the defect; Mode B measures what Mode A missed.
Why in that order. Defective prose measured 5.1% of one corpus's words (731 of 14,288), while
removing it required reading and editing all 14,288. Register also varied by authoring model on the
same repository, so an unwritten standard is not a standard.
references/measurement-baseline.md §5 carries both figures.
Scope boundary — what this is NOT
This is not a jargon or terminology tool. Domain terms are precise and stay verbatim:
singleflight, RTM, дедлайн, throttle. Translating engineering vocabulary into business
language for customer-facing documents is a different problem with a different audience, and this
skill does not attempt it.
1. Red Flags (Anti-Rationalization)
A definition list, not a table: the reality column is prose, and documentation-standards §5.1
prescribes converting such a column rather than widening it.
- "I'll write it, then run the formalizer." That is the expensive order. §Purpose gives the measured cost. Mode A first.
- "The scan is clean, so the text is clean." Every zero is reported next to what the detector
actually saw. Read
DIAGNOSTICS: a corpus whose longest sentence equals the limit was written for the gate. - "I'll add a pattern for this new phrasing." First ask whether a test in the authoring contract already forbade it. If it did, the lexicon gains a faster detector and nothing else changes. If it did not, the contract is amended (§6).
- "«Шов», «нога», «мост» — это термины проекта." A term appears in ARCHITECTURE.md, a public
API, or a cited standard. Run
--terms docs/ARCHITECTURE.mdand let the scanner apply that test. - "I formalized the section that was quoted at me." The pass that produced this skill's own
worked example did exactly that. It left seventeen occurrences of one metaphor in the same task
set. Coverage is per section:
--sections. - "The whole document reads badly — I'll rewrite it." Conforming sentences stay verbatim. Over-rewriting introduces errors that were not there.
- "I'll shorten it by trimming the requirement." Register changes, substance does not. Numbers, identifiers, obligations and scope limits survive verbatim.
- "This word is fine — I'll widen the threshold." A failing scan is fixed in the prose. Thresholds move only when a measurement moves them.
- "It flagged a false positive, so the rule is wrong." The scanner is advisory by design. Judge
the finding;
infoexists because the author decides.
What ships with it
60 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.
- data/register-en.json 19 KB
- data/register-ru.json 22 KB
- evals/corpus-wi12/A1/with_contract/rep-1.md 19 KB
- evals/corpus-wi12/A1/with_contract/rep-1.meta.json 484 B
- evals/corpus-wi12/A1/with_contract/rep-2.md 17 KB
- evals/corpus-wi12/A1/with_contract/rep-2.meta.json 484 B
- evals/corpus-wi12/A1/with_contract/rep-3.md 24 KB
- evals/corpus-wi12/A1/with_contract/rep-3.meta.json 484 B
- evals/corpus-wi12/A5/with_contract/rep-1.md 19 KB
- evals/corpus-wi12/A5/with_contract/rep-1.meta.json 494 B
- evals/corpus-wi12/A5/with_contract/rep-2.md 18 KB
- evals/corpus-wi12/A5/with_contract/rep-2.meta.json 484 B
- evals/corpus-wi12/A5/with_contract/rep-3.md 23 KB
- evals/corpus-wi12/A5/with_contract/rep-3.meta.json 493 B
- evals/corpus/A1/baseline/rep-1.md 43 KB
- evals/corpus/A1/baseline/rep-1.meta.json 479 B
- evals/corpus/A1/baseline/rep-2.md 30 KB
- evals/corpus/A1/baseline/rep-2.meta.json 479 B
- evals/corpus/A1/baseline/rep-3.md 32 KB
- evals/corpus/A1/baseline/rep-3.meta.json 479 B
- evals/corpus/A1/with_contract/rep-1.md 22 KB
- evals/corpus/A1/with_contract/rep-1.meta.json 494 B
- evals/corpus/A1/with_contract/rep-2.md 16 KB
- evals/corpus/A1/with_contract/rep-2.meta.json 495 B
- evals/corpus/A1/with_contract/rep-3.md 17 KB
- evals/corpus/A1/with_contract/rep-3.meta.json 484 B
- evals/corpus/A2/baseline/rep-1.md 59 KB
- evals/corpus/A2/baseline/rep-1.meta.json 489 B
- evals/corpus/A2/with_contract/rep-1.md 34 KB
- evals/corpus/A2/with_contract/rep-1.meta.json 484 B
- evals/corpus/A3/baseline/rep-1.md 12 KB
- evals/corpus/A3/baseline/rep-1.meta.json 477 B
- evals/corpus/A3/with_contract/rep-1.md 13 KB
- evals/corpus/A3/with_contract/rep-1.meta.json 483 B
- evals/corpus/A4/baseline/rep-1.md 12 KB
- evals/corpus/A4/baseline/rep-1.meta.json 479 B
- evals/corpus/A4/with_contract/rep-1.md 5.1 KB
- evals/corpus/A4/with_contract/rep-1.meta.json 483 B
- evals/corpus/A5/baseline/rep-1.md 24 KB
- evals/corpus/A5/baseline/rep-1.meta.json 480 B
- evals/corpus/A5/baseline/rep-2.md 25 KB
- evals/corpus/A5/baseline/rep-2.meta.json 491 B
- evals/corpus/A5/baseline/rep-3.md 21 KB
- evals/corpus/A5/baseline/rep-3.meta.json 480 B
- evals/corpus/A5/with_contract/rep-1.md 19 KB
- evals/corpus/A5/with_contract/rep-1.meta.json 484 B
- evals/corpus/A5/with_contract/rep-2.md 18 KB
- evals/corpus/A5/with_contract/rep-2.meta.json 495 B
- evals/corpus/A5/with_contract/rep-3.md 17 KB
- evals/corpus/A5/with_contract/rep-3.meta.json 495 B
- evals/corpus/A6/baseline/rep-1.md 12 KB
- evals/corpus/A6/baseline/rep-1.meta.json 479 B
- evals/corpus/A6/with_contract/rep-1.md 15 KB
- evals/corpus/A6/with_contract/rep-1.meta.json 494 B
- evals/corpus/B1/rep-1/answer.json 483 B
- evals/corpus/B1/rep-1/meta.json 338 B
- evals/corpus/B2/rep-1/answer.json 242 B
- evals/corpus/B2/rep-1/meta.json 338 B
- evals/corpus/B3/rep-1/answer.json 417 B
- evals/corpus/B3/rep-1/meta.json 327 B
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- today Changed · +4 tokens per session 721456843f49
- 4d ago First seen · 286 lines · 122 tokens per session scan A bc0e788b4446
artifact-formalizer is a skill published in the GitHub repository MatrixFounder/Agentic-development (5 stars, last pushed yesterday), licensed Apache-2.0. It adds 126 tokens to every session and 4,429 once invoked, about $0.0006 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other skills, from other repositories
Learn what defines effective BDD scenarios
Interactive guidance on writing complete, effective BDD scenarios for story-flow.
add-tests
Generates tests for existing code. Analyzes the target function, method, or class to identify the happy path, error cases, and edge cases, then writes test cases following the project's testing framework and naming conventions. Invoked when the user asks to add tests, write tests, cover code, or increase coverage.
tdd-implementation
Use when implementing any code change for a task – new behavior, a bugfix, or a revision after review – before writing the production code. Defines the failing-test-first cycle, public-seam testing, vertical slices, scaffold and infra handling, and counters to common excuses for tests-after-code.
test-writer
Writes tests that fail before a fix and pass after it.
test-generation
Generate and run unit/integration tests TDD-style across Python, Node, and .NET. Use when adding tests to untested code, implementing a feature test-first, or finding coverage gaps.
argent-tv-interact
Control and inspect TV apps via argent — Apple TV (tvOS), Android TV (leanback), and Amazon Fire TV (Vega). Boot the target, read focus, navigate with the D-pad remote, type, screenshot, and on Vega debug the JS runtime (evaluate, console logs, network inspector). Use when a task targets a TV (runtimeKind "tv", or…