Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add cdeust/zetetic-team-subagents --skill measurement-disciplinegit clone --depth 1 https://github.com/cdeust/zetetic-team-subagentsWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/cdeust/zetetic-team-subagents/measurement-discipline)<a href="https://agentmods.dev/skills/cdeust/zetetic-team-subagents/measurement-discipline"><img src="https://agentmods.dev/badge/skills/cdeust/zetetic-team-subagents/measurement-discipline/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/cdeust/zetetic-team-subagents/measurement-discipline"><img src="https://agentmods.dev/badge/skills/cdeust/zetetic-team-subagents/measurement-discipline.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00071 | $0.00708 |
| Opus 5 | $0.00036 | $0.00354 |
| Sonnet 5 | $0.00014 | $0.00142 |
| Haiku 4.5 | $0.00007 | $0.00071 |
Grade A, and why
measurement-discipline scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 12d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 49 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Measurement Discipline
Problem shape: a quantity is being read, improved, or argued about, but the instrument, the unit, or the conservation ledger behind it has never been audited. Symptoms: residuals outside noise, metrics without operational definitions, inputs and outputs that don't balance, observer effects, one-method-only results.
Relevant geniuses
| Agent | Use when |
|---|---|
| curie | measured > predicted from known parts; instrument missing or unit undefined; measurement may perturb the system (Heisenbugs, observability overhead) |
| shannon | "improving X" where X has no formal definition; method proposed without knowing the theoretical limit; a metric with no repeatable procedure |
| lavoisier | money, data, requests, or time "disappearing"; inputs and outputs never balanced; the residual needs a name and a carrier |
| galileo | phenomenon obscured by secondary effects; too fast/large/rare to observe directly; qualitative claims that need a number |
| einstein | a concept in the metric has no measurement procedure; the rule gives different answers from different viewpoints |
| deming | reacting to noise as if it were signal — common vs special cause not separated |
| ekman | a "subjective" domain needs objective coding; signal hides below normal temporal resolution; per-subject baselines missing |
| wu | a "law" or assumption everyone trusts has never actually been tested; increased precision could refute it |
Invocation
- Pick the best-fit agent above. If two or more fit, run
tools/genius-invoker.sh route "<problem>"and take the top ranked match. - Load it:
tools/genius-invoker.sh invoke <agent> "<problem>", then readagents/genius/<agent>.mdin full. - Apply the agent's
<workflow>step by step — do not skip steps — and answer in its<output-format>. - Chain when the problem spans shapes (e.g. lavoisier finds the residual,
curie isolates its carrier):
tools/genius-invoker.sh compose lavoisier curie -- "<problem>". - If no shape above matches, do not force a genius — use a standard team agent (INDEX.md routing rule).
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 12d ago First seen · 49 lines · 71 tokens per session scan A 8e28a53f2c6a
measurement-discipline is a skill published in the GitHub repository cdeust/zetetic-team-subagents (7 stars, last pushed 2d ago), licensed MIT. It adds 71 tokens to every session and 708 once invoked, about $0.0004 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other skills, from other repositories
create-site
Creates a new Power Pages code site (SPA) using React, Angular, Vue, or Astro. Guides through the full process from initial concept to deployed site: requirements discovery, scaffolding, component planning, design, implementation, validation, and deployment. Use when the user wants to create, build, or scaffold a new…
review
5-pass structured code review — correctness, security, performance, readability, consistency.
alive:system-upgrade
Upgrade ALIVE to the current version. Handles v1/v2/v3.x source states, multi-surface aware (alive-mcp / Hermes / Codex), retroactive version detection, partial-failure resume, dry-run previews, and rollback inspection.
extract-resume
Parse a resume's uploaded PDF into structured JSON (basics, experience, projects, skills, education) and save it to the editor.
ensure-pipelines-host
Ensures the tenant has a usable Power Platform Pipelines host environment before any pipeline operation runs. Detects host state via the same resolution order as the Power Apps UI (org-db setting → BAP env metadata → default-custom-host setting); if any existing host (Platform or Custom) is found, uses it. If no host…
codex-test-gen
Generate unit tests for specified functions using Codex exec.