Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add Totes-MickGOATs/opus-pocus --skill the-pensievegit clone --depth 1 https://github.com/Totes-MickGOATs/opus-pocusWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/totes-mickgoats/opus-pocus/the-pensieve)<a href="https://agentmods.dev/skills/totes-mickgoats/opus-pocus/the-pensieve"><img src="https://agentmods.dev/badge/skills/totes-mickgoats/opus-pocus/the-pensieve/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/totes-mickgoats/opus-pocus/the-pensieve"><img src="https://agentmods.dev/badge/skills/totes-mickgoats/opus-pocus/the-pensieve.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00139 | $0.01192 |
| Opus 5 | $0.00069 | $0.00596 |
| Sonnet 5 | $0.00028 | $0.00238 |
| Haiku 4.5 | $0.00014 | $0.00119 |
Grade A, and why
the-pensieve scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 12d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 75 lines — stays where its author put it; the contents beside it link to each section on GitHub.
🔮 The Pensieve
"A week of numbers in the basin. Vibes are not telemetry."
Every other spell changes the instruction layer. This one finds out whether the changes worked. The alternative — another audit — just re-applies the same judgment that wrote the instructions; only behavior counts.
The counters
Instrument what the spells claim to improve. Core set:
| Counter | Needs | What it validates | Source |
|---|---|---|---|
| Stop/advisory re-fires per session — same advisory, unchanged state | hooks + transcripts | Finite Recursum's fingerprint fixes | Hook logs, or grep transcripts for the advisory headers |
| Review-loop rounds per task — review→fix→re-review depth | transcripts | Loop caps working | Transcripts / task logs |
| Subagent bounces per branch — how often incomplete work goes back | agents | Definition-of-done clarity | Dispatch logs, PR/branch history |
| Tokens injected per prompt by hooks | hooks | Muffliato Hookus | Instrument emitters to log output bytes; or measure injected blocks in transcripts |
| Wrong-agent routings — dispatched then redirected | agents + transcripts | Descriptio Reducio | Transcripts: dispatches followed by a different agent doing the work |
| Rule citations of stale facts — model quotes something no longer true | none | Obliviate Fossilium | Review of session outputs; user corrections |
| User corrections per session — times the human had to redirect | transcripts | Everything | Transcripts: user messages contradicting the previous assistant turn |
| Session outcome — task done / partial / abandoned | none | The bottom line | End-of-session state |
Add repo-specific counters for whatever the audit flagged worst — measure where the disease was.
Read the Needs column before promising anything. A repo with no hooks cannot measure advisory
re-fires or injected tokens; a repo with no subagents cannot measure bounces or wrong routings.
Five of these eight have a prerequisite most small repos do not meet.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 12d ago First seen · 75 lines · 139 tokens per session scan A efe61a756593
the-pensieve is a skill published in the GitHub repository Totes-MickGOATs/opus-pocus (10 stars, last pushed 1mo ago), licensed MIT. It adds 139 tokens to every session and 1,192 once invoked, about $0.0007 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other skills, from other repositories
claude-md-architect
Design, create, extract, or audit CLAUDE.md files and their supporting docs/ folder structure for Claude Code projects. Use this skill whenever the user asks to create a new CLAUDE.md, extract context from a long session into a CLAUDE.md system, audit/refactor an existing CLAUDE.md that has bloated, review whether…
mle-workflow
Production machine-learning engineering workflow for data contracts, reproducible training, model evaluation, deployment, monitoring, and rollback. Use when building, reviewing, or hardening ML systems beyond one-off notebooks.
react-patterns
React 18/19 patterns including hooks discipline, server/client component boundaries, Suspense + error boundaries, form actions, data fetching, state management decision trees, and accessibility-first composition. Use when writing or reviewing React components.
article-writing
Write articles, guides, blog posts, tutorials, newsletter issues, and other long-form content in a distinctive voice derived from supplied examples or brand guidance. Use when the user wants polished written content longer than a paragraph, especially when voice consistency, structure, and credibility matter.
huggingface-llm-trainer
Train or fine-tune language models with TRL or Unsloth on Hugging Face Jobs, including SFT, DPO, GRPO, reward models, and GGUF conversion. Use for cloud LLM training; use huggingface-vision-trainer for vision tasks.
feature-dev
Guide a feature implementation through a structured seven-phase workflow with deep codebase understanding, clarifying questions, parallel architecture design, and quality review. Use this skill when the user asks to build a new feature, add functionality, or wants a methodical approach to implementation rather than…