Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add instructions/yoloshii/clawmem/agents-mdgit clone --depth 1 https://github.com/yoloshii/ClawMemWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/instructions/yoloshii/clawmem/agents-md)<a href="https://agentmods.dev/instructions/yoloshii/clawmem/agents-md"><img src="https://agentmods.dev/badge/instructions/yoloshii/clawmem/agents-md.svg" alt="Measured on agentmods" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.07138 | $0.07138 |
| Opus 5 | $0.03569 | $0.03569 |
| Sonnet 5 | $0.01428 | $0.01428 |
| Haiku 4.5 | $0.00714 | $0.00714 |
Grade B, and why
ClawMem AGENTS.md scanned grade B with 1 finding against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 4d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Reads agent configuration directoriesmediumAgent snooping
.claude/, .codex/, .gemini/ hold keys, settings and other credentials a mod has no legitimate need for.
- **Intermittent `UserPromptSubmit hook timed out after 8s — output discarded`** → **fixed in v0.16.0** (upgrade). Root cause was not inference or host RAM alone: the `context-surfacing` vector leg ran a *synchronous* `s How it starts
The opening of the file, as written. The whole thing — 252 lines — stays where its author put it; the contents beside it link to each section on GitHub.
ClawMem — Agent Quick Reference
On-device retrieval-augmented memory for Claude Code, OpenClaw, and Hermes agents. Hooks auto-inject context (~90%); MCP tools cover targeted recall (~10%). TypeScript on Bun, MIT.
This file is the lean root SSOT — agent-facing essentials only. Deep reference lives in docs/ (linked throughout, indexed at the bottom). CLAUDE.md is an @AGENTS.md import; SKILL.md is the portable on-demand operating reference.
Inference at a glance
Three services — embedding, LLM (query expansion / intent / A-MEM), reranker. Default: all three as llama-server with an in-process node-llama-cpp fallback that auto-downloads on first use (works with no GPU). The bin/clawmem wrapper points at localhost:8088/8089/8090. Always run via bin/clawmem — it sets the endpoints.
Choose a stack:
- native (default) — EmbeddingGemma-300M + qmd-query-expansion-1.7B + qwen3-reranker-0.6B · ~4 GB or in-process · permissive, commercial OK · zero-config.
- z / SOTA — zembed-1 + qmd-query-expansion-1.7B + zerank-2 seq-cls sidecar · ~16 GB · CC-BY-NC-4.0, non-commercial only · best recall.
- cloud embedding — Jina/OpenAI/Voyage/Cohere · embedding only (LLM + reranker stay local) · no local GPU needed.
Landmines:
- The zerank-2 GGUF is inert — llama.cpp drops the score head → ranking silently collapses to RRF. Use the seq-cls sidecar; verify with
clawmem rerank-health(liveness ≠ correctness). - A squatted port answers HTTP but serves nothing — an unrelated service on 8088/8089 used to disable enrichment silently forever. Since v0.37.0 persistent HTTP errors trip the 60s cooldown (405/501 instantly, other non-2xx after 3 consecutive) so the fallback engages, and
clawmem doctorshape-probesCLAWMEM_LLM_URLwith a real completion. -ubmust equal-bfor embedding/reranking (non-causal attention) orllama-serverasserts.- Changing embedding dimensions →
clawmem embed --force(full re-embed). - Changing the embedding model (even at the same dimension) →
clawmem embed --force; querying otherwise now throwsVecReadModelMismatchErrorinstead of serving cosine-meaningless results (v0.18.0). CLAWMEM_NO_LOCAL_MODELS=trueto fail fast instead of silent CPU fallback.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 4d ago First seen · 252 lines · 7,138 tokens per session scan B 87f1020b70f9
ClawMem AGENTS.md is an instructions file published in the GitHub repository yoloshii/ClawMem (206 stars, last pushed 16d ago), licensed MIT. It adds 7,138 tokens to every session, about $0.0357 per session on Opus 5. A static security scan graded it B with 1 finding (reads agent configuration directories). No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other instructions, from other repositories
SecureContext CLAUDE.md
Instructions for iampantherr/SecureContext, covering securecontext — agent instructions, what this project is, directory layout, build commands and 13 mcp tools.
subcog CLAUDE.md
Instructions for zircote/subcog, covering claude.md, project overview, key capabilities, build commands and primary development workflow.
ariadne AGENTS.md
Instructions for mclaut/ariadne, covering agents.md — ariadne, what this is, build / test / lint, architecture and runtime layout (not the repo).
mnemon-memory-mcp CLAUDE.md
Claude Code instructions for nikitacometa/mnemon-memory-mcp, covering mnemon-mcp, commands, tech stack, architecture and database.
ariadne CLAUDE.md
Instructions for mclaut/ariadne, covering claude.md — ariadne, what this is, build / test / lint, architecture and runtime layout (not the repo).
llm-wiki-memory AGENTS.md
Instructions for ctxr-dev/llm-wiki-memory, covering agents.md, development discipline (.agents/), layout, tests and conventions.