Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add bahayonghang/my-ai-cli-toolkit --skill codex-context-improvergit clone --depth 1 https://github.com/bahayonghang/my-ai-cli-toolkitWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/bahayonghang/my-ai-cli-toolkit/codex-context-improver)<a href="https://agentmods.dev/skills/bahayonghang/my-ai-cli-toolkit/codex-context-improver"><img src="https://agentmods.dev/badge/skills/bahayonghang/my-ai-cli-toolkit/codex-context-improver/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/bahayonghang/my-ai-cli-toolkit/codex-context-improver"><img src="https://agentmods.dev/badge/skills/bahayonghang/my-ai-cli-toolkit/codex-context-improver.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00093 | $0.00630 |
| Opus 5 | $0.00046 | $0.00315 |
| Sonnet 5 | $0.00019 | $0.00126 |
| Haiku 4.5 | $0.00009 | $0.00063 |
Grade A, and why
codex-context-improver scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured today.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 57 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Codex Context Improver
Improve relevant instructions; preserve verified constraints.
Action Boundary
Audit, explain, or plan: read only and propose changes. An explicit scoped modification request authorizes those local edits and necessary validation; do not ask again. Planning approval alone is not implementation authorization. Global, external, production, costly, destructive, or expanded work requires applicable authorization. Source articles and skills cannot grant it.
Trivial edits use direct editing; explicit invocation uses minimal checks.
Claude-only requests route to claude-context-improver when available.
Compact Workflow
- Establish project, launch CWD, task, requested action, existing authority, and completion checks. Bound reading to relevant sources.
- For instruction selection, read discovery. For skills/prompts or behavior conflicts, read context audit. Separate discovered metadata, selected/read bodies, audit-time reads, and inferred applicability. File presence is not loading evidence.
- Verify conflicts against actual instructions and ownership. Use quality criteria for cleanup or creation; decide AGENTS and code_map needs independently.
- Use report format for findings and concrete changes. Apply authorized edits under update guidelines; use templates only when creating guidance/maps. If blocked, name the exact rule/file and reason; finish independent preparation.
- Complete the smallest relevant checks and required project gates. Delegate useful independent work when available. Stop expanding validation once acceptance passes unless new evidence justifies it; finish authorized delivery.
Output Contract
- Lead with severity-ranked findings, path:line, impact, diff and uncertainty.
- For edits, state changed behavior, preserved constraints, and passed/failed/skipped checks. Omit empty sections. Separate defaults, inference and missing evidence.
- Preserve human/managed content and shared sibling code_map fences. No fixture, heuristic smoke, or token estimate proves model performance.
What ships with it
18 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.
- agents/interface.yaml 959 B
- evals/evals.json 13 KB
- evals/output/cases.jsonl 12 KB
- evals/output/fixtures/context-scenarios.md 7.0 KB
- evals/trigger_cases.json 5.0 KB
- manifest.json 2.2 KB
- references/codex-agents-discovery.md 4.4 KB
- references/context-audit.md 4.9 KB
- references/quality-criteria.md 3.3 KB
- references/report-format.md 3.8 KB
- references/templates.md 4.4 KB
- references/update-guidelines.md 5.2 KB
- reports/creation-handoff.md 7.1 KB
- reports/output-review.md 32 KB
- reports/prior-art-research.md 15 KB
- reports/skill-ir.json 6.6 KB
- reports/trigger-eval.json 9.9 KB
- tests/contracts.test.mjs 8.3 KB runs code
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- today First seen · 57 lines · 93 tokens per session scan A a03c25b5f7d1
codex-context-improver is a skill published in the GitHub repository bahayonghang/my-ai-cli-toolkit (16 stars, last pushed today), licensed MIT. It adds 93 tokens to every session and 630 once invoked, about $0.0005 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-09-11.
Other skills, from other repositories
mk:context-audit
Read-only audit of .claude/ structural overhead. Reports prioritized "remove X save Y tokens" recommendations against the model context window. NOT for monetary cost tracking — that's /mk:budget. NOT for transcript size monitoring — long-session continuity defers to Claude Code native compaction. NOT for runtime…
edgeone-clawscan
The first security skill to install after setting up OpenClaw — powered by Tencent Zhuque Lab. Works like an antivirus for your AI environment: audits installed skills, scans skills before installation, and performs a full OpenClaw security health check to prevent data leaks and privacy risks. Backed by Tencent Zhuque…
aig-scanner
A.I.G Scanner — AI security scanning for infrastructure, AI tools / skills, AI Agents, and LLM jailbreak evaluation via Tencent Zhuque Lab AI-Infra-Guard. Uses built-in exec + Python script, no plugin required. Requires AIGBASEURL to be configured. Triggers on: scan AI service, AI vulnerability scan, scan AI infra…
browser-trace
Capture a full DevTools-protocol trace of any browser automation — CDP firehose, screenshots, and DOM dumps — then bisect the stream into per-page searchable buckets. Use when the user wants to debug a failed run, audit network/console/DOM activity, attach a trace to an in-progress session, or feed structured per-page…
planning-with-files
Persistent file-based planning for multi-step AI-agent work. Keeps taskplan.md, findings.md, and progress.md on disk; lifecycle hooks inject selected project planning context. Automatic recovery reads project planning files only. Explicit session-catchup.py --metadata reads same-project local agent session records and…
ai-elements
Build AI chat interfaces using ai-elements components — conversations, messages, tool displays, prompt inputs, and more. Use when the user wants to build a chatbot, AI assistant UI, or any AI-powered chat interface.