Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add skills/shipwithai/shipwithai-plugins/optimize-harnessnpx skills add ShipWithAI/shipwithai-plugins --skill optimize-harnessgit clone --depth 1 https://github.com/ShipWithAI/shipwithai-pluginsWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/shipwithai/shipwithai-plugins/optimize-harness)<a href="https://agentmods.dev/skills/shipwithai/shipwithai-plugins/optimize-harness"><img src="https://agentmods.dev/badge/skills/shipwithai/shipwithai-plugins/optimize-harness.svg" alt="Measured on agentmods" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.00064 | $0.02188 |
| Opus 5 | $0.00032 | $0.01094 |
| Sonnet 5 | $0.00013 | $0.00438 |
| Haiku 4.5 | $0.00006 | $0.00219 |
Grade A, and why
optimize-harness scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 4d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 192 lines — stays where its author put it; the contents beside it link to each section on GitHub.
/optimize-harness
Reads .claude/logs/*.jsonl (written by observe.py) and surfaces actionable
suggestions: hooks to add, workflow gates to enable — based on what you actually do.
Does NOT repeat static analysis from
/review. Hook binaries, MCP alignment, schema drift → run/shipwithai-starter:review. Runtime usage patterns (frequency, directories, commands) → run this skill.
Flag Handling
--verbose— include LOW confidence suggestions (hidden by default)--window N— analysis window in days, minimum 7 (default: 30)--json— output as JSON instead of markdown report--force— bypass the 7-day cooldown advisory without prompting
Step 1 — Guard Clauses
Evaluate in order. Stop on first failure and output the message below.
Guard 1 — Observability not enabled:
Check: .claude/logs/ exists AND .claude/hooks/observe.py exists
Fail output:
Observability not enabled. Run /shipwithai-starter:setup-observability
first. optimize-harness requires at least 7 days of log data.
Guard 2 — Not enough data:
Check: ≥7 distinct YYYY-MM-DD log filenames within window
AND ≥50 total parsed events
Fail output:
Not enough data yet.
Current: [N] days of logs, [T] tool calls
Required: 7 days AND 50 tool calls
Come back after [earliest_date + 7 days].
N = count of YYYY-MM-DD filenames within window. T = successfully parsed events.
Guard 3 — Cooldown (advisory only, not a hard block):
Check: starter-context.json → optimizer.last_run exists
AND today - last_run < 7 days
Warn: Last optimization was [N] days ago (last run: [date]).
Results may not reflect new patterns. Run anyway? [y/N]
Proceed if user says yes, passes --force, or optimizer.last_run is absent.
Step 2 — Parse & Aggregate
Parse:
- List
.claude/logs/*.jsonlsorted by filename date; filter to window - Read line by line — skip malformed lines silently (no error output)
- Discard:
toolabsent or not in{Edit, Write, Bash} - Discard:
tsabsent or unparseable
What ships with it
2 files beside SKILL.md in the same directory: the scripts, references and assets a skill reads on demand. Not counted in the per-session cost; read them before you install if any of them is executable.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 4d ago First seen · 192 lines · 64 tokens per session scan A 5231f078615d
optimize-harness is a skill published in the GitHub repository ShipWithAI/shipwithai-plugins (10 stars, last pushed 22d ago), licensed MIT. It adds 64 tokens to every session and 2,188 once invoked, about $0.0003 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other skills, from other repositories
claude-code-session-broker
Use when running Arcgentic V2 in Claude Code and fixed Planner, Developer, and Auditor role sessions must be coordinated through a broker.
arcgentic
Use when the user says Arcgentic, asks to use Arcgentic, or wants an idea taken through a complete plan → development → self-audit → external audit workflow in Codex.
espalier-migrate
Migrate an existing harness/espalier install to the current Espalier version — auto-detects which of v0.1→v0.2, v0.3→v0.4, v0.4→v0.5, the v0.5.3 coder-agent patch, v0.5→v0.6 (Stage 1 grill), v0.6→v0.7 (read-only /espalier-ask lane), v0.7→v0.8 (requirements approval gate), the v0.8.1 impact-analysis agent patch, the…
verify-gates
Runs the mechanical quality gates that the arcgentic state machine requires for state transitions. Invoked indirectly by transition.sh OR directly by orchestrator agent before declaring a state transition. Use when about to call transition.sh OR when manually verifying that a round artifact meets the gate criteria.…
session-mode
Use when a project has not yet stored session mode, when a user asks for complete arcgentic workflow execution, or when role identity handoff prompts are needed.
cross-session-handoff
Read, write, snapshot, and lock .arcgentic/state.yaml across planner, dev, audit, and optional test sessions.