Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx agentmods add skills/open-gitagent/opengap/compute-laddernpx skills add open-gitagent/opengap --skill compute-laddergit clone --depth 1 https://github.com/open-gitagent/opengapWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/open-gitagent/opengap/compute-ladder)<a href="https://agentmods.dev/skills/open-gitagent/opengap/compute-ladder"><img src="https://agentmods.dev/badge/skills/open-gitagent/opengap/compute-ladder.svg" alt="Measured on agentmods" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5 | $0.00048 | $0.00677 |
| Opus 5 | $0.00024 | $0.00338 |
| Sonnet 5 | $0.00010 | $0.00135 |
| Haiku 4.5 | $0.00005 | $0.00068 |
Grade C, and why
compute-ladder scanned grade C with 2 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 4d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Downloads and executes remote codehighSupply chain
curl | sh runs whatever the server returns today, which is not necessarily what it returned when this was reviewed.
curl -s localhost:11434/api/tags | python3 -c "import json,sys; d=json.load(sys.stdin); print('TIER-0 OK:', len(d['models']), 'models')" Makes network callslowCapability
Not a fault in itself. Listed so you know the mod talks to something, and to what.
curl -s localhost:11434/api/tags | python3 -c "import json,sys; d=json.load(sys.stdin); print('TIER-0 OK:', len(d['models']), 'models')" What it actually says
Compute Ladder
Tier Definitions
Tier 0 — Local (never dies, zero cost)
ollama/qwen3-coder:latest primary, MoE, 128 TPS, 64k ctx
ollama/gpt-oss:20b fallback, 32k ctx HARD LIMIT
Tier 1 — Fast Free Cloud (up to 2100 TPS)
cerebras/qwen-3-235b-a22b-instruct-2507 235B MoE, fast free
cerebras/llama3.1-8b 8B, ultra-fast light tasks
Tier 2 — Free Cloud (normal latency)
openrouter/z-ai/glm-4.5-air
openrouter/qwen/qwen3-coder
Tier 3 — Free Cloud Deep Reasoning
openrouter/nousresearch/hermes-3-llama-3.1-405b:free
Tier 4 — Break-Glass (paid, restricted use)
openrouter/anthropic/claude-opus-4.6 [narco-check and audit ONLY]
Fallback Rules
DO fallback when:
- HTTP 429 (rate limited)
- Connection timeout (> 90s)
- Stream death / incomplete response
DO NOT fallback when:
- Task seems "complex" or "important"
- You want "better" output quality
- Previous attempt gave a poor answer
Use the primary model. Iterate. Fallback is for infrastructure failure, not preference.
Health Check
# Tier 0
curl -s localhost:11434/api/tags | python3 -c "import json,sys; d=json.load(sys.stdin); print('TIER-0 OK:', len(d['models']), 'models')"
# Tier 1
curl -s -o /dev/null -w "%{http_code}" \
-H "Authorization: Bearer $CEREBRAS_API_KEY" \
https://api.cerebras.ai/v1/models
# Tier 4
curl -s -o /dev/null -w "%{http_code}" \
https://openrouter.ai/api/v1/models
Cost Guard
# Check today's OpenRouter spend
curl -s "https://openrouter.ai/api/v1/auth/key" \
-H "Authorization: Bearer $OPENROUTER_API_KEY" \
| python3 -c "
import json,sys
d=json.load(sys.stdin)['data']
print(f'today: \${d[\"usage_daily\"]:.2f} | week: \${d[\"usage_weekly\"]:.2f} | month: \${d[\"usage_monthly\"]:.2f}')
"
Daily > $5: flag to Ludo. Weekly > $50: flag immediately — tier-4 model is likely being over-used.
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 4d ago First seen · 79 lines · 48 tokens per session scan C 63151cc4c69c
compute-ladder is a skill published in the GitHub repository open-gitagent/opengap (2,925 stars, last pushed 2mo ago), licensed MIT. It adds 48 tokens to every session and 677 once invoked, about $0.0002 per session on Opus 5. A static security scan graded it C with 2 findings (downloads and executes remote code, makes network calls). No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-30.
Other skills, from other repositories
hive.slack-notifications-setup
Set up a Slack notification channel (Sentinel) for a colony by driving the browser — reuse or create the "Hive Sentinel" Slack app from a JSON manifest, install it, capture the bot + app tokens, create/select the channel via the Slack API, and turn Sentinel on so the colony can ping the user on Slack and accept…
hive.chart-creation-foundations
Required reading whenever any chart tool is available. Teaches the one-tool embedding contract (call chartrender → live chart appears in chat AND a downloadable PNG lands in the queen session dir), the ECharts (data viz) vs Mermaid (structural diagrams) decision, the BI/financial-grade aesthetic baseline (no…
browser-edge-cases
SOP for debugging browser automation failures on complex websites. Use when browser tools fail on specific sites like LinkedIn, Twitter/X, SPAs, or sites with Shadow DOM.
hive.pdf
Read, write, merge, split, rotate, watermark, encrypt, and OCR PDF files using Python (pypdf, pdfplumber, reportlab, pypdfium2) and command-line tools (poppler-utils, qpdf). Use when the user asks to extract text/tables/images from a PDF, create or modify a PDF, combine or split PDFs, OCR a scanned PDF…
session-investigator
Investigate fast-agent session and history files to diagnose issues. Use when a session ended unexpectedly, when debugging tool loops, when correlating sub-agent traces with main sessions, or when analyzing conversation flow and timing. Covers session.json metadata, history JSON format, message structure, tool…
hive.note-taking
Maintain a free-form scratchpad of decisions, extracted values, and open questions so context pruning doesn't lose anything you still need.