cost optimization skills

253 tagged cost optimization, measured the same way as everything else here.

Browse within: gbrain 58hermes-agent 58knowledge-graph 58Multi-Agent 46azure 38aws 22agentic-devops 20ai-credits 20ai-observability 20aspire 20agentic 17aiops 17deepseek 17gemini 11

clawrouter

01

BlockRunAI/ClawRouter

Skill Claude CodeCodex

Hosted-gateway LLM router — save 88% on inference costs. A local proxy that forwards each request to the blockrun.ai gateway, which routes to the cheapest capable model across 71 models from OpenAI, Anthropic, Google, DeepSeek, xAI, NVIDIA, and more. 5 free NVIDIA models included. Also exposes realtime market data…

6.6k 2d ago A 219 tokens original MIT

phone

02

BlockRunAI/ClawRouter

Skill Claude CodeCodex

Verify phone numbers (carrier + SIM-swap fraud signals) and place AI-powered outbound voice calls via BlockRun's gateway (Twilio + Bland.ai). Trigger when the user asks to look up a number, check fraud risk, buy/rent a phone number, or place an AI voice call. Payment is automatic via x402 from the wallet.

6.6k 2d ago A 72 tokens original MIT

surf

03

BlockRunAI/ClawRouter

Skill Claude CodeCodex

Use this skill — NOT browser or webfetch — for ALL Surf crypto-data calls. 83 endpoints at localhost:8402/v1/surf/ covering CEX/DEX markets, on-chain SQL over 80+ ClickHouse tables (Ethereum, Base, Arbitrum, BSC, TRON, HyperEVM, Tempo), 100M+ labeled wallets, prediction markets (Polymarket + Kalshi), social/CT…

6.6k 2d ago A 130 tokens original MIT

CommonstackAI/UncommonRoute

Skill Claude CodeCodex

Use when publishing UncommonRoute. A release is only complete after all required steps are done: version sync, validation, GitHub push/tag/release, PyPI publish, and npm publish.

690 2mo ago A 44 tokens original MIT

lynkr

05

Fast-Editor/Lynkr

Skill Claude CodeCodex

Universal LLM gateway with intelligent routing, Graphify code intelligence, Distill compression, routing telemetry, Code Mode, and 12+ provider support. 60-80% cost reduction for Claude Code, Cursor, and Codex.

545 2d ago A 50 tokens original Apache-2.0

console-audit

06

nudgebee/nudgebee

Skill Claude CodeCodex

Audit the running app via chrome-devtools MCP — console errors/warnings + failed/slow network by default; perf (LCP/CLS) and Lighthouse (a11y/SEO) are opt-in. Token-efficient (compacts findings in-browser via a console/network hook); the report is shown inline in the response — no files to open. Target one tab / an…

387 2d ago A 0 tokens

create-issue

07

nudgebee/nudgebee

Skill Claude CodeCodex

Create GitHub issues using repo templates (feature, bug, spike).

387 2d ago A 14 tokens

create-pr

08

nudgebee/nudgebee

Skill Claude CodeCodex

Create a pull request with proper formatting, validation, and conventions for this monorepo.

387 2d ago A 17 tokens

k8s-pod-rightsizer

09

initializ/forge

Skill Claude CodeCodex

Analyze Kubernetes workload metrics and produce policy-constrained CPU/memory rightsizing recommendations with optional patch generation and rollback-safe apply.

156 4d ago A 33 tokens original Apache-2.0

AlephantAI/AIephant-AI-Agent-Gateway

Skill Claude CodeCodexCursor

Plan-first, one-time human plan approval; batched execution in either Gated (human confirms between batches) or Auto-loop (continuous run after plan approval); strict task-state updates; automatic 3-round code review. Use for "audit then implement" or "plan first, then execute". Say "auto-loop" or "frad-dotclaude" for…

111 2mo ago A 83 tokens GPL-3.0

route

11

ypollak2/llm-router

Skill Claude CodeCodex

Route a task to the best LLM based on task type and complexity.

75 4d ago A 16 tokens original MIT

routing

12

ypollak2/llm-router

Skill Claude CodeCodex

Route tasks to the cheapest capable model automatically using llm-router MCP tools.

75 4d ago A 0 tokens original MIT

savings

13

ypollak2/llm-router

Skill Claude CodeCodex

Track and report how much you've saved by routing tasks to cheaper models.

75 4d ago A 0 tokens original MIT

vibe

14

sebastienrousseau/dotfiles

Skill Claude CodeCodex

Delegate a coding task to a cheap AI model (Mistral Vibe by default, but any provider Vibe knows about — DeepSeek, Gemini Flash, etc.) and supervise the result via git diff. Claude orchestrates, the cheap model codes. Claude consumes 500-1500 tokens per delegation regardless of how many file reads the delegate does…

74 2d ago A 137 tokens original Apache-2.0

credit-optimizer

15

rafsilva85/credit-optimizer-v5

Skill Claude CodeCodex

Automatically optimize AI agent credit usage by routing tasks to the most cost-efficient execution path. Use when you want to reduce AI API costs by 30-75% without quality loss, classify task complexity before execution, route simple tasks to free or low-cost models, split complex tasks into optimized sub-tasks, or…

49 3mo ago A 73 tokens original MIT

BusyBee3333/sol-governed-codex

Skill Claude CodeCodex

Use when Codex work should route bounded implementation to lower-cost workers and reserve Sol for risky plan review or final evidence gates, or when installing or validating this workflow.

40 1mo ago A 39 tokens original MIT

cost-mode

17

Sagargupta16/claude-cost-optimizer

Skill Claude CodeCodex

Cost-conscious Claude Code mode. Reduces output tokens 40-70% and overall costs 30-60% by enforcing concise responses, smart model routing, and efficient workflow patterns. Keeps full technical accuracy. Activate with /cost-mode or "enable cost mode". Auto-triggers on mentions of budget, cost, tokens, or spending.

35 6d ago A 71 tokens original MIT

eco-audit

18

sup3x/claude-code-eco

Skill Claude CodeCodex

Audit this machine's Claude Code configuration for token waste - effort level, oversized CLAUDE.md files, MCP servers, tool-output caps, env keys that do nothing, startup skill load - and print the exact settings.json edit that fixes each finding. Read-only; it never writes settings. Use when the user asks why their…

33 16d ago A 91 tokens original MIT

eco-max

19

sup3x/claude-code-eco

Skill Claude CodeCodex

Maximum-savings variant of /eco - the same frugality rules PLUS a low reasoning-effort override for the invoked task. Use for routine chores (rename, small fix, quick question, boilerplate) when the user wants absolute minimum token spend; prefer plain /eco for hard or high-stakes work. Works in any language.

33 16d ago A 71 tokens original MIT

eco-report

20

sup3x/claude-code-eco

Skill Claude CodeCodex

Show where the tokens actually went - per-session output, thinking, input and cache accounting read from this machine's Claude Code transcripts. Use when the user asks what they spent, which sessions were expensive, how their prompt cache is doing, or whether /eco is helping. Works in any language.

33 16d ago A 62 tokens original MIT

tokenhabit

21

epoko77-ai/tokenhabit

Skill Claude CodeCodex

A coaching tool that finds habits in past coding-agent conversations that waste tokens, the text units used to process requests and responses, and suggests ways to reduce that waste.

19 15d ago A 371 tokens original MIT

beastmode

22

lac5q/beastmode

Skill Claude CodeCodex

Multi-agent orchestration framework for high-intensity feature implementation. Routes work across model tiers: frontier models (Claude Fable, Kimi 3, Opus; Codex only when explicitly selected) own design, architecture, and review sign-off, while the pinned Luna Max economy lane handles implementation and mechanical…

19 19d ago A 125 tokens

distil-setup

23

dshakes/distil

Skill Claude CodeCodex

Install distil and route an AI coding agent or SDK app through it to cut LLM token costs with certified, reversible context compression. Use when the user wants to set up distil, install distil-llm, reduce their agent/Claude Code/Codex/Gemini token spend, point a baseurl at the distil proxy, or check how much distil…

16 2d ago A 85 tokens