fkiene/llmtrim

Local proxy that compresses your LLM API requests so you pay less, with no change to the answers. Trims wasted tokens from prompts, history, tool output, and code before they're sent: -31% input / -74% output, measured live. Any provider, no extra model calls. Also an MCP server and embeddable library (Rust, Python, Ruby, Kotlin, Swift, JS/TS).

222Stars on the repository
1Mods indexed here, across every type
15d agoLast push, which is what freshness is scored on
MPL-2.0Licence, which decides whether bodies are shown

llmtrim

01

fkiene/llmtrim

MCP server Claude CodeCodexCursor +2

MCP server and proxy that compresses LLM prompts, tool output, and replies to cut token cost. Runs locally from the @llmtrim/cli npm package.

222 15d ago A tokens not measured original MPL-2.0