Madhan230205/token-reducer

⚡ Cut Claude token usage by 90%+ — free, open-source, local-first context compression for Claude Code. Hybrid RAG (BM25 + ONNX vectors), AST chunking, reranking. No API needed.

44Stars on the repository
9Mods indexed here, across every type
4mo agoLast push, which is what freshness is scored on
MITLicence, which decides whether bodies are shown

context-compressor

01

Madhan230205/token-reducer

Agent

Compresses top retrieval chunks into citation-rich summary packets that preserve intent while cutting token usage.

44 4mo ago A 22 tokens original MIT

hybrid-retriever

02

Madhan230205/token-reducer

Agent

Runs hybrid retrieval with strict FTS-first policy, BM25 lexical ranking, vector merge, and top 3-5 reranking for token-efficient context selection.

44 4mo ago A 38 tokens original MIT

noise-chunker

03

Madhan230205/token-reducer

Agent

Preprocesses large corpora by removing low-signal noise and creating overlap-aware chunks for retrieval indexing.

44 4mo ago A 25 tokens original MIT

se-ops-delegate

04

Madhan230205/token-reducer

Agent

Delegates software engineering operations to focused subagents using compressed context packets to keep the main conversation lean.

44 4mo ago A 26 tokens original MIT