headroomlabs-ai/headroom

Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.

View on GitHub
Python
Stars 61.2k
Forks 4.6k
License Apache-2.0
Open Issues 498
Updated 1h ago