headroomlabs-ai/headroom

Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.

View on GitHub
Python
Stars 62.2k
Forks 4.7k
License Apache-2.0
Open Issues 528
Updated 8h ago