headroomlabs-ai/headroom

Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.

View on GitHub
Python
Stars 63.0k
Forks 4.8k
License Apache-2.0
Open Issues 580
Updated 23h ago