headroomlabs-ai/headroom

Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.

View on GitHub
Python
Stars 60.8k
Forks 4.6k
License Apache-2.0
Open Issues 483
Updated 1d ago