Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.

MCP server published in headroomlabs-ai/headroom. Follow the README for the claude mcp add line that matches your setup.