Headroom compresses tool outputs, logs, files, and RAG chunks before they hit LLMs, slashing token counts by 60-95% while maintaining answer accuracy. It's a library, proxy, and MCP server rolled into one Python package. For builders juggling LLM costs and performance, this is a Swiss Army knife.