Make every token countforAI agent

Get started

Part of

Trusted by engineers at

It starts with compression

Open source · Apache 2.0

Headroom strips repeated boilerplate from tool calls, DB rows, file reads, and RAG payloads before they hit the model. Same answers, fewer tokens, cleaner traces.

pip install "headroom-ai[all]"
GitHub24.9k stars

Start today

Use Headroom your way

Start as a wrapper, proxy, SDK call, or MCP server. Same compression pipeline, whichever surface your stack needs.

Coding agents

Wrap the agent you already use

Launch Claude Code, Codex, Cursor, Aider, Copilot, or OpenClaw through Headroom and compress every tool read before it reaches the model.

pip install "headroom-ai[all]"

headroom wrap claude   # Claude Code
headroom wrap codex    # OpenAI Codex
headroom wrap cursor   # Cursor
headroom wrap aider    # Aider

Enterprise & Scale

Have enterprise needs?

Deploy Headroom in your VPC, customize compression pipelines for proprietary formats, or get dedicated SLA support for high-volume agent workloads.