headroomlabs-ai/headroom
Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.
- Stars
- 74.6k
- Gained, 7 days
- +410
- Gained, 30 days
- n/a
- Reddit posts
- 1
Python · Apache-2.0 · last push 2026-10-08 · created 2026-01-07
GitHub stars, last 90 days
Does it work with local models?
Reports from people who ran it with a model on their own hardware. Tool calling depends on the model as much as on the server, so say which model and quant you used.
No reports yet.
Add your report
Sign in to say whether it worked for you.
Reddit posts linking it
In the official MCP registry
No registry entry points at this repository, so how it runs is not listed here. Check its README.