-
headroom
Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.
Python ★ 66k 18m agoExplain → -
tokview
See where your LLM tokens actually go — down to the individual tool call. A small, local, zero-config proxy + dashboard for token/cost tracking across Claude, OpenAI, Gemini.
Python ★ 68 1mo agoExplain → -
headroom-switchyard ⑂
No description.
★ 0 4h agoExplain → -
strands-headroom
Headroom context compression for AWS Strands Agents
Python ★ 0 10d agoExplain →
No repos match these filters.