Skip to content
#

context-compression

Here are 34 public repositories matching this topic...

Local streaming reverse proxy between AI coding agents (Claude Code, Cursor, Codex) and model APIs (Anthropic, OpenAI, Gemini, MiniMax). Meters every token + USD cost, compacts bloated context to cut pay-per-token API spend, and runs shadow-eval to prove quality held. ccusage-style metering + live local dashboard.

  • Updated Jun 18, 2026
  • TypeScript

✨🧠Zenith GPT is a chatbot, but not the kind you dismiss with a shrug. Every interaction has weight here. A 3D orb breathes in response to your mouse, to the cadence of speech, to the rhythm of thought. While the model thinks, the interface doesn't freeze — it waits, visibly, alive.

  • Updated Jul 3, 2026
  • TypeScript

Reduce Claude Code token usage by 70-90% using a free local LLM (Ollama). MCP server + Stop hook with codebase indexing, tool output compression, and turn summarization.

  • Updated Jun 5, 2026
  • TypeScript

Improve this page

Add a description, image, and links to the context-compression topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the context-compression topic, visit your repo's landing page and select "manage topics."

Learn more