Skip to content
#

context-compression

Here are 307 public repositories matching this topic...

Rolling context compression for Claude Code: a zero-dependency proxy that summarizes old messages and keeps recent turns verbatim — beat the context-window limit and cut token & cache costs. Turn-aware keep policy, concurrency-hardened.

  • Updated Jul 29, 2026
  • Python

Unified compression pipeline for LLM inputs: trim irrelevant rows and fields, re-encode the rest in the cheapest lossless format (TOON / JSON / CSV, measured on your tokenizer), and account for every token saved with an audit trail of what was removed. 95% fewer tokens on realistic payloads, 100% needle recall.

  • Updated Sep 22, 2026
  • Python

LLM token compression via images: a transparent Antrhopic/OpenAI proxy that renders agent context (system prompt, tool docs, tool output, history) to pixels, so coding-agent CLIs like OpenCode send far fewer input tokens. Keeps tool calls and multi-turn intact. ~39% fewer tokens on real runs.

  • Updated Jul 10, 2026
  • Python

Add this topic to your repo

To associate your repository with the context-compression topic, visit your repo's landing page and select "manage topics."

Learn more