Projects in Awesome Lists tagged with context-compression
A curated list of projects in awesome lists tagged with context-compression .
https://github.com/open-compress/claw-compactor
🦞 LLM Token Compression & Reduction Tool — Cut AI agent token costs by up to 97%. 6-layer deterministic context compression for AI agent workspaces. No LLM required. Prompt compression, context window optimization & cost reduction for any LLM pipeline.
ai-agent-tools ai-cost-saving ai-infrastructure claw-compactor context-compression context-pruning context-window-optimization developer-tools llm-compression llm-context-compression llm-cost-reduction llm-token-compression llm-tools openclaw prompt-compression python-tools token-compression token-optimization token-reduction token-saving
Last synced: 01 Apr 2026
https://github.com/manojmallick/sigmap
97% token reduction for AI coding sessions — zero deps, 31 languages, MCP server
ai claude cli code-context code-intelligence context-compression copilot cursor developer-experience developer-tools gemini github-copilot llm mcp nodejs openai retrieval token-reduction vscode zero-dependencies
Last synced: 17 Jun 2026
https://github.com/learnprompt/cc-harness-skills
Portable CC-inspired skills for memory, verification, multi-agent coordination, context compression, and proactive coding-agent workflows.
agent-harness agent-memory ai-agent codex coding-agent context-compression developer-tools multi-agent openclaw prompt-engineering
Last synced: 12 Jul 2026
https://github.com/borhen68/TokenTamer
A drop-in proxy that compresses bloated code context in real-time, cutting LLM API costs by 50–80% without losing what the model actually needs to know.
ai-coding-agent anthropic context-compression cost-reduction developer-tools llm openai proxy python token-optimization
Last synced: 16 Jul 2026
https://github.com/shouvik12/trooper
LLM reliability layer -keeps agents alive with smart routing, context compaction, and local fallback
ai-gateway context-compression fallback go golang llm llm-proxy local-llm ollama proxy
Last synced: 22 Jul 2026
https://github.com/NodeNestor/claude-rolling-context
Rolling context compression for Claude Code — never hit the context wall. Auto-compresses old messages while keeping recent context verbatim. Zero config, zero latency. Works as a Claude Code plugin.
ai-agent ai-coding anthropic claude claude-code claude-code-extension claude-code-plugin context-compression context-management context-window llm-context prompt-compression rolling-context
Last synced: 16 Jul 2026
https://github.com/dshakes/distil
Compression with a quality contract — cache-aware, causally-pruned LLM context compression for agentic runtimes, certified non-inferior across 7 domains. Works with any SDK.
agents ai-infrastructure anthropic claude-code conformal-prediction context-compression cost-optimization llm llmops mcp openai prompt-caching token-optimization
Last synced: 25 Jul 2026
https://github.com/prantikmedhi/auto-skill-finder
Universal AI skill router — auto-detects best installed skill per prompt + activates caveman mode for ~75% token reduction. Works with Claude, Codex, Cursor, OpenCode, Gemini CLI.
ai-agent auto-skill caveman claude claude-code codex context-compression cursor gemini-cli llm-tools npx-skills opencode prompt-engineering skill-routing token-optimization
Last synced: 15 Jul 2026
https://github.com/yuchen20/context-crumb
Save Token Usage on Unstructured Document 😎. Let agent read docs, memories, prompts with in ultra-compressed mode through a tiny local model.
agent ai context-compaction context-compression skills token token-optimization
Last synced: 01 Jun 2026
https://github.com/mikkoparkkola/ultracos
Lossless, on-device token-cost reduction for Claude Code and LLM coding agents. Free plugin: compresses tool-result output, dedups context, compacts the system prompt — stacks on Anthropic prompt caching. Rust hot path, Python fallback, fail-open. PolyForm Noncommercial.
agentic ai-agents anthropic claude claude-code context-compression cost-optimization developer-tools llm llm-tools mcp prompt-compression rust token-compression token-optimization
Last synced: 02 Jun 2026
https://github.com/ashlrai/ashlr-core-efficiency
Token-efficiency primitives for Claude Code: genome, compression, provider-aware budgeting.
ai anthropic bun claude-code context-compression genome library rag token-efficiency typescript
Last synced: 21 Jun 2026
https://github.com/pozii/tokensaver
MCP server that cuts AI agent token costs by up to 97% — compression, caching, pruning, web extraction
ai claude context-compression cost-reduction fastmcp llm mcp mcp-server openai python token-optimization
Last synced: 24 Jun 2026
https://github.com/castnettech/mnemosyne
LLM context compression and retrieval engine. Zero dependencies. Sub-100ms queries. 40-70% token reduction.
bm25 code-retrieval context-compression developer-tools llm open-source python tfidf token-optimization zero-dependencies
Last synced: 07 Apr 2026
https://github.com/anvanster/compressor
Reduce token usage in AI coding agents (Claude Code, Copilot, Cursor) with mode-switchable instruction packs and tool-output compression hooks — with a benchmark harness that measures the savings. Every optimization is measured or it doesn't ship.
agents-md ai-coding-agents anthropic claude-code cli context-compression cursor developer-tools github-copilot llm prompt-compression token-optimization
Last synced: 04 Jul 2026
https://github.com/anvanster/compressor-vscode
Save tokens in GitHub Copilot agent mode: compressed read/search/outline tools, a savings ticker and report, and instruction-pack management for the compressor toolchain. No network calls.
ai-coding-assistant claude context-compression copilot developer-tools github-copilot language-model-tools llm token-optimization vscode vscode-extension
Last synced: 04 Jul 2026
https://github.com/jpoindexter/winnow
Local-first context compression for AI agents — content-aware, reversible, zero runtime deps. Cuts agent token usage 40-95% while keeping the signal.
ai-agents context-compression llm llmops mcp tokens
Last synced: 04 Jul 2026
https://github.com/alpertarhan/pi-smart-compact
Verification-oriented smart compaction extension for the Pi Coding Agent.
ai-agent bun context-compression llm pi pi-coding-agent pi-extension smart-compaction typescript
Last synced: 08 Jun 2026
https://github.com/chawuciren/evoduck
Local-first AI agent framework for personal workflows and enterprise support, built in Go with memory, knowledge base, MCP, plugins, webchat, WeChat/WeCom channels, and self-update.
agent-framework ai-agent context-compression customer-support enterprise-ai go golang knowledge-base llm mcp memory plugins
Last synced: 09 May 2026
https://github.com/Highlydeveloped-trowel635/context-graph-compressor
Compress chat histories into minimal JSON graphs to reduce token costs and maintain context across different LLM sessions.
age agent ai chatgpt claude-ai claude-skill claude-skills claude-skills-creator compression compression-algorithm context context-compression context-engineering context-management json llm problem-solving prompt-engineering
Last synced: 16 Jul 2026
https://github.com/robbiebusinessacc/justllm
Production LLM calls. Just the three lines. Cross-provider fallback, native caching, and reversible context compression on by default.
ai context-compression litellm llm llm-orchestration prompt-caching
Last synced: 27 Jun 2026
https://github.com/mussolene/contextir
Local-first adaptive context compiler and privacy gateway for LLMs and agents
agents context-compression llm privacy python semantic-ir
Last synced: 19 Jul 2026
https://github.com/seonglae/resrer
Retriever, Summarizer, Reader for LLM ODQA(Open-Domain Question Answering) to increase Information Density
context-compression llm odqa qa question-answering summarizer
Last synced: 30 Jan 2026
https://github.com/samuraiwriter7/kazene-memory-breathing-protocol
A memory breathing protocol for AI systems: structuring what to remember, forget, compress, and flow into implicit behavioral layers.
agentic-ai ai ai-governance ai-safety cognitive-architecture context-compression forgetting kazene memory trace-compaction
Last synced: 27 Jun 2026