An open API service indexing awesome lists of open source software.

Projects in Awesome Lists tagged with context-compression

A curated list of projects in awesome lists tagged with context-compression .

https://github.com/open-compress/claw-compactor

🦞 LLM Token Compression & Reduction Tool — Cut AI agent token costs by up to 97%. 6-layer deterministic context compression for AI agent workspaces. No LLM required. Prompt compression, context window optimization & cost reduction for any LLM pipeline.

ai-agent-tools ai-cost-saving ai-infrastructure claw-compactor context-compression context-pruning context-window-optimization developer-tools llm-compression llm-context-compression llm-cost-reduction llm-token-compression llm-tools openclaw prompt-compression python-tools token-compression token-optimization token-reduction token-saving

Last synced: 01 Apr 2026

https://github.com/learnprompt/cc-harness-skills

Portable CC-inspired skills for memory, verification, multi-agent coordination, context compression, and proactive coding-agent workflows.

agent-harness agent-memory ai-agent codex coding-agent context-compression developer-tools multi-agent openclaw prompt-engineering

Last synced: 12 Jul 2026

https://github.com/borhen68/TokenTamer

A drop-in proxy that compresses bloated code context in real-time, cutting LLM API costs by 50–80% without losing what the model actually needs to know.

ai-coding-agent anthropic context-compression cost-reduction developer-tools llm openai proxy python token-optimization

Last synced: 16 Jul 2026

https://github.com/shouvik12/trooper

LLM reliability layer -keeps agents alive with smart routing, context compaction, and local fallback

ai-gateway context-compression fallback go golang llm llm-proxy local-llm ollama proxy

Last synced: 22 Jul 2026

https://github.com/NodeNestor/claude-rolling-context

Rolling context compression for Claude Code — never hit the context wall. Auto-compresses old messages while keeping recent context verbatim. Zero config, zero latency. Works as a Claude Code plugin.

ai-agent ai-coding anthropic claude claude-code claude-code-extension claude-code-plugin context-compression context-management context-window llm-context prompt-compression rolling-context

Last synced: 16 Jul 2026

https://github.com/dshakes/distil

Compression with a quality contract — cache-aware, causally-pruned LLM context compression for agentic runtimes, certified non-inferior across 7 domains. Works with any SDK.

agents ai-infrastructure anthropic claude-code conformal-prediction context-compression cost-optimization llm llmops mcp openai prompt-caching token-optimization

Last synced: 25 Jul 2026

https://github.com/prantikmedhi/auto-skill-finder

Universal AI skill router — auto-detects best installed skill per prompt + activates caveman mode for ~75% token reduction. Works with Claude, Codex, Cursor, OpenCode, Gemini CLI.

ai-agent auto-skill caveman claude claude-code codex context-compression cursor gemini-cli llm-tools npx-skills opencode prompt-engineering skill-routing token-optimization

Last synced: 15 Jul 2026

https://github.com/yuchen20/context-crumb

Save Token Usage on Unstructured Document 😎. Let agent read docs, memories, prompts with in ultra-compressed mode through a tiny local model.

agent ai context-compaction context-compression skills token token-optimization

Last synced: 01 Jun 2026

https://github.com/mikkoparkkola/ultracos

Lossless, on-device token-cost reduction for Claude Code and LLM coding agents. Free plugin: compresses tool-result output, dedups context, compacts the system prompt — stacks on Anthropic prompt caching. Rust hot path, Python fallback, fail-open. PolyForm Noncommercial.

agentic ai-agents anthropic claude claude-code context-compression cost-optimization developer-tools llm llm-tools mcp prompt-compression rust token-compression token-optimization

Last synced: 02 Jun 2026

https://github.com/ashlrai/ashlr-core-efficiency

Token-efficiency primitives for Claude Code: genome, compression, provider-aware budgeting.

ai anthropic bun claude-code context-compression genome library rag token-efficiency typescript

Last synced: 21 Jun 2026

https://github.com/pozii/tokensaver

MCP server that cuts AI agent token costs by up to 97% — compression, caching, pruning, web extraction

ai claude context-compression cost-reduction fastmcp llm mcp mcp-server openai python token-optimization

Last synced: 24 Jun 2026

https://github.com/castnettech/mnemosyne

LLM context compression and retrieval engine. Zero dependencies. Sub-100ms queries. 40-70% token reduction.

bm25 code-retrieval context-compression developer-tools llm open-source python tfidf token-optimization zero-dependencies

Last synced: 07 Apr 2026

https://github.com/anvanster/compressor

Reduce token usage in AI coding agents (Claude Code, Copilot, Cursor) with mode-switchable instruction packs and tool-output compression hooks — with a benchmark harness that measures the savings. Every optimization is measured or it doesn't ship.

agents-md ai-coding-agents anthropic claude-code cli context-compression cursor developer-tools github-copilot llm prompt-compression token-optimization

Last synced: 04 Jul 2026

https://github.com/anvanster/compressor-vscode

Save tokens in GitHub Copilot agent mode: compressed read/search/outline tools, a savings ticker and report, and instruction-pack management for the compressor toolchain. No network calls.

ai-coding-assistant claude context-compression copilot developer-tools github-copilot language-model-tools llm token-optimization vscode vscode-extension

Last synced: 04 Jul 2026

https://github.com/jpoindexter/winnow

Local-first context compression for AI agents — content-aware, reversible, zero runtime deps. Cuts agent token usage 40-95% while keeping the signal.

ai-agents context-compression llm llmops mcp tokens

Last synced: 04 Jul 2026

https://github.com/alpertarhan/pi-smart-compact

Verification-oriented smart compaction extension for the Pi Coding Agent.

ai-agent bun context-compression llm pi pi-coding-agent pi-extension smart-compaction typescript

Last synced: 08 Jun 2026

https://github.com/chawuciren/evoduck

Local-first AI agent framework for personal workflows and enterprise support, built in Go with memory, knowledge base, MCP, plugins, webchat, WeChat/WeCom channels, and self-update.

agent-framework ai-agent context-compression customer-support enterprise-ai go golang knowledge-base llm mcp memory plugins

Last synced: 09 May 2026

https://github.com/robbiebusinessacc/justllm

Production LLM calls. Just the three lines. Cross-provider fallback, native caching, and reversible context compression on by default.

ai context-compression litellm llm llm-orchestration prompt-caching

Last synced: 27 Jun 2026

https://github.com/mussolene/contextir

Local-first adaptive context compiler and privacy gateway for LLMs and agents

agents context-compression llm privacy python semantic-ir

Last synced: 19 Jul 2026

https://github.com/seonglae/resrer

Retriever, Summarizer, Reader for LLM ODQA(Open-Domain Question Answering) to increase Information Density

context-compression llm odqa qa question-answering summarizer

Last synced: 30 Jan 2026

https://github.com/samuraiwriter7/kazene-memory-breathing-protocol

A memory breathing protocol for AI systems: structuring what to remember, forget, compress, and flow into implicit behavioral layers.

agentic-ai ai ai-governance ai-safety cognitive-architecture context-compression forgetting kazene memory trace-compaction

Last synced: 27 Jun 2026