- Conduid
- Marketplace
- #context-compression
MCP servers tagged context-compression
29 live MCP servers tagged "context-compression", ranked by trust score. Tags come from package metadata and repository topics, so this list covers servers across every category that work with context-compression.
Prompt Optimizer MCP
The prompt linter for LLM applications. Score, analyze, and standardize prompt quality. Open-source MCP server + Node.js API.
Lean Ctx
Hybrid Context Optimizer — Shell Hook + MCP Server. Reduces LLM token consumption by 89-99%. Single Rust binary, zero dependencies.
Tokless
An effective and efficient coding-agent setup in under 30 seconds: best tools, unified instructions, any OS.
Headroom Zed
Zed extension for Headroom — context compression for AI agents
DevProjex
Build safe, token-efficient codebase context for LLMs, AI chats, and coding agents — local-first GUI, TUI, CLI, and a read-only MCP server with Smart Ignore, secret/PII redaction,…
Context Diamond
Auditable context capsules for LLM handoffs, coding agents, and OpenCode MCP workflows.
Fak
Agentic Runtime
TokenPack
TokenPack packs long documents, codebases, PDFs, and folders into compact, evidence-dense LLM context using local embeddings, evidence scoring, and budget-aware selection.
Token Reducer
⚡ Cut Claude token usage by 90%+ — free, open-source, local-first context compression for Claude Code. Hybrid RAG (BM25 + ONNX vectors), AST chunking, reranking. No API needed.
Sensei
Self-hosted AI workspace that compresses prompts before they leave your machine — 79% fewer tokens, measured. Drop-in gateway for Claude Code, Cursor, Aider and any OpenAI/Anthrop…
Compress Tokens
MCP server that compresses text by removing unnecessary tokens using local LLM surprisal scoring
Codex Agent Mem
Local-first Model Context Protocol (MCP) memory layer for Codex CLI/Desktop, Claude Code, Gemini CLI, Qwen/DeepSeek/Ollama and agent workflows. SQLite + FTS5 compact context packs…
SaveContext
Git LFS for your LLM context — 20x Context Savings — an MCP server that keeps big documents out of the model and returns byte-exact quotes for a fraction of the tokens.
Knitbrain
The local-first brain for coding agents: project memory + workflow intelligence + lossless ~50% context savings. MCP server + LLM proxy, zero cloud.
Lm Resizer
Rust-native context compression for Claude Code, Codex & MCP agents: filters & compresses noisy tool output (tests, diffs, logs, JSON, provider traffic) before it reaches the LLM…
Guppi
GUPPI (General-purpose Unifying Pluggable Intelligence) — Local agentic sidecar daemon, Lossless Semantic Tree memory engine, and telemetry deck.
Capsule
Codex plugin MCP server (codex-plugin-mcp) for token efficiency, token reduction, context compression: exact-recoverable terminal, file, project, web, memory.
Furl Ctx
Context compression for AI agents — cut token usage costs with retrievable CCR compression. Rust core + Python API, Claude Code & Codex plugin, MCP server. PyPI: furl-ctx
Agentflare
Run AI coding agents efficiently, and coordinate more than one of them. Zero fluff, zero runtime deps.
Syntavra
Adaptive context and runtime optimization for coding agents—reducing token waste, externalizing tool output, preserving exact session history, and enforcing secure tool routing.
Contextomizer
Contextomizer is an ultra-fast, deterministic library for transforming bloated tool outputs, raw APIs, documents, and messy logs into perfectly optimized context for AI Agents 🤖🚀
Graphsift
Token Saver for Claude, GPT-5 & Gemini. 80-150x code context reduction, F1 0.85. AST dependency graph, ranked context selection, 19 CLI compressors, MCP server, agent memory. Save…
Cycgraph
Agent context compressor and memory persistent orchestration
Session Migrator
A cross-agent session memory migration layer: migrate or compress a conversation based on the target model's context window. MCP server + Python library.
Lelouch
⚡ Absolute command over your AI agent's token budget — local-first compression of noisy tool output before it reaches the model. Claude Code / OpenCode / Codex / Gemini.
Nuxs
Up to 99% token savings on AI agent context. Free. Provider-agnostic (Claude, Cursor, Codex, Cline, any MCP agent). 91.62% over 626.8M auditable tokens. Try: nuxs.ai/playground
Meshmind
A single MCP server merging local codebase mapping, recency research, and token compression for AI agents.
Tare
Lossless-by-default context compression for LLM coding agents — proxy, library, CLI, and MCP server. Local-first, cache-correct, reversible.
Context Squeeze
Deterministic context-compression MCP server for Claude: squeeze codebases, files, and logs to fit token budgets using tree-sitter and tree-walking — no LLM in the loop. (Rust)