1. Conduid
  2. Marketplace
  3. #context-compression
Tag

MCP servers tagged context-compression

29 live MCP servers tagged "context-compression", ranked by trust score. Tags come from package metadata and repository topics, so this list covers servers across every category that work with context-compression.

29
live servers
38
avg trust
2
most stars
1

Prompt Optimizer MCP

rishiatlan

The prompt linter for LLM applications. Score, analyze, and standardize prompt quality. Open-source MCP server + Node.js API.

★ 2updated 6 months agoMIT
70
2

Lean Ctx

Hybrid Context Optimizer — Shell Hook + MCP Server. Reduces LLM token consumption by 89-99%. Single Rust binary, zero dependencies.

52
3

Tokless

An effective and efficient coding-agent setup in under 30 seconds: best tools, unified instructions, any OS.

52
4

Headroom Zed

Zed extension for Headroom — context compression for AI agents

44
5

DevProjex

Build safe, token-efficient codebase context for LLMs, AI chats, and coding agents — local-first GUI, TUI, CLI, and a read-only MCP server with Smart Ignore, secret/PII redaction,…

39
6

Context Diamond

Auditable context capsules for LLM handoffs, coding agents, and OpenCode MCP workflows.

39
7

Fak

Agentic Runtime

39
8

TokenPack

TokenPack packs long documents, codebases, PDFs, and folders into compact, evidence-dense LLM context using local embeddings, evidence scoring, and budget-aware selection.

39
9

Token Reducer

⚡ Cut Claude token usage by 90%+ — free, open-source, local-first context compression for Claude Code. Hybrid RAG (BM25 + ONNX vectors), AST chunking, reranking. No API needed.

39
10

Sensei

SenseiIssei

Self-hosted AI workspace that compresses prompts before they leave your machine — 79% fewer tokens, measured. Drop-in gateway for Claude Code, Cursor, Aider and any OpenAI/Anthrop…

★ 2
34
11

Compress Tokens

MCP server that compresses text by removing unnecessary tokens using local LLM surprisal scoring

34
12

Codex Agent Mem

Local-first Model Context Protocol (MCP) memory layer for Codex CLI/Desktop, Claude Code, Gemini CLI, Qwen/DeepSeek/Ollama and agent workflows. SQLite + FTS5 compact context packs…

34
13

SaveContext

Git LFS for your LLM context — 20x Context Savings — an MCP server that keeps big documents out of the model and returns byte-exact quotes for a fraction of the tokens.

34
14

Knitbrain

The local-first brain for coding agents: project memory + workflow intelligence + lossless ~50% context savings. MCP server + LLM proxy, zero cloud.

34
15

Lm Resizer

Rust-native context compression for Claude Code, Codex & MCP agents: filters & compresses noisy tool output (tests, diffs, logs, JSON, provider traffic) before it reaches the LLM…

34
16

Guppi

GUPPI (General-purpose Unifying Pluggable Intelligence) — Local agentic sidecar daemon, Lossless Semantic Tree memory engine, and telemetry deck.

34
17

Capsule

Codex plugin MCP server (codex-plugin-mcp) for token efficiency, token reduction, context compression: exact-recoverable terminal, file, project, web, memory.

34
18

Furl Ctx

Context compression for AI agents — cut token usage costs with retrievable CCR compression. Rust core + Python API, Claude Code & Codex plugin, MCP server. PyPI: furl-ctx

34
19

Agentflare

Run AI coding agents efficiently, and coordinate more than one of them. Zero fluff, zero runtime deps.

34
20

Syntavra

Adaptive context and runtime optimization for coding agents—reducing token waste, externalizing tool output, preserving exact session history, and enforcing secure tool routing.

34
21

Contextomizer

Contextomizer is an ultra-fast, deterministic library for transforming bloated tool outputs, raw APIs, documents, and messy logs into perfectly optimized context for AI Agents 🤖🚀

34
22

Graphsift

Token Saver for Claude, GPT-5 & Gemini. 80-150x code context reduction, F1 0.85. AST dependency graph, ranked context selection, 19 CLI compressors, MCP server, agent memory. Save…

34
23

Cycgraph

Agent context compressor and memory persistent orchestration

34
24

Session Migrator

A cross-agent session memory migration layer: migrate or compress a conversation based on the target model's context window. MCP server + Python library.

34
25

Lelouch

⚡ Absolute command over your AI agent's token budget — local-first compression of noisy tool output before it reaches the model. Claude Code / OpenCode / Codex / Gemini.

34
26

Nuxs

Up to 99% token savings on AI agent context. Free. Provider-agnostic (Claude, Cursor, Codex, Cline, any MCP agent). 91.62% over 626.8M auditable tokens. Try: nuxs.ai/playground

34
27

Meshmind

A single MCP server merging local codebase mapping, recency research, and token compression for AI agents.

34
28

Tare

Lossless-by-default context compression for LLM coding agents — proxy, library, CLI, and MCP server. Local-first, cache-correct, reversible.

34
29

Context Squeeze

Deterministic context-compression MCP server for Claude: squeeze codebases, files, and logs to fit token budgets using tree-sitter and tree-walking — no LLM in the loop. (Rust)

34
← Previous Page 1 of 1 Next →