- Conduid
- Marketplace
- #token-savings
MCP servers tagged token-savings
8 live MCP servers tagged "token-savings", ranked by trust score. Tags come from package metadata and repository topics, so this list covers servers across every category that work with token-savings.
Memos
Self-evolving memory OS for LLM & AI Agents: ultra-persistent memory, hybrid-retrieval, and cross-task skill reuse, with 35.24% token savings
Opentoken
Token-saving companion for OpenCode — 42 compression layers, zero risk, no caveman speak
Omni
The Context OS for Autonomous AI Agents. Distill terminal noise into pure semantic signal, stop agent hallucinations, and cut token costs by up to 90%.
Houtini Lm
MCP server that saves Claude Code tokens by delegating bounded tasks to local or cloud LLMs. 93% token savings benchmarked. Works with LM Studio, Ollama, vLLM, DeepSeek, Groq, Cer…
Git Courer
MCP server for Git with local Ollama — zero tokens for git operations
PsYcGoD/sage
`, stores command history locally, and returns compressed terminal output to reduce context noise.
Scribe
MCP server that delegates bulk file reading and doc/boilerplate writing to cheaper OpenAI-compatible models
Codex Agent Mem
Local-first Model Context Protocol (MCP) memory layer for Codex CLI/Desktop, Claude Code, Gemini CLI, Qwen/DeepSeek/Ollama and agent workflows. SQLite + FTS5 compact context packs…