- Conduid
- Marketplace
- #token-efficiency
MCP servers tagged token-efficiency
24 live MCP servers tagged "token-efficiency", ranked by trust score. Tags come from package metadata and repository topics, so this list covers servers across every category that work with token-efficiency.
Remindb
An agentic memory database that cuts session tokens by 75–99%. One portable SQLite file — your agent's memory, anywhere.
Agent Engineering Infrastructure
Deploys your OS, databases, and SSL on your VPS in just 10 minutes. Orchestrates a team of AI agents for coding, marketing, and sales. The built-in optimizer saves up to 90% on to…
Hedl
Token-efficient data serialization for LLM/AI. 50% fewer tokens than JSON, 93% better value/token. Rust, schema validation, LSP.
Agent Coherence
The data integrity layer that stops AI agents from silently corrupting shared state — across sessions, time, and concurrent runs.
Gcf
The AI-native wire format for structured data. 100% comprehension on every frontier model. 50-92% fewer tokens than JSON. 43B+ lossless round-trips across 17 formats. Spec v3.5.1…
Cdc
Convert MCP servers into Claude Code skills — not MCP connections. Token-minimal, code-execution pattern.
Rag Runtime Kernel
Persistent memory and deterministic state control for any LLM, in two tiers. Tier 1: a single markdown spec you paste into any chat - no install. Tier 2: a formally-verified Pytho…
Qarinah
Qarinah is evidence-linked project memory for coding agents - compact enough to save context, inspectable enough to trust.
Agentbrain
Local-first long-term memory for AI agents - plain Markdown vault + thin MCP server. CJK-aware BM25, token-efficient by design.
Dotnet Toolkit
AI-native .NET code intelligence for faster, more accurate codebase navigation with lower token usage.
Capsule
Codex plugin MCP server (codex-plugin-mcp) for token efficiency, token reduction, context compression: exact-recoverable terminal, file, project, web, memory.
Gcf Rust
GCF Rust implementation. 100% LLM comprehension on every frontier model. 50-92% fewer tokens than JSON. 43B+ round-trips verified. Zero dependencies.
Gcf Python
GCF Python implementation. 100% LLM comprehension on every frontier model. 50-92% fewer tokens than JSON. 43B+ round-trips verified. Zero dependencies.
Gcf Go
GCF Go implementation. 100% LLM comprehension on every frontier model. 50-92% fewer tokens than JSON. 43B+ round-trips verified. Zero dependencies.
Furl Ctx
Context compression for AI agents — cut token usage costs with retrievable CCR compression. Rust core + Python API, Claude Code & Codex plugin, MCP server. PyPI: furl-ctx
Gcf Typescript
GCF TypeScript implementation. 100% LLM comprehension on every frontier model. 50-92% fewer tokens than JSON. 43B+ round-trips verified. Zero dependencies.
Gcf Swift
GCF Swift implementation. 100% LLM comprehension on every frontier model. 50-92% fewer tokens than JSON. 43B+ round-trips verified. Zero dependencies.
Winnow
Winnow — an embedded TypeScript SDK that lets an agent use many MCP servers without context bloat.
Tree Sitter Gcf
Tree-sitter grammar for GCF. Full spec compliance, WASM build for browser playground.
Urusilla
Open AI-agent interoperability research. Bring your own agent, complete verified public work, and earn contribution credit; if URSL launches, 1 credit = 1 URSL.
Axiolex
Shared tool discovery and execution for AI clients and applications across MCP, A2A, and internal tools—intent-ranked Top-K retrieval with centrally managed provider and tool chan…
Mugil Ide
Token-efficient autonomous AI IDE — a browser-based coding agent with an xterm.js terminal, smart caching, automatic model routing, and MCP tooling. Minimizes token cost while max…
Claude Token Efficient Setup
A tested setup guide and prompt workflow for pairing Claude Desktop with code-review-graph and filesystem MCP on macOS — for token-efficient AI-assisted development on large codeb…
Tokenectomy
Autonomous M2M MCP server that scrubs framework noise & redacts secrets from AI agent error logs before they hit your context window