- Conduid
- Marketplace
- #cost-reduction
MCP servers tagged cost-reduction
6 live MCP servers tagged "cost-reduction", ranked by trust score. Tags come from package metadata and repository topics, so this list covers servers across every category that work with cost-reduction.
Prompt Optimizer MCP
The prompt linter for LLM applications. Score, analyze, and standardize prompt quality. Open-source MCP server + Node.js API.
Tokenfold
Local, provider-neutral compression for LLM prompts, tool schemas, JSON, logs, and diffs. Lossless, receipt-audited, no ML tax.
Prompt Caching
Automatic prompt caching for Claude Code. Cuts token costs by up to 90% on repeated file reads, bug fix sessions, and long coding conversations - zero config.
Llmtrim
Local proxy that compresses your LLM API requests so you pay less, with no change to the answers. Trims wasted tokens from prompts, history, tool output, and code before they're s…
Graphsift
Token Saver for Claude, GPT-5 & Gemini. 80-150x code context reduction, F1 0.85. AST dependency graph, ranked context selection, 19 CLI compressors, MCP server, agent memory. Save…
Model Finops
Intelligent LLM router that reduces AI API costs by up to 60% through smart model selection and caching. FastAPI service with multi-provider support (Gemini, Claude, OpenRouter) a…