- Conduid
- Marketplace
- #llama-cpp
MCP servers tagged llama-cpp
10 live MCP servers tagged "llama-cpp", ranked by trust score. Tags come from package metadata and repository topics, so this list covers servers across every category that work with llama-cpp.
OGAM
The Swiss Army Knife of Offline AI. Chat, see, speak, and generate images on your phone or Mac — GGUF LLMs, vision, Whisper speech-to-text, Stable Diffusion, tool calling, and loc…
Chat Studio
A Next.js chat workspace featuring server-side AI orchestration, cross-session memory, and native visualization of llama-server metrics.
Atomic Agent
Local First Ai Agent. Optimized for Local Ai models. Long context window. Proper tools callings. Runs privately on your device.
ScreenMind
AI-powered screen memory — captures, analyzes, and lets you search/chat your screen history. Powered by Gemma 4 . 100% local, 100% private.
Omniglass
AI that acts on your screen, not just reads it. Snip → execute. MCP plugins, sandboxed, local LLM.
Universal Intelligence
◉ Universal Intelligence: AI made simple.
Deskdrop
Android keyboard with local AI (Ollama, Whisper, MCP) or cloud (Gemini, Groq, OpenAI)
Lilbee
Run local AI models, search your files and code, and crawl the web, all in one program. Cited answers, local-first, with an MCP server for your coding agent. TUI, CLI, REST API, a…
Opendot
A terminal AI agent you can fully undo - every file and shell action is snapshotted and reversible. Model-agnostic (any LLM), and connects to 1000+ app tools and MCP servers.
Diffusion LLM MCP
FastMCP fleet MCP server for diffusion LMs (dLLM). DiffusionGemma on Goliath RTX 4090 — batch inference, HLE-shaped reasoning, ~200–400 tok/s. Doc phase; llama-diffusion-cli sidec…