1. Conduid
  2. Marketplace
  3. #llama-cpp
Tag

MCP servers tagged llama-cpp

10 live MCP servers tagged "llama-cpp", ranked by trust score. Tags come from package metadata and repository topics, so this list covers servers across every category that work with llama-cpp.

10
live servers
48
avg trust
18
most stars
1

OGAM

The Swiss Army Knife of Offline AI. Chat, see, speak, and generate images on your phone or Mac — GGUF LLMs, vision, Whisper speech-to-text, Stable Diffusion, tool calling, and loc…

64
2

Chat Studio

JoeCastrom

A Next.js chat workspace featuring server-side AI orchestration, cross-session memory, and native visualization of llama-server metrics.

★ 18updated 7 months agoMIT
61
3

Atomic Agent

Local First Ai Agent. Optimized for Local Ai models. Long context window. Proper tools callings. Runs privately on your device.

59
4

ScreenMind

AI-powered screen memory — captures, analyzes, and lets you search/chat your screen history. Powered by Gemma 4 . 100% local, 100% private.

52
5

Omniglass

goshtasb

AI that acts on your screen, not just reads it. Snip → execute. MCP plugins, sandboxed, local LLM.

★ 3updated 6 months ago
51
6

Universal Intelligence

◉ Universal Intelligence: AI made simple.

44
7

Deskdrop

Android keyboard with local AI (Ollama, Whisper, MCP) or cloud (Gemini, Groq, OpenAI)

39
8

Lilbee

Run local AI models, search your files and code, and crawl the web, all in one program. Cited answers, local-first, with an MCP server for your coding agent. TUI, CLI, REST API, a…

39
9

Opendot

A terminal AI agent you can fully undo - every file and shell action is snapshotted and reversible. Model-agnostic (any LLM), and connects to 1000+ app tools and MCP servers.

39
10

Diffusion LLM MCP

FastMCP fleet MCP server for diffusion LMs (dLLM). DiffusionGemma on Goliath RTX 4090 — batch inference, HLE-shaped reasoning, ~200–400 tok/s. Doc phase; llama-diffusion-cli sidec…

34
← Previous Page 1 of 1 Next →