- Conduid
- Marketplace
- #local-llm
MCP servers tagged local-llm
54 live MCP servers tagged "local-llm", ranked by trust score. Tags come from package metadata and repository topics, so this list covers servers across every category that work with local-llm.
Anything LLM
The all-in-one Desktop & Docker AI application with built-in RAG, AI agents, No-code agent builder, MCP compatibility, and more.
Awesome Openclaw
A curated list of OpenClaw resources, tools, skills, tutorials & articles. OpenClaw (formerly Moltbot / Clawdbot) — open-source self-hosted AI agent for WhatsApp, Telegram, Discor…
Client For Ollama
A text-based user interface (TUI) client for interacting with MCP servers using Ollama. Features include agent mode, multi-server, model switching, streaming responses, tool manag…
Tome
a magical LLM desktop client that makes it easy for *anyone* to use LLMs and MCP
Ollama MCP Bridge
Extend the Ollama API with dynamic AI tool integration from multiple MCP (Model Context Protocol) servers. Fully compatible, transparent, and developer-friendly, ideal for buildin…
Local Deep Research
Local Deep Research achieves ~95% on SimpleQA benchmark (tested with GPT-4.1-mini). Supports local and cloud LLMs (Ollama, Google, Anthropic, ...). Searches 10+ sources - arXiv, P…
Ouroboros
Ouroboros — self-creating AI agent. Born Feb 16, 2026.
Awesome Local LLM
A curated list of awesome platforms, tools, practices and resources that helps run LLMs locally
Open Multi Agent
TypeScript AI agent orchestration framework with dynamic workflows. Describe the goal, not the graph: a coordinator plans the task DAG at runtime and runs it on any LLM (Claude, C…
Llmc
One stop shop - Local-first RAG stack with intelligent polyglot-code/docs, remote code execution, local llama enrichment, progressive disclosure tools, mcp server, sandboxed secur…
Atomic Chat
Local AI app and inference engine for agents. Run open-weight LLMs locally — private, 100% offline on your computer.
Atomic Agent
Local First Ai Agent. Optimized for Local Ai models. Long context window. Proper tools callings. Runs privately on your device.
Dark Moon
Autonomous AI pentesting engine, continuous offensive security across web, cloud, AD & Kubernetes. Agentic reasoning + real exploit execution deliver proof-based vulnerabilities.…
Local Analyst
Talk to your data locally 💬📊. A private AI Data Analyst built with the Model Context Protocol (MCP), Ollama, and SQLite. Turn natural language into SQL queries without data leavin…
Prism MCP Server
Prism-Coder v12.5: The world's first O(1) Cognitive Memory Architecture for AI Agents. 100% Tool-Call Accuracy (BFCL Gold Certified), 54 Agent Skills, Zero-Search Retrieval (HDC/H…
Roampal
Memory that learns what works.
Awesome Second Brain
A curated solutions to building a self-evolving second brain that helps AI agents understand your personal and team context.
Synto
More than just Karpathy’s LLM Wiki, 100% local with Ollama. Drop Markdown notes → AI extracts concepts → your Obsidian wiki auto-links and grows. Zero sharing. Your notes stay you…
Bagidea Office
A living AI-agent office on your desktop wallpaper — Claude Code agents that walk, work, delegate, learn & hold meetings. Per-agent swappable models (Claude/GLM/DeepSeek/Qwen/Kimi…
Captain Claw
Self-hosted framework for orchestrating fleets of specialist AI agents — ensemble reasoning and a full agentic coding pipeline, model-agnostic and local-friendly.
Omk
Evidence-gated runner for Codex, Claude Code, OpenCode, and local coding agents. Routes tasks into scoped DAG lanes with replayable artifacts.
Agentty
AI pair programming in your terminal — one static binary, sub-ms startup, any model
Ollamarama Matrix
Local AI chatbot for Matrix with infinite personalities, tool calling, and MCP — powered by Ollama
HARTOS
An AI-native OS. Models run on your own hardware, nodes federate peer-to-peer with no broker, and the API is OpenAI-compatible. Boots, has its own Wayland compositor, and runs on…
LLM CLI
Easily pipe anything into an LLM
Houtini Lm
MCP server that saves Claude Code tokens by delegating bounded tasks to local or cloud LLMs. 93% token savings benchmarked. Works with LM Studio, Ollama, vLLM, DeepSeek, Groq, Cer…
Ramibot
RamiBot v3.8.0 is a local-first AI security operations platform integrating multi-LLM support, a dynamic red/blue team skill pipeline, MCP tool orchestration, Docker terminal acce…
AgentRunKit
Lightweight Swift 6 framework for building LLM-powered agents — cloud + on-device inference via MLX on Apple Silicon
Llcat
cat for LLMs
Git Courer
MCP server for Git with local Ollama — zero tokens for git operations
Deskdrop
Android keyboard with local AI (Ollama, Whisper, MCP) or cloud (Gemini, Groq, OpenAI)
Simple LLM CLI
Easily pipe anything into an LLM
Awesome Hermes Usecases
Curated real-world use cases for Hermes Agent — the self-improving AI agent from Nous Research. Backed by primary sources.
Lilbee
Run local AI models, search your files and code, and crawl the web, all in one program. Cited answers, local-first, with an MCP server for your coding agent. TUI, CLI, REST API, a…
Brow
Experimental Chrome side-panel browser agent with LangGraph.js, MCP/WebMCP, local models, and semantic browser automation.
Opencode LLM Proxy
Local AI gateway for OpenCode with tool/function calling — use any model via OpenAI, Anthropic, or Gemini API
Codex Universal Proxy
Use Ollama,OpenRouter, Gemini, cohere and moonshot ai with Codex plugins, MCP tools, tool_search and apply_patch. Use any plugin you want with any OpenAI compatible api.
Ollama Intern
MCP control plane for local cognitive labor — job-shaped tools with tiered Ollama models (instant/workhorse/deep/embed), server-enforced guardrails, and measured economics so Clau…
Ollama Agent Kit
Minimal Node.js library for building autonomous tool-calling AI agents on local LLMs with Ollama. A unified registry merges local (Zod) tools and MCP tools: define a tool once — y…
Mcpscope
A local-first workbench for developing, inspecting, and benchmarking MCP servers against local (LM Studio, Ollama) or remote (OpenRouter) models — Web UI, CLI, and MCP interface o…
Local AI
Unified MCP server for managing local model runtimes (Ollama, LM Studio, and more): provider-agnostic discovery, lifecycle, hardware-fit, and delegated inference.
Openmanus MCP
FastMCP 3.1 Server + Webapp wrapping OpenManus (FOSS CLI, 100% local LLM when you configure Ollama / LM Studio in OpenManus). Not Manus.im.
Local LLM
MCP server for delegating mechanical tasks to local LLMs via Ollama. Claude does the thinking, your local model does the grunt work.
EchoMuse
A calm voice for busy minds.
Awesome AI Workflows That Works
Practical AI & more workflows for coding, MCP, local AI, productivity, research, media, and ops.
TARILIO
AI-powered Information Retrieval with integrated AI Assistant, MCP Client, Local LLM server
Aero Ref
Multi-server MCP agent: AeroAPI live flight data + BigQuery catalog, orchestrated with LangGraph and Ollama local inference. FastMCP · mcp-use · qwen2.5
Lmsmob Chat
Android chat client for LM Studio local server with OpenAI-compatible API, MCP/server tools, chat history, and LAN support.
Cybersec Toolkit
580+ cybersecurity tools, one command. Modular bash installer for Linux & Termux with 14 profiles, 18 modules, and an MCP server for AI-assisted ethical hacking.
Diffusion LLM MCP
FastMCP fleet MCP server for diffusion LMs (dLLM). DiffusionGemma on Goliath RTX 4090 — batch inference, HLE-shaped reasoning, ~200–400 tok/s. Doc phase; llama-diffusion-cli sidec…