- Conduid
- Marketplace
- #inference-server
MCP servers tagged inference-server
4 live MCP servers tagged "inference-server", ranked by trust score. Tags come from package metadata and repository topics, so this list covers servers across every category that work with inference-server.
Omlx
LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar
Molla
Local inference server in pure Mojo. Speaks OpenAI, Anthropic, and MCP. Models are OCI artifacts. Same source runs on CPU, NVIDIA, AMD, and Apple GPUs. Apache-2.0 all the way down.
GroundCortex
Continuous LoRA fine-tuning as a local service. Watches source files for changes, retrains on change, and hot-swaps the adapter into a live OpenAI-compatible inference endpoint. P…
Caix
caix — native Apple Core AI inference server for Apple silicon (beta): OpenAI/Anthropic API, dashboard, streaming chat with tools/skills/MCP, MTP speculative decoding.