- Conduid
- Marketplace
- #retrieval-augmented-generation
MCP servers tagged retrieval-augmented-generation
34 live MCP servers tagged "retrieval-augmented-generation", ranked by trust score. Tags come from package metadata and repository topics, so this list covers servers across every category that work with retrieval-augmented-generation.
Memoryos
[EMNLP 2025 Oral] MemoryOS is designed to provide a memory operating system for personalized AI agents.
Local Deep Research
Local Deep Research achieves ~95% on SimpleQA benchmark (tested with GPT-4.1-mini). Supports local and cloud LLMs (Ollama, Google, Anthropic, ...). Searches 10+ sources - arXiv, P…
Txtai
💡 All-in-one AI framework for semantic search, LLM orchestration and language model workflows
LEANN
[MLsys2026]: RAG on Everything with LEANN. Enjoy 97% storage savings while running a fast, accurate, and 100% private RAG application on your personal device.
PocketFlow
Pocket Flow: 100-line LLM framework. Let Agents build Agents!
Swirl Search
AI Search & RAG Without Moving Your Data. Get instant answers from your company's knowledge across 100+ apps while keeping data secure. Deploy in minutes, not months.
Haystack
Open-source AI orchestration framework for building context-engineered, production-ready LLM applications. Design modular pipelines and agent workflows with explicit control over…
Llmc
One stop shop - Local-first RAG stack with intelligent polyglot-code/docs, remote code execution, local llama enrichment, progressive disclosure tools, mcp server, sandboxed secur…
RAG MCP
Simple RAG implementation from scratch using MCP, focusing on Perception, Memory, Decision and Action
Redis Vl Python
Redis Vector Library (RedisVL) -- the AI-native Python client for Redis.
Redevops Rag
Hybrid RAG (DuckDB vector + BM25 + RRF + recency/keyword priors + optional cross-encoder rerank) as an installable library + CLI.
PharosRAG
Pharos — local-first agentic RAG for your team's document library: multi-format ingest, hybrid retrieval, enterprise ACL, dual HTTP + MCP exits.
OpenDocuments
Open source RAG tool for AI document search - connect GitHub, Notion, Google Drive and ask questions with cited answers. Self-hosted with Ollama/OpenAI/Claude.
Jurisd
Model Context Protocol server for Australian legislation and case law search.
Reporecall
Local codebase memory for Claude Code and MCP - AST indexing, call graphs, hybrid search. 0 tool calls, 3-8x fewer tokens.
Knowledge Rag
Local RAG System for Claude Code — Hybrid search + Cross-encoder Reranking + Markdown-aware Chunking + 12 MCP Tools. No external servers, pure ONNX in-process.
Pinecone Claude Code Plugin
The official Pinecone marketplace for Claude Code Plugins
Synaptic Memory
Brain-inspired knowledge graph: spreading activation, Hebbian learning, memory consolidation.
TabulaRAG
A fast-ingesting tabular data MCP RAG tool backed with cell citations.
EstateWise Chapel Hill Chatbot
🏠 An AI real estate app featuring secure auth, conversation management, & personalized property recommendations. Powered by Agentic AI, RAG (w/ Pinecone), GraphRAG (w/ Neo4j), MCP…
cdeust/Cortex
coding write gate, WRRF retrieval. PostgreSQL + pgvector, 33 MCP tools, 7 lifecycle hooks. Benchmarked 97.8% R@10 on LongMemEval. `claude plugin marketplace add cdeust/Cortex`
TrueMemory
A living memory system that ingests long-horizon data to infer insights, enabling more decisive action, all while running on a single SQLite file locally.
Gptme Rag
Local RAG as a simple CLI, for standalone use or as a gptme tool
Lilbee
Run local AI models, search your files and code, and crawl the web, all in one program. Cited answers, local-first, with an MCP server for your coding agent. TUI, CLI, REST API, a…
Fortemi
Self-hosted AI knowledge base with hybrid semantic search (pgvector + FTS + RRF), MCP server, multi-provider LLM inference (Ollama, OpenAI, OpenRouter, llama.cpp), multimodal inge…
Rag Eval MCP Server
MCP server exposing RAG evaluation tools (judge, suite, gate)
Zerodb Sequential Thinking
Persistent sequential thinking MCP — chain-of-thought reasoning that survives sessions, resumes across agents, and saves conclusions as plan artifacts. Powered by ZeroDB. Zero-con…
The Pensieve
One simply siphons the excess thoughts from one's mind, pours them into the basin, and examines them at one's leisure. It becomes easier to spot patterns and links, you understand…
Ferrex
Persistent memory for AI agents. Stores facts, events, and procedures with intelligent retrieval, entity resolution, and staleness scoring. Fully local.
Graphrag Code
🕸️ Code knowledge graph for Claude Code & AI coding agents — index TypeScript, NestJS, React into Neo4j and query architecture in Cypher
Mosaic
MCP server for agent memory over HexxlaDB—ring retrieval, embeddings + lexical search, seams, facets, YAML persistence policy, localhost HTTP transport.
Rag Forge
Production-grade RAG pipelines with evaluation baked in
Fetchium
Rust-native universal retrieval layer for humans and AI agents — search, extract, rank, and synthesize the web. CLI + REST API + MCP server, with adaptive extraction (CEP), token-…
andyliszewski/grounding-ai
first, no cloud dependency.