- Conduid
- Marketplace
- #speech-to-text
MCP servers tagged speech-to-text
19 live MCP servers tagged "speech-to-text", ranked by trust score. Tags come from package metadata and repository topics, so this list covers servers across every category that work with speech-to-text.
Stripe
Official one-stop shop for AI Agents and developers building with Telnyx.
Advanced Homeassistant MCP
An advanced MCP server for Home Assistant. 🔋 Batteries included.
Patter
Open-source voice-AI SDK. The Vapi/Retell alternative for builders who want to own the stack. Give your AI agent a phone number in 4 lines — Python and TypeScript, MIT licensed, T…
Kuon
久远:一个开发中的大模型语音助手,当前关注易用性,简单上手,支持对话选择性记忆和Model Context Protocol (MCP)服务。 KUON:A large language model-based voice assistant under development, currently focused on ease of use and s…
DeLive
System audio capture + multi-provider ASR + local-first AI review workspace. Floating live captions, 6 ASR backends, 60+ languages, AI summary/chat/mindmap, Open API, MCP server,…
Markitdown MCP
📄 Professional MCP server for converting 29+ file formats to Markdown - Perfect for Claude Desktop and AI workflows!
T2t
Voice-to-text with MCP support. System-wide dictation (hold fn) and AI agent mode (hold fn+ctrl) that connects to any MCP server. Cross-platform desktop app with local Whisper tra…
Folio
Local-first meeting notes for macOS. Records system audio + mic, transcribes on-device with Whisper, and writes markdown to your vault — your audio never leaves your Mac.
AI Model Advisor
An intelligent Model Context Protocol (MCP) Server that acts as an AI model advisor. Compare LLM and media pricing, features, and quality across 1,000+ models from OpenRouter, fal…
Voicelayer
Voice I/O layer for AI coding assistants — local TTS, STT, session booking. MCP server.
Trtc AI Build Quickly
Quickly build real-time AI voice conversation apps with TRTC. Supports multi-agent selection, LLM integration, and TTS/STT services.
fasuizu-br/brainiall-mcp-server
speech with multiple voices.
Speech MCP
Fastmcp 3.2 server plus webapp for speech in/out
Elevenlabs CLI
Community-built CLI for the ElevenLabs AI audio platform with TTS, STT, voice tools, and MCP server mode.
eviscerations/whisper-windows-mcp
native local audio and video transcription using whisper.cpp with Vulkan GPU acceleration. No cloud APIs, no Python. Batch processing, multilingual support, model management, and…
mohitbadwal/ringback
hosted pieces (SIP via pjsua2/Linphone + whisper.cpp STT + macOS say TTS). No paid telephony, no extra API key for the conversation.
spokenmd/spoken
Fetch published podcast transcripts as clean Markdown with real speaker names (not "Speaker 1") via the [Spoken](https://spoken.md) API. Search episodes, get transcripts, check cr…
Hearsay
crawl4ai for video & audio — one command turns any YouTube video, podcast, or local recording into clean, timestamped, LLM-ready markdown
ankurmans/pepys-mcp
mcp`. 99+ languages.