Audio and voice
40 servers. Where no tool list is shown, it is undisclosed rather than empty.
40 servers
-
Audio Analyzer
Gives LLMs ears. Spectral, harmonic, rhythm, stereo, and structural audio analysis.
-
Audio File MCP App
Inspect local audio files, playback, metadata, loudness, spectrogram. npx -y @counterpoint-studio/audio-file-mcp-app
-
Cochlea
Render, analyze, and verify audio (WAV/FLAC/mp3/ogg) offline and deterministically via MCP tools.
-
Developers
Gaudio Lab Audio AI, Stem Separation, DME Separation, AI Text Sync. npx -y @gaudiolab/mcp-developers
-
ElevenLabs
ElevenLabs MCP server: TTS, music, sound effects, voices, and audio transcription. npx -y @mindstone/mcp-server-elevenlabs
-
Elevenlabs
RunAPI MCP server for ElevenLabs: create tasks, poll status, check pricing. npx -y @runapi.ai/elevenlabs-mcp
-
Elevenlabs
RunAPI MCP server for ElevenLabs: create tasks, poll status, check pricing. npx -y @runapi.ai/elevenlabs-mcp
-
Elevenlabs
Convert text to speech, transcribe audio, and dub videos with AI voices. uvx mcparmory-elevenlabs
-
Fish
MCP server exposing the AceDataCloud Fish Audio API (text-to-speech with voice conditioning). uvx mcp-fish
-
Gandr
Text to speech for MCP clients. 23 languages, six voices, every render watermarked. uvx gandr-mcp
-
Gemini Audio
High-performance audio, music, and voice generation MCP server for Gemini 2.5 and Lyria 3. docker run -i --rm ghcr.io/jxoesneon/gemini-audio-mcp:0.1.0
-
Gemini Tts
RunAPI MCP server for Gemini TTS: create tasks, poll status, check pricing. npx -y @runapi.ai/gemini-tts-mcp
-
Jellypod
Create, import, and publish Jellypod podcast episodes from your AI assistant.
-
Kokoro Tts
Local Kokoro-82M TTS MCP server that synthesizes and plays speech on your machine. uvx mcp-kokoro-tts
-
Kooma, Bambara AI
Bambara AI over MCP: text-to-speech, transcription and translation (Bamanankan + more).
-
Kurdish Tts Stt
Kurdish (Sorani & Kurmanji) text-to-speech & speech-to-text, 664 AI voices. API key required.
-
Labs
MCP Server with AI-powered tools including ElevenLabs text-to-speech for realistic voice synthesis. uvx labs-mcp-server
-
Listen
Give your AI agents the ability to listen. Microphone capture and speech-to-text. npx -y mcp-listen
-
Listen
Give your AI agents the ability to listen. Microphone capture and speech-to-text. npx -y mcp-listen
-
Local Voice
Give your MCP clients the ability to speak by running local voice models using Chatterbox TTS. npx -y @codecraftersllc/local-voice-mcp
-
Orcadub
AI video dubbing via OrcaRouter (model orca/dub): upload a video or URL, poll, download the MP4. npx -y @orcadub/cli
-
Salutespeech
MCP server for Sber SaluteSpeech API, speech recognition and synthesis. npx -y @theyahia/salutespeech-mcp
-
Sleeper Hit Studio
Make podcasts, video shows, audio drama, and documentaries just by chatting. Script to episode.
-
Supertone TTS
Composable Supertone TTS toolkit: synthesis, voice search/preview/clone, usage, 31 languages. uvx supertone-mcp
-
Talkies
Self-hosted MCP server for speech: ASR transcription, TTS synthesis, and file staging tools. docker run -i --rm docker.io/psyb0t/talkies:v0.13.3
-
Text to Speech
Reads text aloud locally on Windows, macOS, and Linux. No API key, account, or cloud service. uvx text-to-speech-mcp
-
Text to Speech API
Convert text to MP3 speech audio in 50+ languages. x402 micropayment.
-
Texttospeech Mcp
MCP server for Text-to-Speech.
-
Tts
Hosted pay-per-use TTS: 54 neural voices, 9 languages incl. Brazilian Portuguese. $10 free credits.
-
UnlimitedTTS
Quote-first, non-custodial x402 text-to-speech with spend policy, MP3 artifacts, and receipts. npx -y @unlimitedtts/mcp
-
Viberadio Fm
Turns an agent's output into audio: narration, podcast, or music. npx -y viberadio-fm mcp
-
Voice Audio
Voice Audio MCP server. text to speech, list voices, transcribe. Built by MEOK AI Labs. python voice-audio-mcp
-
Voice Bridge
Multi-engine TTS for AI coding assistants. 5 engines, free engine included. npx -y ai-voice-bridge
-
VoiceLabs
AI voice generation: text-to-speech and voice cloning from any MCP client.
-
Voicemode
Natural voice conversations for AI assistants - STT/TTS via MCP. uvx voice-mode
-
Vox
Native macOS MCP server for voice I/O, Swift binary, SFSpeechRecognizer + ElevenLabs TTS.
-
Yandex Speechkit
MCP server for Yandex SpeechKit API, speech recognition and synthesis. npx -y @theyahia/yandex-speechkit-mcp
-
cast0 - Turn Text into a Published Podcast
Send text, get a podcast episode. cast0 converts it to audio using TTS (text-to-podcast), publishes it to an automatically hosted RSS feed, and delivers it to Apple Podcasts, Spotify, Overcast, or.
-
three.ws Audio
Text-to-speech, speech-to-text, audio-to-face lipsync, and motion-capture clips for 3D agents. npx -y @three-ws/audio-mcp
-
voiceover
Generate highly realistic Text to Speech voiceovers.
Nothing matches that. Try a shorter search.