AudioPod AI
Hosted MCP server for AudioPod's audio AI: text-to-speech, voice cloning, music generation, stem and speaker separation, transcription, denoise, and voice…
Directory
Hosted MCP server for AudioPod's audio AI: text-to-speech, voice cloning, music generation, stem and speaker separation, transcription, denoise, and voice…
Text to speech for MCP clients. Reads numbers and dates correctly. Six voices, WAV out, every render watermarked.
Integrates with ElevenLabs to provide high-quality text-to-speech, voice cloning, and conversational capabilities with customizable voice profiles and audio…
Advanced MCP Server with AI-powered Limitless API features: natural time queries, meeting detection, action item extraction, daily summaries, and speaker…
Official SynClub server for AI generation, including text-to-speech, voice cloning, video, and image creation.
A simple text-to-speech server that plays audio from text, supporting multiple voice models.
A server for text-to-speech (TTS) using the VoiceVox engine.
Access Whissle API for speech-to-text, diarization, translation, and text summarization.
A text-to-speech (TTS) server using the VOICEVOX engine. Requires a running VOICEVOX instance and is currently macOS only.
A text-to-speech server for VOICEROID2 via the voiceroid_daemon.
A server for text-to-speech generation using the AivisSpeech engine.
A dictionary server using the Merriam-Webster API to provide definitions, parts of speech, and pronunciations for words.
Connects AI systems to VOICEVOX text-to-speech engine for Japanese voice synthesis, supporting both default transport and Server-Sent Events with configurable…
Generate high-quality text-to-speech and text-to-voice outputs using the [DAISYS](https://www.daisys.ai/) platform.
Bitcoin-powered tools marketplace. Image, text, video, music, speech, 3D, file conversion, SMS — all via Lightning micropayments. No signup required.
Generates text-to-speech audio with automatic playback using the Chatterbox TTS model.
One API key gives agents access to 80+ tools: web search, deep search, browser automation, screenshots, 400+ LLM models, image generation, text-to-speech…
Real-time speech-to-text for AI assistants. Transcribe audio files with production-grade accuracy. Pay per use with USDC via x402 — no API keys needed.
Provides a full suite of AI tools via DeepInfra’s OpenAI-compatible API, including image generation, text processing, embeddings, and speech recognition.
Search, read, and create speech-to-text transcripts on-device with the Whisper Notes Mac app — 100% offline, no open ports.
<h1 align="center">ListenHub MCP Server</h1> Official MCP Server for [ListenHub](https://listenhub.ai/), supporting AI podcast generation (single or…
# Dota 2 MCP Server & Discord Coach Bot This project provides a Dota 2 Game State Integration (GSI) MCP server and a Discord bot that coaches you with…
A iOS/MacOS Swift MCP Client using voice interacting with python MCP servers both natively
Generate images, video, and audio directly in Claude Code, Cursor, Windsurf, or any MCP-compatible AI agent.