Directory

Search MCP Servers

Explore 23,068 servers by name, category, or capability

Showing 1–24 of 6,855 for Speech Mcp

Text To Speech (Windows)

Integrates with Windows speech services to enable text-to-speech and speech-to-text capabilities using native system features and PowerShell commands.

Speech Interface (Faster Whisper)

Integrates voice interaction capabilities using faster-whisper and PyAudio for speech recognition and synthesis, enabling natural language voice interfaces for…

Kokoro TTS

Integrates with the Kokoro TTS engine to provide customizable text-to-speech capabilities, supporting cross-platform audio playback and file output for…

Kokoro Speech

Provides text-to-speech capabilities using the Kokoro TTS model, enabling natural-sounding voice output with customizable playback speed and voice selection…

Chatty MCP

a MCP server enable your AI code editor (e.g., Cursor, Cline) with voice capabilities and voice response summaries

Minimax

Official MiniMax Model Context Protocol (MCP) server that enables interaction with powerful Text to Speech and video generation APIs.

AivisSpeech

Enables AI systems to generate and play speech audio from text input through the AivisSpeech API, with configurable speaker settings for voice output…

Kokoro TTS

Converts text to speech using the Kokoro TTS engine with configurable voices, speeds, and languages, supporting both local storage and S3 cloud integration…

TTS Say

Integrates with OpenAI's API and local sound playback to convert text into audible speech, enabling voice output for various applications.

ElevenLabs Text-to-Speech

Integrates ElevenLabs' text-to-speech capabilities for high-quality, customizable voice output in interactions, featuring voice selection and model choice.

Pixelle Mcp

An omnimodal AIGC framework that seamlessly converts ComfyUI workflows into MCP tools with zero code, enabling full-modal support for Text, Image, Sound, and…

MiniMax-MCP

Official MiniMax Model Context Protocol (MCP) server that enables interaction with powerful Text to Speech and video/image generation APIs. This server allows…

macOS Say

Leverages macOS 'say' command for customizable text-to-speech functionality, enabling dynamic voice output.

Say (Text-to-Speech)

Provides text-to-speech capabilities through both native system voices and ElevenLabs integration, enabling vocalization of responses without leaving the…

Voice Recorder (Whisper)

Integrates with OpenAI's Whisper model to provide voice recording and transcription capabilities for applications requiring speech-to-text functionality.

Gpu Bridge Mcp Server

Access 47 AI models and 30 services via MCP. LLM, image gen, video, speech, embeddings, reranking, PDF parsing & more. Pay-per-use with x402 (USDC) or API key.

TTS MCP

Text-to-Speech protocol server that synthesizes text from LLMs and plays audio natively through the host system's desk speakers.

ClickSend MCP Server

Integrates with ClickSend's API to enable sending SMS messages and initiating Text-to-Speech calls for automated communication workflows.

ElevenLabs

Integrates with ElevenLabs to provide high-quality text-to-speech, voice cloning, and conversational capabilities with customizable voice profiles and audio…

DAISYS

Generate high-quality text-to-speech and text-to-voice outputs using the [DAISYS](https://www.daisys.ai/) platform.

Sapiom

One API key gives agents access to 80+ tools: web search, deep search, browser automation, screenshots, 400+ LLM models, image generation, text-to-speech…