Text To Speech (Windows)
Integrates with Windows speech services to enable text-to-speech and speech-to-text capabilities using native system features and PowerShell commands.
Directory
Integrates with Windows speech services to enable text-to-speech and speech-to-text capabilities using native system features and PowerShell commands.
MCPServer is a Python-based server that leverages Alibaba's FunASR library to provide speech processing services through the FastMCP framework.
Integrates voice interaction capabilities using faster-whisper and PyAudio for speech recognition and synthesis, enabling natural language voice interfaces for…
Integrates with the Kokoro TTS engine to provide customizable text-to-speech capabilities, supporting cross-platform audio playback and file output for…
Provides text-to-speech capabilities using the Kokoro TTS model, enabling natural-sounding voice output with customizable playback speed and voice selection…
a MCP server enable your AI code editor (e.g., Cursor, Cline) with voice capabilities and voice response summaries
Official MiniMax Model Context Protocol (MCP) server that enables interaction with powerful Text to Speech and video generation APIs.
Enables AI systems to generate and play speech audio from text input through the AivisSpeech API, with configurable speaker settings for voice output…
Converts text to speech using the Kokoro TTS engine with configurable voices, speeds, and languages, supporting both local storage and S3 cloud integration…
Integrates with OpenAI's API and local sound playback to convert text into audible speech, enabling voice output for various applications.
Integrates ElevenLabs' text-to-speech capabilities for high-quality, customizable voice output in interactions, featuring voice selection and model choice.
An omnimodal AIGC framework that seamlessly converts ComfyUI workflows into MCP tools with zero code, enabling full-modal support for Text, Image, Sound, and…
Official MiniMax Model Context Protocol (MCP) server that enables interaction with powerful Text to Speech and video/image generation APIs. This server allows…
Leverages macOS 'say' command for customizable text-to-speech functionality, enabling dynamic voice output.
Provides text-to-speech capabilities through both native system voices and ElevenLabs integration, enabling vocalization of responses without leaving the…
Integrates with OpenAI's Whisper model to provide voice recording and transcription capabilities for applications requiring speech-to-text functionality.
Provides speech-to-text transcription capabilities using OpenAI's Whisper API with configurable language settings and optional file saving
Access 47 AI models and 30 services via MCP. LLM, image gen, video, speech, embeddings, reranking, PDF parsing & more. Pay-per-use with x402 (USDC) or API key.
A JavaScript/TypeScript server for MiniMax MCP, offering image/video generation, text-to-speech, and voice cloning.
Text-to-Speech protocol server that synthesizes text from LLMs and plays audio natively through the host system's desk speakers.
Integrates with ClickSend's API to enable sending SMS messages and initiating Text-to-Speech calls for automated communication workflows.
Integrates with ElevenLabs to provide high-quality text-to-speech, voice cloning, and conversational capabilities with customizable voice profiles and audio…
Generate high-quality text-to-speech and text-to-voice outputs using the [DAISYS](https://www.daisys.ai/) platform.
One API key gives agents access to 80+ tools: web search, deep search, browser automation, screenshots, 400+ LLM models, image generation, text-to-speech…