Gives Claude Desktop a maid personality with Japanese-accented text-to-speech, an interactive visual avatar with 16+ poses and animations, and speech recognition for voice input. Designed for fun rather than productivity.
全面解析MCP YoutubeMCP Server的核心功能、安装配置和实用案例。作为顶级Model Context Protocol服务器,MCP Youtube能让AI助手访问实时数据、执行操作,为您提供更智能的工作体验和自动化解决方案。
全面解析MCP SayMCP Server的核心功能、安装配置和实用案例。作为顶级Model Context Protocol服务器,MCP Say能让AI助手访问实时数据、执行操作,为您提供更智能的工作体验和自动化解决方案。
全面解析MCP VoiceMCP Server的核心功能、安装配置和实用案例。作为顶级Model Context Protocol服务器,MCP Voice能让AI助手访问实时数据、执行操作,为您提供更智能的工作体验和自动化解决方案。
Enables comprehensive audio file analysis and metadata extraction with specialized game audio development features, supporting batch processing of multiple formats and providing platform-specific optimization recommendations.
Enables batch audio processing and optimization using FFmpeg with preset configurations for game audio, voice processing, and music mastering, including specialized optimization for ElevenLabs AI voice output.
An enhanced server for ElevenLabs that enables high-quality text-to-speech, voice cloning, and multi-speaker dialogue management. It features advanced conversational tools for transcript retrieval, history tracking, and emotional audio synthesis using the v3 model.
MCP server for ElevenReader TTS service. Enables document management, URL/ebook upload, voice configuration, and progress tracking.
A Python MCP server for invoice and receipt processing that uses OCR technology to extract data from PDFs and images, offering AI assistants the ability to process, extract text from, and merge invoice documents.
Enables AI assistants to convert text to high-quality speech audio using MeloTTS. Automatically splits long texts into segments, generates WAV files, and merges them using ffmpeg with support for multiple languages and customizable speech parameters.
A Docker-containerized MCP proxy that provides AI image generation, text generation, vision analysis, and text-to-speech capabilities through REST endpoints using Pollinations AI services. Enables multimodal AI interactions including image creation, transformation, OCR, and audio generation through standard HTTP APIs.
Combines phish.net and phish.in APIs into twelve tools for setlists, songs, jam-charts, reviews, and audio.
全面解析Props Labs MCP ServersMCP Server的核心功能、安装配置和实用案例。作为顶级Model Context Protocol服务器,Props Labs MCP Servers能让AI助手访问实时数据、执行操作,为您提供更智能的工作体验和自动化解决方案。
An MCP server that exposes speech-to-text and text-to-speech capabilities using a local speaches instance, allowing AI assistants to transcribe audio and generate speech.
This service provides fast and reliable transcriptions for audio/video files and voice memos. It allows LLMs to interact with the text content of audio/video file.
Enables video generation from text prompts or images using Google's Veo 3 API. Supports multiple models, audio generation, and various aspect ratios for creating high-quality videos.
Voice Mode for Claude Code
Provides voice notifications using Grok's text-to-speech API to alert users when Claude Code completes tasks, with support for both local and remote server configurations.
Enables users to list and manage Zoom cloud recordings through the Model Context Protocol. It allows for searching recordings by date and retrieving specific meeting details, including download URLs for video, audio, and transcripts.
Enables control of MentraOS smart glasses through tools for managing text displays, audio output, and input transcriptions. It integrates with Mentra Cloud to facilitate seamless interaction with smart glass hardware using natural language commands.



