全面解析Sheet Music MCPMCP Server的核心功能、安装配置和实用案例。作为顶级Model Context Protocol服务器,Sheet Music MCP能让AI助手访问实时数据、执行操作,为您提供更智能的工作体验和自动化解决方案。
A FastMCP-based service that extracts no-watermark links from 20+ video platforms (including TikTok, Kuaishou, etc.) and can convert video speech to text.
全面解析Speech Interface (Faster Whisper)MCP Server的核心功能、安装配置和实用案例。作为顶级Model Context Protocol服务器,Speech Interface (Faster Whisper)能让AI助手访问实时数据、执行操作,为您提供更智能的工作体验和自动化解决方案。
Privacy-first audio intelligence: BPM, key, waveform. Audio never stored. Pay per second.
Suno AI music generation with custom lyrics, song extension, cover/remix creation, lyrics generation, and persona management for reusable voice styles.
Integrates Claude Desktop with Super Singularity's course creation API, enabling creation and management of courses with multiple card types (content, quiz, poll, form, video, audio, link), ElevenLabs text-to-speech generation, and Azure Blob Storage for audio hosting.
全面解析Systemprompt MCP InterviewMCP Server的核心功能、安装配置和实用案例。作为顶级Model Context Protocol服务器,Systemprompt MCP Interview能让AI助手访问实时数据、执行操作,为您提供更智能的工作体验和自动化解决方案。
Enables LLM applications to make voice calls and send SMS messages through the Vonage API, allowing AI assistants to perform real-world telephony operations with support for speech recognition and customizable voice parameters.
The Tencent RTC MCP Server provides real-time communication capabilities such as audio and video chat. By integrating this MCP Server into your application or project, you can easily implement secure, high-quality real-time voice and video communication.
An MCP server that analyzes your unique Twitter voice to generate, manage, and post AI-powered tweets and quote tweet drafts. It supports multiple AI providers and provides tools for draft management, voice profiling, and automated content creation from images.
Enables seamless integration with Typecast API through the Model Context Protocol, allowing clients to manage voices, convert text to speech, and play audio in a standardized way.
Enables LLMs to compose and play multi-track MIDI music through natural language prompts. Supports outputting to software or hardware synthesizers for enhanced audio quality.
A Model Context Protocol server that enables AI assistants to perform comprehensive video and audio editing operations including trimming, effects, overlays, audio processing, and YouTube downloads.
An MCP server that turns Claude into a hands-on video editor for short-form videos, enabling music generation, script writing, voiceover synthesis, and video stitching with FFmpeg. It also features a text-to-documentary skill that converts long-form text into structured documentary videos.
全面解析VideolingoMCP Server的核心功能、安装配置和实用案例。作为顶级Model Context Protocol服务器,Videolingo能让AI助手访问实时数据、执行操作,为您提供更智能的工作体验和自动化解决方案。
A Model Context Protocol (MCP) server that provides comprehensive video tools: transcript retrieval, video downloading, and automatic subtitle generation using AI speech-to-text. Works with YouTube, Bilibili, Vimeo, and any platform supported by yt-dlp.
vidlizer pulls frames out of any video, image, or PDF using ffmpeg, sends them to a vision LLM, and returns a flow array — one entry per scene. Each entry tells you what happened, who was on screen, what text was visible, and what changed. If the video has audio, it transcribes it with Apple MLX Whisper and merges the speech into each step.
Enables AI-powered translation of YouTube videos into localized versions with synthesized voiceovers and avatar videos. Supports the full content pipeline from transcript extraction and translation to video generation and publishing across social platforms.
Enables AI agents to generate speech, transcribe audio, and manage voices via the Vocea API.
VocoType 是一款运行在本地端侧的隐私安全语音输入工具,通过快捷键即可将语音实时转换为文字并自动输入到当前应用。支持语音转文字MCP、AI 优化文本、自定义替换词典、录音视频转文字等功能,让语音输入更高效、更安全。



