ms-voice-tts使用 edge-tts 生成高质量中文语音消息并发送。当用户要求发语音、语音回复、TTS、文字转语音、语音播报、语音消息时使用。支持多种中文声音(男声/女声/方言),可调节语速音调,适用于飞书/Telegram/Discord 等渠道的语音消息发送。
Install via ClawdBot CLI:
clawdbot install binbin1213/ms-voice-ttsGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Generated Mar 21, 2026
Businesses can use this skill to send automated voice notifications to customers via messaging platforms like Telegram or Discord. For example, sending order confirmation or shipping updates as voice messages, which can be more engaging and accessible than text, especially for users on the go.
Companies can integrate this skill into their internal communication tools like Feishu to send voice reminders for meetings, deadlines, or system alerts. This ensures critical notifications are heard promptly, reducing missed messages in busy work environments.
Content creators or social media managers can generate voice messages for platforms like Discord to announce updates, share news, or engage communities with personalized audio content. The ability to adjust voice styles and parameters allows for creative, brand-consistent messaging.
Organizations can use this skill to provide audio versions of text-based communications, aiding users with visual impairments or literacy challenges. The support for multiple Chinese voices, including dialects, also enables localized outreach in diverse linguistic communities.
Developers can enhance chatbots or gaming bots on platforms like Discord by adding voice response capabilities. For instance, a game bot could announce events or results using dynamic voice messages, making interactions more immersive and fun for users.
Offer this skill as part of a subscription-based service where companies pay a monthly fee to integrate voice TTS into their existing communication tools like Feishu or Slack. Revenue comes from tiered plans based on usage volume, voice options, and support levels.
Provide a freemium API that allows developers to access the TTS functionality programmatically, with free limited usage and paid tiers for higher volumes or advanced features like custom voices. This model targets app builders and tech startups looking to add voice capabilities.
License the skill as a white-label solution to marketing agencies or content creators who can rebrand it for their clients. Revenue is generated through one-time setup fees and ongoing maintenance contracts, catering to businesses that want customized voice messaging without in-house development.
💬 Integration Tip
Ensure the edge-tts dependency is installed globally and test voice generation in a sandbox environment before deploying to production to avoid system conflicts.
Scored Jun 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Search and manage Spotify playlists, tracks, albums, artists, and playback state via the Spotify Web API. Use this skill when users want to search for music,...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Transcribe long audio files safely on 16GB RAM machines using auto-chunking with Whisper’s base model and seamless transcript merging.
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...