voice-tts语音输入(Whisper ASR)+ 语音输出(Edge TTS)技能,支持 agent 专属音色,可调用 send_voice_reply.mjs 发送 Telegram 语音消息。
Install via ClawdBot CLI:
clawdbot install believe3344/voice-ttsGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Generated Mar 21, 2026
Integrate the skill into customer service bots on platforms like Telegram or WhatsApp to automatically transcribe customer voice inquiries and respond with synthesized voice and text, improving accessibility and efficiency for support teams.
Use the skill in educational apps or Discord communities to convert text-based learning materials into audio lectures, enabling students to listen to lessons and ask questions via voice for interactive learning experiences.
Deploy the skill in healthcare systems to send automated voice and text reminders for appointments via platforms like WhatsApp, enhancing patient engagement and reducing no-show rates in clinics.
Implement the skill in corporate tools like Lark or Discord to transcribe voice messages from employees in meetings and generate voice summaries, streamlining internal communication and documentation processes.
Leverage the skill in assistive applications to convert text from websites or documents into speech using Edge TTS, and allow users to input commands via voice, improving digital accessibility on platforms like Telegram.
Offer the skill as a cloud-based service with tiered pricing based on usage volume (e.g., number of voice transcriptions or TTS requests per month), targeting businesses that need scalable voice processing solutions.
Provide basic voice processing features for free with limited usage, and charge for advanced capabilities like higher accuracy models (e.g., Whisper turbo), custom voices, or priority support to monetize heavy users.
Generate revenue by offering custom integration services to businesses wanting to embed the skill into their existing platforms (e.g., CRM systems or mobile apps), along with ongoing maintenance and support contracts.
💬 Integration Tip
Ensure proper configuration of audio directories and platform-specific parameters like Telegram's asVoice flag to avoid common deployment issues.
Scored Apr 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Search and manage Spotify playlists, tracks, albums, artists, and playback state via the Spotify Web API. Use this skill when users want to search for music,...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Transcribe long audio files safely on 16GB RAM machines using auto-chunking with Whisper’s base model and seamless transcript merging.
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...