day253-volcengine-ai-audio-ttsText-to-speech generation on Volcengine (ByteDance) speech services. Use when users need narration, multi-language speech output, voice selection, or TTS tro...
Install via ClawdBot CLI:
clawdbot install day253/day253-volcengine-ai-audio-ttsGrade Limited — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Sends data to undocumented external endpoint (potential exfiltration)
POST → https://openspeech.bytedance.com/api/v1/tts`Calls external URL not in known-safe list
https://www.volcengine.com/docs/6561/79824Audited Apr 18, 2026 · audit v1.0
Generated Mar 21, 2026
Generate voiceovers for online courses and educational videos, supporting multiple languages and customizable voices to enhance engagement. Useful for creating accessible content for diverse learners.
Produce automated voice prompts for interactive voice response systems in call centers, enabling natural-sounding instructions in various languages. Helps improve customer experience with clear, localized audio.
Convert text manuscripts into spoken audio for audiobooks, allowing for voice selection and parameter tuning to match narrative tone. Speeds up production for publishers and independent authors.
Create speech output for screen readers and assistive technologies, helping visually impaired users access digital content. Supports customization of speech rate and pitch for user comfort.
Generate promotional audio for advertisements, podcasts, and social media videos with tailored voices and emotional tones. Enables quick iteration for marketing campaigns across regions.
Offer pay-per-use API access for developers to integrate TTS into applications, charging based on request volume or audio duration. Generates recurring revenue from SaaS subscriptions.
License the TTS technology to enterprises for embedding in their products, such as e-learning platforms or customer service software. Provides upfront licensing fees and ongoing support contracts.
Provide managed audio production services for clients needing bulk narration, like audiobook publishers or video producers, using automated TTS with human oversight. Revenue comes from project-based pricing.
💬 Integration Tip
Ensure environment variables for credentials are set and test with a minimal request to verify cluster and voice_type compatibility before full deployment.
Scored Apr 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Search and manage Spotify playlists, tracks, albums, artists, and playback state via the Spotify Web API. Use this skill when users want to search for music,...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Transcribe long audio files safely on 16GB RAM machines using auto-chunking with Whisper’s base model and seamless transcript merging.
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...