dlazy-qwen-ttsAlibaba Bailian qwen3-tts text-to-speech. Choose from curated system voices (including dialects) or design a custom voice from a natural-language description. 阿里云百炼 qwen3-tts 文本转语音,支持系统音色(含方言)或通过自然语言描述自定义新音色(声音设计)。
Install via ClawdBot CLI:
clawdbot install dlazyai/dlazy-qwen-ttsRequires:
Grade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://github.com/dlazyai/cliAudited Jun 2, 2026 · audit v1.0
Generated Aug 8, 2026
Generate professional voiceovers for audiobooks and podcasts using system voices like 'Ethan' or custom-voice design. The tool supports multiple languages and dialects, making it ideal for multilingual content production.
Create engaging educational content by converting text lessons into natural-sounding speech. Educators can select voices that match their brand's tone or design a friendly, approachable voice from a description.
Produce dynamic voiceovers for promotional videos, advertisements, and explainer content. The ability to design unique voices from natural-language descriptions allows brands to stand out in crowded markets.
Develop applications that convert written content into speech for visually impaired users. The variety of voices and language support ensures inclusivity across different user groups.
Add spoken responses to chatbots, virtual assistants, and IoT devices. Custom voice design enables a unique personality that can be tailored to the product's identity.
Offer TTS services as a cloud-based API or CLI tool on a subscription basis, charging users based on usage volume. Customers pay a monthly fee for access to the API and may have tiered plans depending on the number of characters processed.
Charge developers and businesses per API call or per character of generated speech. This model is attractive for startups or companies with unpredictable usage patterns, as they only pay for what they use.
Leverage the TTS tool to provide custom voiceover creation services for content creators, media companies, and marketing agencies. This could involve a team using the tool to produce high-quality audio for clients, charging a premium for the service.
💬 Integration Tip
Start by using the CLI with the --dry-run flag to understand the payload and cost before invoking the API, and leverage the pipe references to simplify integration with other CLI tools.
Scored Aug 8, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Search and manage Spotify playlists, tracks, albums, artists, and playback state via the Spotify Web API. Use this skill when users want to search for music,...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Transcribe long audio files safely on 16GB RAM machines using auto-chunking with Whisper’s base model and seamless transcript merging.
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...