qwen-tts-lan阿里云千问语音合成(TTS)技能,支持将文本转换为自然语音。当用户要求朗读、语音合成、文字转语音、TTS、读一段话、把文字转成声音时使用。支持多种音色(中文/英文/方言),支持流式输出边合成边播放。
Install via ClawdBot CLI:
clawdbot install lanlan314/qwen-tts-lanGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Sends data to undocumented external endpoint (potential exfiltration)
POST → https://dashscope.aliyuncs.com/api/v1/services/aigc/multimodal-generation/generaCalls external URL not in known-safe list
https://dashscope.console.aliyun.comUses known external API (expected, informational)
open.feishu.cnAI Analysis
The skill sends user text to the documented Alibaba DashScope TTS API, which is consistent with its stated purpose of text-to-speech conversion. While it integrates with an external service (Feishu) for optional message delivery, this is disclosed and requires explicit user configuration of credentials.
Generated May 11, 2026
将文字内容如小说、新闻、博客文章转换为自然语音,用于播客、有声读物或视频配音。支持多种音色和指令控制情感,适合内容创作者快速生成高质量音频。
在客服机器人或IVR系统中,将回复文本实时合成为语音,提供自然流畅的交互体验。支持流式输出,可边合成边播放,降低延迟。
为在线课程、培训视频或电子教材生成语音讲解,支持中英文和方言,适合教育机构快速制作多语言版本的教学内容。
帮助视障人士或有阅读困难用户将屏幕文字、文档或网页内容转换为语音,提升信息可访问性。集成简单,可通过脚本调用。
利用声音设计模型(vd)或声音复刻模型(vc)为品牌打造专属音色,用于广告、品牌视频或虚拟助手,增强品牌辨识度。
按实际调用的字符数或请求次数收费,适合初创企业或中小型应用,无需前期投入硬件和维护成本。
面向内容创作者或企业提供月/年订阅套餐,涵盖一定量级的TTS调用和高级音色使用,可附加飞书集成等增值服务。
为品牌或企业提供从零设计或复刻音色的定制服务,收取一次性项目费用,并持续提供API调用支持。
💬 Integration Tip
最少仅需配置DASHSCOPE_API_KEY即可通过curl或shell脚本快速生成语音,支持同步和异步调用,适合快速原型开发。
Scored May 11, 2026
Audited Apr 17, 2026 · audit v1.0
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Search and manage Spotify playlists, tracks, albums, artists, and playback state via the Spotify Web API. Use this skill when users want to search for music,...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Transcribe long audio files safely on 16GB RAM machines using auto-chunking with Whisper’s base model and seamless transcript merging.
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...