qwenspeakText-to-speech generation via Qwen3-TTS over SSH. Preset voices, voice cloning, voice design. Use when the user wants to generate speech audio, clone voices, or work with TTS.
Install via ClawdBot CLI:
clawdbot install psyb0t/qwenspeakGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Accesses sensitive credential files or environment variables
~/.ssh/id_rsaCalls external URL not in known-safe list
https://github.com/psyb0t/docker-qwenspeakUses known external API (expected, informational)
raw.githubusercontent.comAI Analysis
The skill operates over SSH to a user-controlled instance and does not send data to unauthorized external servers. The credential access signal appears to be a false positive from SSH's normal key usage, and external URLs are only for documentation. No hidden instructions or obfuscation were found.
Audited Apr 16, 2026 · audit v1.0
Generated Mar 1, 2026
Generate custom voice prompts for IVR systems or chatbots in multiple languages using preset speakers like Ryan for English and Vivian for Chinese. This ensures consistent brand voice across global customer interactions without hiring voice actors.
Create narrated audio for e-learning modules or audiobooks by cloning a teacher's voice or designing friendly voices via voice-design mode. This allows scalable production of personalized educational materials in various tones and languages.
Produce dynamic voiceovers for commercials or social media ads using emotion tricks with cloned voices or preset speakers like Aiden for sunny American tones. Enables rapid iteration on ad campaigns with tailored emotional delivery.
Convert text documents or web content into speech audio using voice-clone mode to replicate a familiar voice for personalized assistance. Supports multiple languages and emotional styles to enhance user experience.
Design unique character voices for games or animations using voice-design mode with natural language descriptions. Allows creators to prototype and produce diverse vocal performances without studio recordings.
Offer a cloud-based API service where users pay monthly for access to qwenspeak's TTS features, including voice cloning and design. Revenue comes from tiered plans based on usage limits and advanced features like emotion control.
Provide bespoke integration and voice training services for businesses needing branded or cloned voices. Charge project-based fees for setup, customization, and ongoing support, targeting industries like customer service and education.
Operate a platform where users pay per audio file generated, with pricing based on factors like model size or voice mode. Attract indie creators and small businesses with low upfront costs and scalable usage.
💬 Integration Tip
Ensure SSH access and environment variables are configured properly; use YAML templates for batch processing to streamline job submissions and management.
Scored Apr 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Secure, offline, OpenAI-compatible local Whisper ASR endpoint for OpenClaw. Features faster-whisper (large-v3-turbo), built-in privacy with no cloud telemetr...
Turn your AI into JARVIS. Voice, wit, and personality — the complete package. Humor cranked to maximum.
Text-to-speech conversion using node-edge-tts npm package for generating audio from text. Supports multiple voices, languages, speed adjustment, pitch contro...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。