moss-voice-generatorMOSI Studio 指令式音色生成(moss-voice-generator): 用自然语言描述想要的音色风格,无需指定预设 voice_id, 模型根据描述实时生成对应的声音。 触发词:指令式语音、按描述生成声音、自定义音色、描述一个声音、 "voice generator"、"generate voice...
Install via ClawdBot CLI:
clawdbot install mkkb473/moss-voice-generatorGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://studio.mosi.cnAudited Apr 17, 2026 · audit v1.0
Generated Mar 20, 2026
Audiobook platforms can generate unique voice styles for different genres or authors without pre-recording each narrator. For example, a mystery novel could use a 'low, suspenseful male voice with a slow, deliberate pace' to enhance atmosphere, while a children's book might use a 'sweet, playful female voice with animated intonation'.
Marketing agencies can quickly produce tailored voiceovers for ads or social media content based on brand tone. A luxury brand might request a 'sophisticated, calm female voice with authoritative clarity' for a video ad, while a tech startup could use a 'youthful, energetic male voice with upbeat enthusiasm' for a promotional reel.
E-learning platforms can generate diverse instructional voices to match lesson themes, such as a 'patient, gentle female voice for math tutorials' or a 'dramatic, engaging male voice for history storytelling'. This avoids the need for multiple voice actors and allows real-time customization based on student feedback.
Developers can integrate this skill to create AI assistants with user-defined personalities, like a 'friendly, cheerful voice for customer support' or a 'professional, neutral tone for business analytics'. It enables brands to offer unique voice experiences without relying on preset options.
Offer the voice generation capability via an API, charging per request or through subscription tiers. This model targets developers and businesses needing scalable, on-demand voice synthesis without infrastructure investment, with potential upsells for higher usage limits or premium features.
License the technology to audiobook, podcast, or video platforms as an integrated feature. This allows partners to enhance their offerings with customizable voices, generating revenue through licensing fees or revenue-sharing agreements based on user engagement.
Provide a free web-based tool with basic voice generation, monetizing through premium plans for advanced features like higher-quality outputs, faster processing, or commercial usage rights. This attracts individual creators and small businesses, converting them to paid users as needs grow.
💬 Integration Tip
Ensure the MOSI_TTS_API_KEY is set in the environment and use the provided script with clear text and instruction parameters for quick testing.
Scored Apr 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Search and manage Spotify playlists, tracks, albums, artists, and playback state via the Spotify Web API. Use this skill when users want to search for music,...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Transcribe long audio files safely on 16GB RAM machines using auto-chunking with Whisper’s base model and seamless transcript merging.
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...