gemini-ttsCustom TTS using Gemini 2.5 Flash for high-quality, persona-driven voice output.
Install via ClawdBot CLI:
clawdbot install yangzhe1991/gemini-ttsGrade Limited — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://generativelanguage.googleapis.com/v1beta/models/gemini-2.5-flash-previewUses known external API (expected, informational)
googleapis.comAudited Apr 17, 2026 · audit v1.0
Generated Mar 22, 2026
Generates audiobooks with consistent character voices using the 'little-claw-persona' or custom tones. Ideal for indie authors or publishers seeking affordable, high-quality narration without hiring voice actors.
Creates engaging voiceovers for e-learning modules, tutorials, or language apps with a friendly, persona-driven tone. Enhances user retention by making digital lessons more relatable and immersive.
Integrates custom TTS into chatbots or IVR systems to provide natural, brand-aligned voice responses. Reduces reliance on generic robotic voices, improving customer experience in support interactions.
Converts text content like websites, documents, or emails into clear, high-quality audio using customizable voices. Helps organizations comply with accessibility standards while offering a personalized listening experience.
Produces dynamic voice content for commercials, social media videos, or promotional materials with tailored personas. Enables marketers to quickly iterate on audio assets without studio recording costs.
Offers a cloud-based TTS service with tiered pricing based on usage (e.g., characters per month). Targets businesses needing regular audio generation, with premium features like custom voice training.
Provides a pay-per-use API for developers to integrate Gemini TTS into applications. Generates income through API calls, with volume discounts for high-traffic clients like app developers or content platforms.
Licenses the TTS technology to other companies for rebranding in their products, such as e-learning platforms or voice assistant devices. Involves upfront licensing fees and ongoing support contracts.
💬 Integration Tip
Ensure the GEMINI_API_KEY is securely stored in environment variables and test voice outputs with short texts before scaling to avoid API quota issues.
Scored Apr 19, 2026
Use CodexBar CLI local cost usage to summarize per-model usage for Codex or Claude, including the current (most recent) model or a full model breakdown. Trigger when asked for model-level usage/cost data from codexbar, or when you need a scriptable per-model summary from codexbar cost JSON.
Check Antigravity account quotas for Claude and Gemini models. Shows remaining quota and reset times with ban detection.
使用豆包(火山引擎)语音合成大模型 API 将文本转换为语音音频文件。支持声音复刻音色(S_ 开头的音色ID)和官方预置音色。当用户要求"语音合成"、"文字转语音"、"TTS"、"朗读文本"、"生成语音"、"用我的声音读"、"豆包语音"、"声音复刻合成"等相关请求时,务必使用此 skill。即使用户只是说"帮我把...
Intelligent model routing for sub-agent task delegation. Choose the optimal model based on task complexity, cost, and capability requirements. Reduces costs...
自动生成科技新闻摘要。从多个来源(RSS、Twitter、GitHub、Web Search)抓取科技新闻,整合后生成摘要。
让 AI 代理根据对话内容自动选择最合适的模型。四层识别(系统过滤→关键词→指示词→语义相似度),四池架构(高速/智能/人文/代理),五分支路由,全自动 Fallback 回路。支持 trigger_groups_all 非连续词组命中。