sayitText-to-speech using Kokoro local TTS. Use when the user wants to convert text to audio, read aloud, or generate speech.
Install via ClawdBot CLI:
clawdbot install babysor/sayitGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://github.com/nazdridoy/kokoro-ttsAudited Apr 16, 2026 · audit v1.0
Generated Mar 21, 2026
Teachers and content creators can generate audio versions of textbooks, articles, or study materials for students with visual impairments or those who prefer auditory learning. This supports inclusive education by converting written educational resources into accessible speech using various voices and languages.
Authors and publishers can locally convert EPUB or PDF files into audiobooks without relying on cloud services, ensuring privacy and reducing costs. The tool allows splitting books into chapters and exporting in formats like MP3, making it suitable for self-publishing or small-scale production.
Businesses can generate pre-recorded audio messages or instructions in multiple languages for customer service hotlines or automated systems. Using voices like zf_xiaoni for Chinese or jf_alpha for Japanese helps cater to diverse customer bases locally and efficiently.
Developers can integrate this skill into applications that read aloud digital content such as websites, documents, or emails for users with visual impairments. The local processing ensures data privacy and offline functionality, enhancing accessibility in personal or workplace settings.
Independent filmmakers or podcasters can create voice-overs for videos, advertisements, or audio content using customizable voices and speed adjustments. This enables cost-effective production without hiring voice actors, suitable for small studios or hobbyists.
Offer a cloud-based platform with enhanced features like voice blending, advanced editing, and API access for businesses needing scalable TTS solutions. Revenue comes from monthly or annual subscriptions based on usage tiers, targeting enterprises and developers.
Sell perpetual licenses for the software to organizations requiring offline, secure TTS capabilities, such as government agencies or healthcare providers. This model includes support and updates, generating upfront revenue with optional maintenance fees.
Provide a free version with basic voices and features, while monetizing through the sale of premium voice packs, additional languages, or advanced functionalities like speed control. This attracts individual users and upsells to professionals needing specialized options.
💬 Integration Tip
Ensure model files are downloaded and placed in the working directory before use, and consider using the --stream option for real-time playback without file storage.
Scored Jun 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Turn your AI into JARVIS. Voice, wit, and personality — the complete package. Humor cranked to maximum.
Text-to-speech conversion using node-edge-tts npm package for generating audio from text. Supports multiple voices, languages, speed adjustment, pitch contro...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...