speechUse when the user asks for text-to-speech narration or voiceover, accessibility reads, audio prompts, or batch speech generation via the OpenAI Audio API; ru...
Install via ClawdBot CLI:
clawdbot install patches429/speechGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://platform.openai.com/api-keysAudited Apr 17, 2026 · audit v1.0
Generated Mar 20, 2026
Create voiceovers for online courses, tutorials, or e-learning modules where clear, steady narration enhances learning retention. The batch processing capability allows educators to generate entire lesson series efficiently.
Convert written content like articles, reports, or documentation into spoken audio for visually impaired users or those preferring auditory consumption. The single-file generation is ideal for on-demand accessibility requests.
Generate professional phone system prompts for customer service lines, including hold messages, menu options, and automated responses. Batch processing handles multiple prompt variations needed for complete IVR setups.
Produce engaging narration for software demos, product walkthroughs, or marketing videos where friendly, confident voice guidance improves user engagement. The instruction augmentation ensures appropriate pacing and emphasis.
Convert book chapters or long-form content into consistent audio segments using batch processing with stable voice parameters. The character limit handling allows chunking of longer texts while maintaining vocal consistency.
Offer API-based text-to-speech services to businesses needing regular audio content creation. Charge per audio minute generated with tiered pricing based on volume and voice options. The bundled CLI provides reliable, reproducible outputs for enterprise clients.
Provide audio conversion services to organizations needing WCAG compliance, converting documents, websites, and applications into accessible audio formats. Use batch processing for large-scale projects while maintaining quality standards through validation workflows.
Deliver professional voiceover services for e-learning companies, marketing agencies, and media producers. Leverage the skill's instruction augmentation to quickly adapt to client specifications while using the CLI for consistent, high-quality outputs.
💬 Integration Tip
Set OPENAI_API_KEY as an environment variable before use and utilize the bundled CLI script for consistent results rather than creating custom implementations.
Scored Apr 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Turn your AI into JARVIS. Voice, wit, and personality — the complete package. Humor cranked to maximum.
Text-to-speech conversion using node-edge-tts npm package for generating audio from text. Supports multiple voices, languages, speed adjustment, pitch contro...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...