ai-podcast-creationCreate AI-powered podcasts with text-to-speech, music, and audio editing. Tools: Kokoro TTS, DIA TTS, Chatterbox, AI music generation, media merger. Capabili...
Install via ClawdBot CLI:
clawdbot install okaris/ai-podcast-creationGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://inference.shAudited Apr 16, 2026 · audit v1.0
Generated Feb 25, 2026
News organizations can use this skill to quickly convert written articles into daily audio briefings or podcasts. It enables multi-voice narration with background music, allowing for engaging audio news content without extensive studio resources.
Educators and e-learning platforms can generate narrated lessons, audiobooks, or discussion-based audio content from textbooks or study materials. The multi-voice capability simulates classroom conversations, making learning more interactive and accessible.
Businesses can create branded podcasts, audio newsletters, or promotional voice-overs for ads and social media. The skill allows for consistent voice branding with customizable tones and background music to enhance engagement and reach auditory audiences.
Media producers can transform scripts or research documents into narrated audio segments for documentaries, interviews, or audio series. It supports natural dialogue generation and audio merging, streamlining post-production workflows.
Companies can automate the creation of training modules, internal announcements, or compliance podcasts. The skill enables multi-voice simulations for role-playing scenarios and adds background music to make content more engaging for employees.
Offer a monthly subscription where users receive automated podcasts or audiobooks generated from their provided content, such as news summaries or educational materials. Revenue comes from tiered plans based on usage volume and features like custom voices.
Provide a free tier with basic text-to-speech and limited voices, while charging for advanced features like multi-voice conversations, AI music generation, and high-quality audio exports. Monetize through premium upgrades and API access for developers.
License the skill to marketing agencies, publishers, or e-learning companies as a white-label tool for their clients. Revenue is generated through one-time setup fees and ongoing support contracts, enabling agencies to offer podcast creation as a service.
💬 Integration Tip
Start with simple narration using Kokoro TTS to test voice quality, then gradually incorporate multi-voice workflows and background music for more complex projects.
Scored Apr 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
ElevenLabs text-to-speech with mac-style say UX.
Transcribe audio via OpenAI Audio Transcriptions API (Whisper).
Text-to-speech conversion using node-edge-tts npm package for generating audio from text. Supports multiple voices, languages, speed adjustment, pitch control, and subtitle generation. Use when: (1) User requests audio/voice output with the "tts" trigger or keyword. (2) Content needs to be spoken rather than read (multitasking, accessibility, driving, cooking). (3) User wants a specific voice, speed, pitch, or format for TTS output.
Local text-to-speech via sherpa-onnx (offline, no cloud)
Start voice calls via the OpenClaw voice-call plugin.