podcast-transcribeDownload podcast audio from RSS feeds and transcribe to text using AuralWise API. This skill should be used when the user wants to download podcast episodes, convert podcast audio to text transcripts, or batch-process a podcast library for searchable content. Triggers include downloading podcasts, podcast transcription, audio-to-text conversion, RSS feed downloading, or any request involving podcast audio acquisition and speech-to-text conversion. Covers RSS feed discovery, audio downloading, AuralWise API transcription, and AI-generated content overviews with book references, key concepts, and searchable keywords.
Install via ClawdBot CLI:
clawdbot install dairui1/podcast-transcribeGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://storage.googleapis.com/eleven-public-cdn/audio/marketing/nicole.mp3Uses known external API (expected, informational)
googleapis.comAudited Apr 17, 2026 · audit v1.0
Generated Mar 22, 2026
Podcast creators can use this skill to generate transcripts and subtitles from episode audio files or URLs, enhancing accessibility and SEO. It supports multiple transcription engines, including offline options for Apple Silicon, and includes cleanup for accuracy with episode context.
Researchers can transcribe interviews or lectures from public audio sources, producing SRT and TXT files for analysis. The cleanup feature helps correct ASR errors using contextual information from related web pages, improving data quality.
Content agencies can generate subtitles in SRT format from audio files for video localization projects. The skill handles various input types and offers cleanup to ensure proper nouns are accurately transcribed, supporting multilingual workflows.
Organizations can provide transcripts for audio content to meet accessibility standards, using this skill to process podcast episodes or public audio files. It outputs multiple artifact types and allows conservative cleanup without altering original transcripts.
Offer a cloud-based transcription service with API access, charging per minute of audio processed. Include features like multi-engine support and cleanup with episode context, targeting podcasters and media companies.
Provide a free version for basic transcription with limited features, and a paid tier for advanced options like offline MLX-Whisper engine and priority support. Monetize through upgrades and enterprise licenses.
License this skill as part of a larger content management or video editing platform, enabling seamless transcription within existing workflows. Revenue comes from platform subscriptions and add-on fees for transcription services.
💬 Integration Tip
Ensure API keys for hosted transcription engines are configured, and use the default CLI entry points like npx for ease of integration without local installations.
Scored Apr 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Search and manage Spotify playlists, tracks, albums, artists, and playback state via the Spotify Web API. Use this skill when users want to search for music,...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Transcribe long audio files safely on 16GB RAM machines using auto-chunking with Whisper’s base model and seamless transcript merging.
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...