xiaoyuzhou-asrTranscribe 小宇宙 (Xiaoyuzhou) podcast episodes to text using local Qwen3-ASR speech recognition. Combines xyz API (小宇宙FM API) to fetch episode metadata and aud...
Install via ClawdBot CLI:
clawdbot install worldwonderer/xiaoyuzhou-asrGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://github.com/ultrazg/xyz.gitAudited May 28, 2026 · audit v1.0
Generated Aug 9, 2026
A content creator transcribes podcast episodes to create blog posts, social media snippets, or newsletters. The transcription allows easy quoting and repurposing of audio content into written formats.
A researcher or journalist transcribes episodes from specific podcasts to analyze themes, quotes, or trends. Batch transcription enables quick scanning of multiple episodes for relevant content.
Podcast platforms or individual creators generate transcripts to make content accessible to deaf or hard-of-hearing audiences. This skill provides a local, cost-effective solution for producing text versions of episodes.
Language learners use transcripts to follow along with native audio, improving listening and reading skills. The transcripts can be used as study materials, and the local processing ensures privacy.
Podcast aggregators or platforms build searchable databases of podcast content by transcribing episodes. This enables users to find specific topics or quotes within podcasts.
Offer basic transcription for free with limited features, and charge subscription fees for advanced options like batch processing, multiple format exports, or higher priority processing.
License generated transcripts to content platforms or publishers who want to repurpose podcast content into articles or ebooks. Charged per transcript or via a recurring license.
Expose the transcription capability as an API for developers to integrate into their own applications (e.g., podcast apps, note-taking tools). Charge per API request or via usage-based pricing.
💬 Integration Tip
Ensure the xyz API server is running and properly configured, and verify all dependencies (ffmpeg, Qwen3-ASR model, qwen3-asr-rs) are installed before use.
Scored May 28, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Search and manage Spotify playlists, tracks, albums, artists, and playback state via the Spotify Web API. Use this skill when users want to search for music,...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Transcribe long audio files safely on 16GB RAM machines using auto-chunking with Whisper’s base model and seamless transcript merging.
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...