super-transcribeUnified speech-to-text skill. Use when the user asks to transcribe audio or video, generate subtitles, identify speakers, translate speech, search transcript...
Install via ClawdBot CLI:
clawdbot install theplasmak/super-transcribeGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Potentially destructive shell commands in tool definitions
eval(Accesses system directories or attempts privilege escalation
/proc/Calls external URL not in known-safe list
https://youtube.com/watch?v=...`Uses known external API (expected, informational)
arxiv.orgGenerated Mar 21, 2026
Automatically transcribe podcast episodes for accessibility, SEO, and content repurposing. The skill generates accurate transcripts with speaker diarization, enabling easy editing and creation of show notes or blog posts.
Transcribe business meetings to create searchable records and action items. The diarization feature identifies speakers, while translation supports multilingual teams, improving collaboration and compliance.
Generate subtitles for online courses and video lectures to enhance accessibility and engagement. The skill supports multiple subtitle formats and can translate content for global audiences.
Transcribe large audio/video archives, such as interviews or historical recordings, to enable keyword search and content analysis. The skill handles various formats and languages, making archives more usable.
Transcribe customer service calls to analyze sentiment, identify common issues, and improve training. The skill's accuracy and speed allow for real-time or batch processing of audio data.
Offer a cloud-based transcription service with tiered pricing based on usage, such as hours transcribed per month. Include features like advanced diarization, translation, and API access for enterprise clients.
Provide a free version with basic transcription and limited features, while charging for premium features like high-accuracy backends, bulk processing, or custom integrations. Monetize through in-app purchases or one-time licenses.
Sell on-premise or custom-deployed versions to large organizations, such as media companies or government agencies, with support for compliance, security, and integration into existing workflows.
💬 Integration Tip
Integrate with existing media pipelines by using the skill's command-line interface for batch processing, and consider adding webhooks for real-time transcription in applications like chat platforms.
Scored Jun 19, 2026
AI Analysis
The skill's primary risk is from user-provided URLs (like YouTube) being processed by external tools (ffmpeg, yt-dlp), which is consistent with its transcription purpose. The flagged signals are related to normal system checks (/proc/, eval for environment setup) and expected external API usage, not hidden data exfiltration or credential harvesting.
Audited Apr 16, 2026 · audit v1.0
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Search and manage Spotify playlists, tracks, albums, artists, and playback state via the Spotify Web API. Use this skill when users want to search for music,...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Transcribe long audio files safely on 16GB RAM machines using auto-chunking with Whisper’s base model and seamless transcript merging.
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...