elevenlabs-stt-openclawTranscribe audio files with ElevenLabs Speech-to-Text (Scribe v2) from the local CLI. Supports diarization, events, JSON output, webhooks, and advanced STT o...
Install via ClawdBot CLI:
clawdbot install xhunx/elevenlabs-stt-openclawGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://example.com/audio.mp3Audited Apr 17, 2026 · audit v1.0
Generated Mar 21, 2026
Law firms can use this skill to transcribe recorded depositions and court proceedings with speaker diarization to identify who said what. The JSON output format makes it easy to integrate transcriptions into case management systems for legal review and evidence organization.
Healthcare providers can transcribe patient consultations and medical interviews to create accurate medical records. The diarization feature helps distinguish between doctor and patient speech, while webhook integration allows automatic updating of electronic health records after transcription completion.
Market research firms can process recorded focus group discussions and customer interviews. The transcription with speaker identification helps analyze different participant perspectives, and the JSON output enables easy data extraction for sentiment analysis and thematic coding in research reports.
Businesses can automatically transcribe team meetings, board discussions, and strategy sessions. The skill creates searchable transcripts that can be distributed as meeting minutes, with webhook functionality to trigger follow-up task assignments in project management tools.
Content creators and podcast producers can transcribe episodes for closed captions, show notes, and content repurposing. The real-time streaming capability allows for live transcription during recordings, while the JSON output facilitates integration with content management systems and SEO optimization tools.
Offer professional transcription services to businesses that need audio converted to text but lack in-house capabilities. Charge per minute of audio processed, with premium pricing for features like diarization, JSON formatting, and webhook integrations for enterprise clients.
Build middleware that connects this transcription skill to popular business software like CRM systems, telehealth platforms, or learning management systems. Charge integration fees and monthly subscription costs for businesses that want automated transcription workflows within their existing tools.
Create tailored transcription solutions for specific industries like legal, medical, or education. Bundle the transcription skill with industry-specific templates, compliance features, and custom integrations, then sell complete packages to organizations in those verticals.
💬 Integration Tip
Set up webhooks to automatically trigger downstream processes when transcription completes, and use the JSON output format for easy parsing in your applications rather than plain text.
Scored Apr 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Search and manage Spotify playlists, tracks, albums, artists, and playback state via the Spotify Web API. Use this skill when users want to search for music,...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Transcribe long audio files safely on 16GB RAM machines using auto-chunking with Whisper’s base model and seamless transcript merging.
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...