elevenlabs-transcribeTranscribe audio to text using ElevenLabs Scribe. Supports batch transcription, realtime streaming from URLs, microphone input, and local files.
Install via ClawdBot CLI:
clawdbot install paulasjes/elevenlabs-transcribeRequires:
Grade Good — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://elevenlabs.io/speech-to-textAudited Apr 16, 2026 · audit v1.0
Generated Mar 22, 2026
Transcribe recorded team meetings or conference calls with speaker diarization to identify who said what, enabling automated meeting minutes and action item extraction. Useful for remote teams, corporate compliance, and project management tools.
Stream real-time audio from live events, podcasts, or radio broadcasts to generate instant captions or subtitles, improving accessibility and audience engagement. Supports multiple languages for global content delivery.
Transcribe customer service calls from microphone input or recordings to create searchable logs, analyze sentiment, and train AI models for quality assurance. Helps in compliance monitoring and improving service efficiency.
Convert recorded lectures, seminars, or online courses into text with timestamps for note-taking, study aids, and accessibility for hearing-impaired students. Enables easy content indexing and translation.
Transcribe audio from legal proceedings, interviews, or depositions with high accuracy and speaker identification, producing timestamped transcripts for evidence and record-keeping. Reduces manual transcription costs.
Offer a cloud-based transcription service with tiered pricing based on usage hours, features like diarization or JSON output, and API access. Targets businesses needing scalable, automated transcription without infrastructure setup.
Provide developers with an API to integrate ElevenLabs transcription into their applications, charging per minute of audio processed. Includes support for batch and real-time modes, with custom language and event tagging options.
License the transcription skill to other companies or platforms (e.g., video conferencing tools, content management systems) for embedding into their products under their brand. Includes customization and technical support.
💬 Integration Tip
Set the ELEVENLABS_API_KEY environment variable and ensure ffmpeg is installed; use the --quiet flag in automated scripts to suppress status messages for cleaner output.
Scored May 14, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Search and manage Spotify playlists, tracks, albums, artists, and playback state via the Spotify Web API. Use this skill when users want to search for music,...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Transcribe long audio files safely on 16GB RAM machines using auto-chunking with Whisper’s base model and seamless transcript merging.
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...