elevenlabs-transcribeTranscribe audio to text using ElevenLabs Scribe. Supports batch transcription, realtime streaming from URLs, microphone input, and local files.
Install via ClawdBot CLI:
clawdbot install paulasjes/elevenlabs-transcribeRequires:
Grade Good — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://elevenlabs.io/speech-to-textAudited Apr 16, 2026 · audit v1.0
Generated Mar 22, 2026
Transcribe recorded team meetings or conference calls with speaker diarization to identify who said what, enabling automated meeting minutes and action item extraction. Useful for remote teams, corporate compliance, and project management tools.
Stream real-time audio from live events, podcasts, or radio broadcasts to generate instant captions or subtitles, improving accessibility and audience engagement. Supports multiple languages for global content delivery.
Transcribe customer service calls from microphone input or recordings to create searchable logs, analyze sentiment, and train AI models for quality assurance. Helps in compliance monitoring and improving service efficiency.
Convert recorded lectures, seminars, or online courses into text with timestamps for note-taking, study aids, and accessibility for hearing-impaired students. Enables easy content indexing and translation.
Transcribe audio from legal proceedings, interviews, or depositions with high accuracy and speaker identification, producing timestamped transcripts for evidence and record-keeping. Reduces manual transcription costs.
Offer a cloud-based transcription service with tiered pricing based on usage hours, features like diarization or JSON output, and API access. Targets businesses needing scalable, automated transcription without infrastructure setup.
Provide developers with an API to integrate ElevenLabs transcription into their applications, charging per minute of audio processed. Includes support for batch and real-time modes, with custom language and event tagging options.
License the transcription skill to other companies or platforms (e.g., video conferencing tools, content management systems) for embedding into their products under their brand. Includes customization and technical support.
💬 Integration Tip
Set the ELEVENLABS_API_KEY environment variable and ensure ffmpeg is installed; use the --quiet flag in automated scripts to suppress status messages for cleaner output.
Scored May 14, 2026
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...
Create, manage, and deploy ElevenLabs conversational AI agents. Use when the user wants to work with voice agents, list their agents, create new ones, or manage agent configurations.
Voice note transcription and archival for OpenClaw agents. Powered by Deepgram Nova-3. Transcribes audio messages, saves both audio files and text transcript...
Text-to-Speech (TTS) and Speech-to-Text (ASR) using coze-coding-dev-sdk. Returns results directly to stdout.