toby-audio-transcribeTranscribe, diarise, translate, post-process, and structure audio/video with AssemblyAI. Use this skill when the user wants AssemblyAI specifically, needs hi...
Install via ClawdBot CLI:
clawdbot install tobeyrebecca/toby-audio-transcribeGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://www.assemblyai.com/docsAudited Apr 16, 2026 · audit v1.0
Generated May 6, 2026
Transcribe team meetings or interviews with speaker labels and optional name mapping. The skill produces agent-friendly Markdown and JSON outputs, enabling downstream AI workflows such as automated meeting summaries or action item extraction.
Translate or transcribe audio/video in multiple languages using AssemblyAI's language detection and translation features. Content creators can generate subtitles, translated transcripts, and structured exports for international audiences.
Analyze customer service calls by transcribing with speaker diarization, then extracting topics, entities, and sentiment. The structured JSON output can feed CRM systems or analytics dashboards to improve service quality.
Convert recorded lectures into searchable, timestamped transcripts with speaker names. Educators can produce Markdown notes for students and NLP-friendly JSON for automated quiz generation or content indexing.
Transcribe legal proceedings with high accuracy, speaker labels, and optional translation. The agent-friendly output formats facilitate integration with legal document management systems for review and discovery.
Offer pay-as-you-go or monthly subscription plans based on audio hours processed, with higher tiers unlocking advanced features like speaker identification, translation, and LLM Gateway extraction.
Package the skill as a composable microservice in AI agent marketplaces (e.g., Clawdbot). Charge per API call or via subscription, enabling other agents to invoke transcription and post-processing on demand.
Provide a white-label transcription solution with dedicated support, custom integrations, and SLAs for enterprises needing high accuracy and compliance. Includes on-premise or VPC deployment options.
💬 Integration Tip
To maximize agent compatibility, always use the --bundle-dir flag to generate manifest files and multiple output formats (Markdown, JSON) for easy downstream consumption.
Scored May 6, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Secure, offline, OpenAI-compatible local Whisper ASR endpoint for OpenClaw. Features faster-whisper (large-v3-turbo), built-in privacy with no cloud telemetr...
Turn your AI into JARVIS. Voice, wit, and personality — the complete package. Humor cranked to maximum.
Text-to-speech conversion using node-edge-tts npm package for generating audio from text. Supports multiple voices, languages, speed adjustment, pitch contro...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。