video-sttExtract audio from video URLs and transcribe using STT (Speech-to-Text). Supports local Whisper or cloud APIs. Use when: user provides a video URL and wants...
Install via ClawdBot CLI:
clawdbot install damiencronw/video-sttGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://youtube.com/watch?v=xxxAudited Apr 17, 2026 · audit v1.0
Generated Mar 21, 2026
YouTubers and podcasters can transcribe videos for subtitles, improving accessibility and SEO. This enables easy translation into multiple languages for global audience reach.
Educators and e-learning platforms can transcribe lecture videos to create study notes, transcripts, and captions. This supports students with disabilities and enhances learning materials.
Marketing teams can transcribe competitor videos, webinars, or interviews to analyze trends, keywords, and sentiment. This aids in content strategy and competitive intelligence.
Law firms and compliance departments can transcribe video evidence, depositions, or training sessions for accurate records. This ensures documentation for audits and legal proceedings.
Film and TV studios can generate subtitles and scripts from raw footage, streamlining editing workflows. This reduces manual transcription costs and speeds up production timelines.
Offer free basic transcription with limited features, then charge for advanced models, higher accuracy, or bulk processing. Target small creators and scale to enterprise clients with premium plans.
Provide this skill as an API for developers to integrate into their apps, such as video platforms or CMS. Charge based on API calls, data volume, or custom model deployments.
Sell customized packages to businesses in education, legal, or media for internal use, with added features like security, compliance, and dedicated support. Focus on high-volume, long-term contracts.
💬 Integration Tip
Integrate with video hosting platforms via webhooks for automated transcription, and use environment variables for secure API key management.
Scored Jun 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Turn your AI into JARVIS. Voice, wit, and personality — the complete package. Humor cranked to maximum.
Text-to-speech conversion using node-edge-tts npm package for generating audio from text. Supports multiple voices, languages, speed adjustment, pitch contro...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...