audio-video-to-text音视频转文字技能,使用 Whisper 进行语音识别。支持多种音视频格式,可输出纯文本、SRT/VTT 字幕或 JSON 格式。适用于会议记录、视频字幕生成、采访整理、播客转录等场景。
Install via ClawdBot CLI:
clawdbot install ivan830826/audio-video-to-textGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://ffmpeg.org/download.htmlAudited Apr 17, 2026 · audit v1.0
Generated Mar 20, 2026
Automatically transcribe recorded meetings into text documents for minutes, action item tracking, and compliance archiving. This reduces manual note-taking effort and ensures accurate records across departments.
Generate SRT or VTT subtitles for online videos, enhancing accessibility and viewer engagement on platforms like YouTube or e-learning courses. Supports multiple languages for global content distribution.
Transcribe interviews or focus group recordings for qualitative research, enabling easy text analysis, coding, and data extraction. Useful in social sciences, market research, and ethnographic studies.
Convert podcast episodes or radio broadcasts into text for show notes, SEO optimization, and content repurposing into articles or social media snippets. Speeds up post-production workflows.
Transcribe client consultations, medical dictations, or legal proceedings into accurate text records for documentation, reference, and compliance with regulatory standards in sensitive fields.
Offer a cloud service with tiered plans based on usage hours or features like batch processing and API access. Targets businesses needing regular transcription services with scalable pricing.
Provide an API for developers to integrate transcription into applications, charging per minute of audio processed. Ideal for startups and tech companies building custom solutions.
Sell on-premise or custom deployments with support, training, and integration services for large organizations in regulated industries like healthcare or finance, ensuring data privacy.
💬 Integration Tip
Integrate via command-line scripts or API calls; ensure ffmpeg is installed for format support and optimize with model selection based on accuracy vs. speed needs.
Scored Apr 19, 2026
视频自动剪辑助手。基于 FFmpeg 自动提取精彩片段、生成字幕、裁剪时长、制作短视视频,支持多平台导出。当用户需要:自动剪辑视频、提取关键词片段、生成字幕并烧录、导出短视频、提取视频精华、生成视频摘要时使用此技能。
Perform video editing tasks with ffmpeg, including cutting, merging, converting formats, extracting audio, adding subtitles, resizing, cropping, adjusting sp...
When the user wants to optimize signup, registration, account creation, or trial activation flows. Also use when the user mentions "signup conversions," "reg...
Use when the user has a video + a target-language SRT and wants the video to actually speak that language — generates a time-aligned TTS voice dub. Routes by...
Download videos, extract transcripts, capture frames. Analyze YouTube, tutorials, DD videos with yt-dlp + Whisper + ffmpeg.
Extract frames or short clips from videos using ffmpeg.