video-understanding把视频分析为结构化理解索引:场景检测、ASR 转写、逐场景 VLM 观察、静音窗口、融合时间线和写作 brief。 用于理解、索引或总结视频,也作为后续创作前的分析阶段。输入视频文件;输出 scenes.json、 asr_result.json、vlm_analysis.json、silence_periods.json、timeline_fusion.json、agent_narration_brief.md。 触发词:视频理解、视频分析、视频索引、video understanding、analyze video、看懂视频。
Install via ClawdBot CLI:
clawdbot install bill492/video-understandingGrade Good — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://github.com/yt-dlp/yt-dlp/blob/master/supportedsites.mdAudited Apr 16, 2026 · audit v1.0
Generated Mar 1, 2026
Educators and e-learning platforms can use this skill to automatically transcribe and summarize instructional videos, generating structured notes and answering student questions about video content. It supports platforms like YouTube and Loom, making it ideal for online courses and training materials.
Marketing teams can analyze videos from TikTok, Instagram, and Twitter/X to track brand mentions, understand visual trends, and generate summaries for reporting. The skill extracts transcripts and descriptions, aiding in content strategy and competitive analysis.
Support teams can analyze customer-submitted videos from platforms like Loom to quickly transcribe issues, describe visual elements (e.g., UI errors), and answer specific questions, improving response times and accuracy in troubleshooting.
Production studios can use this skill to generate transcripts and visual descriptions for raw footage from various sources, aiding in editing, subtitling, and content indexing. It handles large videos via Gemini's File API, streamlining post-production processes.
Legal firms can analyze video evidence from depositions or surveillance, extracting verbatim transcripts with timestamps and summarizing key events. This assists in case preparation and compliance reporting by automating video content review.
Offer a cloud-based service where users upload video URLs to receive automated analysis reports via API. Charge based on usage tiers (e.g., number of videos processed per month) and include premium features like custom prompts or higher file size limits.
License the skill as part of a larger enterprise software suite for industries like education or media, providing tailored solutions with dedicated support. Revenue comes from one-time licensing fees or annual contracts with customization options.
Provide a free basic version for individual users with limited features (e.g., max video size or analysis frequency) and upsell to a paid plan for advanced capabilities like batch processing, API access, or priority support. Monetize through premium upgrades.
💬 Integration Tip
Ensure yt-dlp and ffmpeg are installed via brew or pip, and set the GEMINI_API_KEY environment variable before use; for YouTube URLs, leverage direct Gemini processing to avoid download delays.
Scored Sep 28, 2026
视频自动剪辑助手。基于 FFmpeg 自动提取精彩片段、生成字幕、裁剪时长、制作短视视频,支持多平台导出。当用户需要:自动剪辑视频、提取关键词片段、生成字幕并烧录、导出短视频、提取视频精华、生成视频摘要时使用此技能。
Perform video editing tasks with ffmpeg, including cutting, merging, converting formats, extracting audio, adding subtitles, resizing, cropping, adjusting sp...
When the user wants to optimize signup, registration, account creation, or trial activation flows. Also use when the user mentions "signup conversions," "reg...
Use when the user has a video + a target-language SRT and wants the video to actually speak that language — generates a time-aligned TTS voice dub. Routes by...
Extract frames or short clips from videos using ffmpeg.
本地视频转文字 - 使用 OpenAI Whisper 进行语音识别,完全免费、离线运行、保护隐私