pedestrian-traffic-counting-gemini-video-understandingAnalyze videos with Google Gemini API (summaries, Q&A, transcription with timestamps + visual context, scene/timeline detection, video clipping, FPS control,...
Install via ClawdBot CLI:
clawdbot install lnj22/pedestrian-traffic-counting-gemini-video-understandingGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://www.youtube.com/watch?v=VIDEO_IDAudited Apr 18, 2026 · audit v1.0
Generated May 21, 2026
Automatically upload recorded meetings (MP4/MOV) and get concise summaries with timestamps for key topics, decisions, and action items. This saves hours of manual note-taking and ensures no critical point is missed.
Educators can upload lecture recordings to generate transcripts with visual descriptions, chapter markers, and Q&A capabilities. Students can ask specific timestamp-based questions to clarify concepts.
Compliance officers can analyze uploaded videos for policy violations by querying specific timestamps or detecting scene changes. The system can flag non-compliant segments for review.
Analyze public YouTube videos (e.g., product reviews, tutorials) to extract key points, sentiment, and visual context for competitive intelligence or trend analysis.
Editors and content managers can clip long videos to relevant segments (e.g., highlight reels, scene-based cuts) using timestamp queries and custom FPS sampling to reduce processing costs.
Offer a web-based platform with tiered plans based on number of videos processed, video length, or features (basic summarization vs. full timeline analysis). Free tier with limited monthly minutes.
Provide REST API endpoints for businesses to integrate video understanding into their own applications (e.g., LMS, CRM, video platforms). Charge per minute of video processed.
License the full suite to large enterprises for on-premise or private cloud deployment, with custom model fine-tuning, dedicated SLAs, and bulk video processing tiers.
💬 Integration Tip
Start by using the provided Python snippets for batch processing of local videos via the File API, then scale to YouTube URL analysis for public content monitoring.
Scored May 21, 2026
Use CodexBar CLI local cost usage to summarize per-model usage for Codex or Claude, including the current (most recent) model or a full model breakdown. Trigger when asked for model-level usage/cost data from codexbar, or when you need a scriptable per-model summary from codexbar cost JSON.
Check Antigravity account quotas for Claude and Gemini models. Shows remaining quota and reset times with ban detection.
使用豆包(火山引擎)语音合成大模型 API 将文本转换为语音音频文件。支持声音复刻音色(S_ 开头的音色ID)和官方预置音色。当用户要求"语音合成"、"文字转语音"、"TTS"、"朗读文本"、"生成语音"、"用我的声音读"、"豆包语音"、"声音复刻合成"等相关请求时,务必使用此 skill。即使用户只是说"帮我把...
Intelligent model routing for sub-agent task delegation. Choose the optimal model based on task complexity, cost, and capability requirements. Reduces costs...
自动生成科技新闻摘要。从多个来源(RSS、Twitter、GitHub、Web Search)抓取科技新闻,整合后生成摘要。
让 AI 代理根据对话内容自动选择最合适的模型。四层识别(系统过滤→关键词→指示词→语义相似度),四池架构(高速/智能/人文/代理),五分支路由,全自动 Fallback 回路。支持 trigger_groups_all 非连续词组命中。