glmv-groundingA skill that uses GLM-V native grounding capabilities for coordinate conversion, bounding-box visualization, and more. GLM-V native grounding can locate any...
Install via ClawdBot CLI:
clawdbot install jaredforreal/glmv-groundingGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Potentially destructive shell commands in tool definitions
eval(Calls external URL not in known-safe list
https://github.com/zai-org/GLM-V/tree/main/skills/glmv-groundingAI Analysis
The skill uses a documented, official API (Zhipu AI) for its stated grounding/tracking purpose, with no evidence of credential harvesting, hidden instructions, or obfuscation. The primary risk is external data transmission to a third-party API, which is expected but requires user consent.
Audited Apr 18, 2026 · audit v1.0
Generated Aug 8, 2026
Use GLM-V to automatically detect and localize products on store shelves from images, enabling automated inventory management and out-of-stock alerts. The skill can output 2D bounding boxes for each product, streamlining stock monitoring.
In manufacturing, GLM-V can ground defects or specific components in product images, providing coordinates for precise localization. This helps in automated visual inspection, reducing manual effort and improving consistency.
The skill supports video tracking, allowing security teams to follow persons or vehicles across frames with bounding boxes. This can be used for real-time monitoring and forensic analysis, enhancing situational awareness.
Medical professionals can leverage grounding to locate anatomical structures or abnormalities in X-rays or MRI scans, assisting in diagnosis and surgical planning. The skill outputs precise coordinates for analysis.
For autonomous vehicle development, GLM-V can ground objects like pedestrians, vehicles, and traffic signs in images and videos, generating labeled data for training perception models. This accelerates the annotation process.
Offer the grounding capability as a paid API service, where developers integrate image/video grounding into their applications. Pricing can be based on number of API calls or data volume.
Build specialized software solutions for industries like retail or healthcare, embedding the grounding functionality to solve specific problems. Charge subscription or licensing fees for the software.
Provide automated data annotation services for companies needing labeled visual data. Combine the grounding skill with human validation to ensure quality, charging per annotation or per project.
💬 Integration Tip
Start by setting up the ZHIPU_API_KEY environment variable and install dependencies. Test with simple image grounding before trying video.
Scored Aug 8, 2026
视频自动剪辑助手。基于 FFmpeg 自动提取精彩片段、生成字幕、裁剪时长、制作短视视频,支持多平台导出。当用户需要:自动剪辑视频、提取关键词片段、生成字幕并烧录、导出短视频、提取视频精华、生成视频摘要时使用此技能。
Perform video editing tasks with ffmpeg, including cutting, merging, converting formats, extracting audio, adding subtitles, resizing, cropping, adjusting sp...
When the user wants to optimize signup, registration, account creation, or trial activation flows. Also use when the user mentions "signup conversions," "reg...
Use when the user has a video + a target-language SRT and wants the video to actually speak that language — generates a time-aligned TTS voice dub. Routes by...
Download videos, extract transcripts, capture frames. Analyze YouTube, tutorials, DD videos with yt-dlp + Whisper + ffmpeg.
Extract frames or short clips from videos using ffmpeg.