glm-v-model智谱 GLM-4V/4.6V 视觉模型调用技能。用于图像/视频理解、多模态对话、图表分析等任务。 当用户提到:图片理解、图像识别、视觉模型、GLM-4V、GLM-4.6V、多模态分析、看图说话、图表分析、视频理解时使用此技能。
Install via ClawdBot CLI:
clawdbot install baokui/glm-v-modelGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://example.com/image.jpgAudited Apr 16, 2026 · audit v1.0
Generated Mar 21, 2026
Automatically describe product images for online listings, extracting key features like color, style, and materials. This helps sellers create detailed product descriptions efficiently, improving searchability and customer engagement on platforms like Amazon or Taobao.
Assist healthcare professionals by analyzing medical images such as X-rays or MRIs to identify anomalies or patterns. This can provide preliminary insights for diagnosis, reducing workload and supporting faster decision-making in clinics or hospitals.
Generate descriptive captions or explanations for educational images and charts in textbooks or online courses. This enhances learning materials by making visual content more accessible, especially for students with visual impairments or in digital learning platforms.
Monitor and analyze images and videos uploaded to social media platforms to detect inappropriate or harmful content. This helps automate moderation processes, ensuring community guidelines are upheld and reducing manual review efforts for platforms like Facebook or TikTok.
Offer the GLM visual models via a subscription-based API, charging users per token or based on usage tiers. This model targets developers and businesses integrating AI vision capabilities into their applications, providing scalable access with pay-as-you-go flexibility.
License the skill to enterprises for embedding into their proprietary software or platforms, such as customer service bots or analytics tools. This generates revenue through one-time licensing fees or ongoing support contracts, catering to industries needing customized AI vision solutions.
Provide basic image analysis for free to attract individual users or small businesses, while charging for advanced features like video understanding, higher resolution support, or priority processing. This model encourages adoption and upsells to premium tiers for enhanced capabilities.
💬 Integration Tip
Ensure API keys are securely stored and images are optimized to under 10MB to avoid processing errors and reduce token costs.
Scored Apr 19, 2026
Use CodexBar CLI local cost usage to summarize per-model usage for Codex or Claude, including the current (most recent) model or a full model breakdown. Trigger when asked for model-level usage/cost data from codexbar, or when you need a scriptable per-model summary from codexbar cost JSON.
Check Antigravity account quotas for Claude and Gemini models. Shows remaining quota and reset times with ban detection.
使用豆包(火山引擎)语音合成大模型 API 将文本转换为语音音频文件。支持声音复刻音色(S_ 开头的音色ID)和官方预置音色。当用户要求"语音合成"、"文字转语音"、"TTS"、"朗读文本"、"生成语音"、"用我的声音读"、"豆包语音"、"声音复刻合成"等相关请求时,务必使用此 skill。即使用户只是说"帮我把...
Intelligent model routing for sub-agent task delegation. Choose the optimal model based on task complexity, cost, and capability requirements. Reduces costs...
自动生成科技新闻摘要。从多个来源(RSS、Twitter、GitHub、Web Search)抓取科技新闻,整合后生成摘要。
Sync OpenRouter models used by OpenClaw into this installation's config. Fetches the OpenClaw app leaderboard from OpenRouter, verifies model IDs against the...