glm-understand-image使用 GLM 视觉 MCP 进行图像理解和分析。触发条件:(1) 用户要求分析图片、理解图像、描述图片内容 (2) 需要识别图片中的物体、文字、场景 (3) 使用 GLM 的视觉理解功能
Install via ClawdBot CLI:
clawdbot install thincher/glm-understand-imageGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://www.bigmodel.cn/glm-coding?ic=OOKF4KGGTWAudited Apr 16, 2026 · audit v1.0
Generated Mar 21, 2026
Automatically analyze product images to generate detailed descriptions, identify key features, and extract text like brand names or specifications. This helps in cataloging products, improving searchability, and enhancing customer experience on online platforms.
Analyze screenshots of error messages, logs, or system interfaces to diagnose issues and provide troubleshooting steps. This reduces resolution time for IT support teams and improves user self-service capabilities.
Review images and videos for inappropriate content, verify compliance with guidelines, and detect sensitive information. Useful for social media platforms, online communities, and regulatory environments to maintain safety and standards.
Interpret diagrams, charts, and technical drawings in educational content to generate explanations or summaries. Assists in creating accessible learning materials and automating content analysis for educators and trainers.
Analyze visual advertisements, social media posts, or campaign images to extract trends, assess design effectiveness, and generate insights for optimization. Helps marketers refine strategies based on visual content performance.
Offer the image analysis capabilities as a pay-per-use API, charging based on the number of requests or data processed. This model targets developers and businesses needing scalable, on-demand visual understanding without infrastructure overhead.
Provide customized solutions with dedicated support, higher usage limits, and integration assistance for large organizations. This model ensures reliability and compliance, appealing to industries like finance or healthcare with strict data requirements.
Offer basic image analysis for free with limited requests, while charging for advanced tools like video analysis, real-time processing, or priority support. This attracts small users and converts them to paid plans as needs grow.
💬 Integration Tip
Ensure API keys are securely stored and test with sample images before full deployment to validate accuracy and performance.
Scored Apr 19, 2026
Generate/edit images with Nano Banana Pro (Gemini 3 Pro Image). Use for image create/modify requests incl. edits. Supports text-to-image + image-to-image; 1K/2K/4K; use --input-image.
Generate images using multiple AI models — Midjourney (via Legnext.ai), Flux, Nano Banana Pro (Gemini), Ideogram, Recraft, and more via fal.ai. Intelligently...
Perform image manipulation tasks like background removal, resizing, format conversion, rounding corners, watermarking, and color adjustments using ImageMagic...
Improves the quality of images, especially screenshots, by enhancing resolution, sharpness, and clarity. Perfect for preparing images for presentations, docu...
A specialized skill for generating high-quality, consistent children's bilingual picture books using 'banana nano'. Supports 18 visual styles across 4 catego...
从图片中提取表格数据并输出为表格和CSV格式;当用户需要从图片提取表格数据、识别表格内容或导出CSV时使用;触发条件:试验图片表格提取,图片表格提取,试验图片数据提取,提取图片数据,提取试验数据