understand-image-minimax图片理解技能,使用 Minimax Coding Plan VLM API 分析图片
Install via ClawdBot CLI:
clawdbot install xbos1314/understand-image-minimaxRequires:
Grade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://example.com/photo.jpgAudited Apr 17, 2026 · audit v1.0
Generated Mar 22, 2026
Automatically analyze product images to generate detailed descriptions, identify key features, and extract attributes like color, size, and material. This helps in cataloging inventory and improving search functionality on online platforms.
Scan user-uploaded images to detect inappropriate content, verify compliance with platform guidelines, and flag potential violations. This assists in maintaining a safe and respectful online community.
Analyze medical images such as X-rays or scans to assist in identifying anomalies, providing preliminary insights for healthcare professionals. This can speed up diagnosis and support remote consultations.
Process property photos to automatically generate descriptions, identify room types, and highlight features like amenities or architectural details. This streamlines listing creation and enhances buyer engagement.
Use images from textbooks or educational materials to create interactive learning aids, generate quiz questions, or provide visual explanations for complex concepts. This enhances digital learning experiences.
Offer the image analysis skill as a cloud-based service with tiered pricing based on usage volume, such as number of API calls or image processing capacity. This provides recurring revenue and scalability for businesses.
License the skill's underlying technology to other developers or companies for integration into their own applications, charging per API call or through enterprise agreements. This enables broad adoption and customization.
Provide basic image analysis features for free to attract users, with premium features like advanced analytics, higher processing limits, or priority support available for a fee. This drives user acquisition and upsells.
💬 Integration Tip
Ensure the MINIMAX_API_KEY environment variable is set and test with both local and remote image paths to handle different input types seamlessly.
Scored May 14, 2026
Use CodexBar CLI local cost usage to summarize per-model usage for Codex or Claude, including the current (most recent) model or a full model breakdown. Trigger when asked for model-level usage/cost data from codexbar, or when you need a scriptable per-model summary from codexbar cost JSON.
Check Antigravity account quotas for Claude and Gemini models. Shows remaining quota and reset times with ban detection.
使用豆包(火山引擎)语音合成大模型 API 将文本转换为语音音频文件。支持声音复刻音色(S_ 开头的音色ID)和官方预置音色。当用户要求"语音合成"、"文字转语音"、"TTS"、"朗读文本"、"生成语音"、"用我的声音读"、"豆包语音"、"声音复刻合成"等相关请求时,务必使用此 skill。即使用户只是说"帮我把...
Intelligent model routing for sub-agent task delegation. Choose the optimal model based on task complexity, cost, and capability requirements. Reduces costs...
自动生成科技新闻摘要。从多个来源(RSS、Twitter、GitHub、Web Search)抓取科技新闻,整合后生成摘要。
让 AI 代理根据对话内容自动选择最合适的模型。四层识别(系统过滤→关键词→指示词→语义相似度),四池架构(高速/智能/人文/代理),五分支路由,全自动 Fallback 回路。支持 trigger_groups_all 非连续词组命中。