minimax-image-understanding使用多模态大模型理解图片内容,生成业务含义描述。支持多种模型:(1) MiniMax VLM (2) OpenAI GPT-4V (3) Claude Vision。用于理解截图、图表、文档照片等,生成精准的文字描述。
Install via ClawdBot CLI:
clawdbot install aidescend/minimax-image-understandingGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://api.minimaxi.comUses known external API (expected, informational)
api.anthropic.comAudited Apr 17, 2026 · audit v1.0
Generated Mar 20, 2026
Analyze product images from online stores to generate detailed descriptions for listings, including features, materials, and usage scenarios. This helps improve SEO and customer engagement by providing accurate, context-rich content.
Process medical scans like X-rays or MRIs to generate descriptive reports summarizing findings, such as anomalies or conditions, aiding healthcare professionals in documentation and preliminary analysis for faster decision-making.
Interpret screenshots of stock charts, graphs, or financial reports to extract trends, key data points, and insights, enabling analysts to quickly summarize market conditions or performance metrics for reports.
Analyze diagrams, illustrations, or textbook images to generate explanatory descriptions for learning materials, helping educators create accessible content for students with visual or textual summaries.
Review user-uploaded images on platforms to detect inappropriate content or generate captions for accessibility, enhancing user safety and compliance with community guidelines through automated analysis.
Offer the image understanding capability as a cloud-based API, charging per request or through subscription tiers. This model targets developers and businesses needing scalable, on-demand image analysis without infrastructure management.
Provide customized solutions for large organizations, such as healthcare or finance firms, with dedicated support, integration services, and compliance features. This includes one-time licensing fees and ongoing maintenance contracts.
Offer a free tier with limited usage or basic models, while charging for advanced features like higher accuracy models, batch processing, or priority support. This attracts small users and converts them to paid plans as needs grow.
💬 Integration Tip
Ensure API keys are securely stored as environment variables and test with sample images to verify model compatibility before full deployment.
Scored Apr 19, 2026
Use CodexBar CLI local cost usage to summarize per-model usage for Codex or Claude, including the current (most recent) model or a full model breakdown. Trigger when asked for model-level usage/cost data from codexbar, or when you need a scriptable per-model summary from codexbar cost JSON.
Check Antigravity account quotas for Claude and Gemini models. Shows remaining quota and reset times with ban detection.
Safely rotate the OpenRouter API key across all config files in an OpenClaw installation. Finds every location where the key is stored, updates them, restart...
Memory graph engine with caller-provided embed and LLM callbacks; core is pure, with real-time correction flow and optional OpenAI integration.
使用豆包(火山引擎)语音合成大模型 API 将文本转换为语音音频文件。支持声音复刻音色(S_ 开头的音色ID)和官方预置音色。当用户要求"语音合成"、"文字转语音"、"TTS"、"朗读文本"、"生成语音"、"用我的声音读"、"豆包语音"、"声音复刻合成"等相关请求时,务必使用此 skill。即使用户只是说"帮我把...
Intelligent model routing for sub-agent task delegation. Choose the optimal model based on task complexity, cost, and capability requirements. Reduces costs...