ollama-vision本地调用 Ollama qwen3-vl:4b 模型自动压缩并分析图片,支持描述、OCR 文字提取和自定义信息抽取。
Install via ClawdBot CLI:
clawdbot install lzm2023/ollama-visionGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
http://localhost:11434/api/generateAudited Apr 17, 2026 · audit v1.0
Generated Mar 22, 2026
Businesses can use this skill to automatically extract text from scanned documents, receipts, or handwritten notes, converting them into editable digital formats for easy storage and retrieval. It supports OCR mode to handle various document types without manual data entry.
Online retailers can analyze product images to generate detailed descriptions, extract key features like brand names or prices via OCR, or customize extractions for inventory management. This helps in automating catalog updates and improving searchability.
This skill can describe images in detail or extract text from them, providing audio or text descriptions to assist visually impaired individuals in understanding visual content on websites or in apps. It enhances digital accessibility and inclusivity.
Educators and students can use it to analyze diagrams, charts, or textbook images, extracting information or generating explanations to aid in learning. Custom prompts can focus on specific data points, making it useful for research and study materials.
Security teams can apply this skill to analyze surveillance footage or photos, extracting text from license plates or signs, or describing scenes to identify anomalies. It supports real-time or post-event analysis for enhanced safety measures.
Offer this skill as a cloud-based service where users pay a monthly fee for API access to image analysis features, including OCR and custom extractions. It can target small businesses needing affordable automation tools.
Sell customized licenses to large organizations for integrating the skill into their internal systems, such as document management or customer service platforms, with premium support and scalability options.
Provide basic image analysis features for free to attract individual users or small teams, then charge for advanced modes like custom extraction, higher usage limits, or priority processing to generate revenue from power users.
💬 Integration Tip
Ensure Ollama is running locally and the model is downloaded before deployment; use the automatic compression feature to handle large images efficiently.
Scored Apr 19, 2026
Use CodexBar CLI local cost usage to summarize per-model usage for Codex or Claude, including the current (most recent) model or a full model breakdown. Trigger when asked for model-level usage/cost data from codexbar, or when you need a scriptable per-model summary from codexbar cost JSON.
Check Antigravity account quotas for Claude and Gemini models. Shows remaining quota and reset times with ban detection.
使用豆包(火山引擎)语音合成大模型 API 将文本转换为语音音频文件。支持声音复刻音色(S_ 开头的音色ID)和官方预置音色。当用户要求"语音合成"、"文字转语音"、"TTS"、"朗读文本"、"生成语音"、"用我的声音读"、"豆包语音"、"声音复刻合成"等相关请求时,务必使用此 skill。即使用户只是说"帮我把...
Intelligent model routing for sub-agent task delegation. Choose the optimal model based on task complexity, cost, and capability requirements. Reduces costs...
自动生成科技新闻摘要。从多个来源(RSS、Twitter、GitHub、Web Search)抓取科技新闻,整合后生成摘要。
Sync OpenRouter models used by OpenClaw into this installation's config. Fetches the OpenClaw app leaderboard from OpenRouter, verifies model IDs against the...