image-content-extractor统一图片内容提取技能。智能识别终端/文档/通用模式,自动提取内容生成Markdown。
Install via ClawdBot CLI:
clawdbot install zhaog100/image-content-extractorGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://github.com/zhaog100/openclaw-skillsAudited Apr 16, 2026 · audit v1.0
Generated Mar 21, 2026
Automates the conversion of screenshots from software documentation, tutorials, or API references into structured Markdown. This helps technical writers and developers build and maintain knowledge bases efficiently by extracting text, code blocks, and headings from images.
Processes terminal or command-line interface screenshots shared by users during support tickets. It extracts error messages, commands, and paths, formatting them for clear analysis in ticketing systems, speeding up resolution and documentation.
Assists educators and trainers in converting lecture slides, whiteboard notes, or textbook images into digital notes. It structures content with headings and lists, making it easier to create study materials or online course content from visual sources.
Extracts text from screenshots of regulatory documents, logs, or reports for compliance audits. It ensures accurate digitization and indexing into knowledge bases, aiding legal and financial teams in maintaining organized records.
Converts screenshots of user-generated content, infographics, or promotional materials into text for repurposing. Marketers can quickly generate blog posts, social media captions, or newsletters from visual assets, enhancing content workflows.
Offers a free open-source version under MIT License for personal and small-scale use, with paid annual licenses for commercial enterprises based on team size. This model encourages adoption while generating revenue from businesses needing support and advanced features.
Provides a cloud-based service integrating advanced OCR engines like Baidu or Tencent, as mentioned in future plans. Users pay a monthly or yearly subscription for higher accuracy, faster processing, and API access, targeting medium to large enterprises.
Generates income through professional services for customizing the skill to specific workflows, such as integrating with existing knowledge bases or developing new modes. This includes training, support, and development contracts for tailored solutions.
💬 Integration Tip
Integrate with existing knowledge bases like QMD by using batch processing and auto-indexing features to streamline content updates and maintain consistency across documentation systems.
Scored Apr 19, 2026
Generate images using multiple AI models — Midjourney (via Legnext.ai), Flux, Nano Banana Pro (Gemini), Ideogram, Recraft, and more via fal.ai. Intelligently...
Perform image manipulation tasks like background removal, resizing, format conversion, rounding corners, watermarking, and color adjustments using ImageMagic...
A specialized skill for generating high-quality, consistent children's bilingual picture books using 'banana nano'. Supports 18 visual styles across 4 catego...
从图片中提取表格数据并输出为表格和CSV格式;当用户需要从图片提取表格数据、识别表格内容或导出CSV时使用;触发条件:试验图片表格提取,图片表格提取,试验图片数据提取,提取图片数据,提取试验数据
Generate or edit images with Gemini using the Google GenAI SDK. Use when the user asks to create, transform, render, or save one or more images in an OpenCla...
提供图片生成能力:输入文本描述生成图片,可附加风格提示词。支持 1-4 张图片生成,按图片尺寸对应点数计费。Authorization 使用本系统发放的 Bearer Token。