image-readerImage recognition and understanding tool. Uses a multimodal model (e.g. doubao-seed-2.0-pro, kimi-k2.5) to analyze image content and supports OCR text extrac...
Install via ClawdBot CLI:
clawdbot install simonjoe246/image-readerGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://ark.cn-beijing.volces.com/api/coding/v3`Audited Apr 18, 2026 · audit v1.0
Generated Mar 22, 2026
Extract text from scanned documents, receipts, or forms to automate data entry and archiving. Useful for converting paper records into searchable digital formats, reducing manual transcription errors.
Generate detailed descriptions of images for visually impaired users, enabling screen readers to convey visual content. Helps make websites, apps, and digital media more inclusive and compliant with accessibility standards.
Analyze user-uploaded images to detect inappropriate content or extract text for compliance checks. Supports automated filtering in social media, e-commerce, or community platforms to maintain safe environments.
Extract text from product labels, menus, or posters to update inventory systems or provide digital translations. Assists restaurants and stores in managing offerings and improving customer experience through quick information retrieval.
Analyze screenshots of applications to verify text content and visual elements, aiding in quality assurance. Helps developers and testers automate checks for consistency and functionality across different interfaces.
Offer the skill as a cloud-based service with tiered pricing based on usage volume, such as number of images processed per month. Targets businesses needing scalable OCR and image analysis without infrastructure management.
License the underlying technology to other developers or companies for integration into their own products, charging per API call or through annual licenses. Enables customization and embedding in diverse applications.
Provide basic OCR and description features for free with limited usage, while charging for advanced features like batch processing, higher accuracy, or priority support. Attracts individual users and small businesses before upselling.
💬 Integration Tip
Ensure API keys are securely stored in config.yaml and test with sample images to verify compatibility with your system's Python environment.
Scored Apr 19, 2026
Generate images using multiple AI models — Midjourney (via Legnext.ai), Flux, Nano Banana Pro (Gemini), Ideogram, Recraft, and more via fal.ai. Intelligently...
Perform image manipulation tasks like background removal, resizing, format conversion, rounding corners, watermarking, and color adjustments using ImageMagic...
Improves the quality of images, especially screenshots, by enhancing resolution, sharpness, and clarity. Perfect for preparing images for presentations, docu...
A specialized skill for generating high-quality, consistent children's bilingual picture books using 'banana nano'. Supports 18 visual styles across 4 catego...
从图片中提取表格数据并输出为表格和CSV格式;当用户需要从图片提取表格数据、识别表格内容或导出CSV时使用;触发条件:试验图片表格提取,图片表格提取,试验图片数据提取,提取图片数据,提取试验数据
AI avatar creation and digital persona design powered by CellCog. Character.ai for agents — create characters, clone voices, generate images, build personalities. Persistent avatars reusable across all chats. Brand mascots, digital twins, creative characters, AI spokespersons, podcast hosts. Describe your character in plain language and CellCog builds it.