photo-ocrOCR for photos and images using MinerU. Extract text from photographs, screenshots, camera captures, and image files with high accuracy. Features: image OCR...
Grade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://mineru.netAudited Apr 17, 2026 · audit v1.0
Generated May 22, 2026
Automatically extract text from photos of receipts and invoices, converting them into structured data for expense tracking or accounting. Ideal for small business owners and freelancers who need to digitize paper receipts quickly.
Use the skill to capture notes from whiteboard photos or text from signs and posters, converting them into editable Markdown or text. Useful for meeting notes, brainstorming sessions, or translating signage.
Extract text from screenshots of web pages, presentations, or error messages. Developers and researchers can quickly capture code snippets, quotes, or data without manual typing.
Digitize photos of paper documents, such as contracts, letters, or forms, for storage or search. Archives and libraries can use this to preserve and index physical documents.
Extract text from images containing multiple languages, such as multilingual menus or international signs. Supports English, Chinese, and other languages, making it useful for travel and global business.
Offer basic OCR via flash-extract for free, no token required, with limitations on file size and pages. Charge for premium extract with higher accuracy, VLM mode, and no size limits, monetized via token purchases.
Provide the mineru-open-api as a service, charging per API call or subscription for developers integrating OCR into their apps. The CLI tool can be used directly or via API wrappers.
License the skill to enterprises for automated document processing, such as invoice scanning or data entry. Offer custom integrations, on-premises deployment, and dedicated support.
💬 Integration Tip
Start with flash-extract for quick testing without a token. For production, authenticate with MINERU_TOKEN and use extract for higher accuracy on complex images.
Scored Jun 19, 2026
Generate images using multiple AI models — Midjourney (via Legnext.ai), Flux, Nano Banana Pro (Gemini), Ideogram, Recraft, and more via fal.ai. Intelligently...
Perform image manipulation tasks like background removal, resizing, format conversion, rounding corners, watermarking, and color adjustments using ImageMagic...
A specialized skill for generating high-quality, consistent children's bilingual picture books using 'banana nano'. Supports 18 visual styles across 4 catego...
从图片中提取表格数据并输出为表格和CSV格式;当用户需要从图片提取表格数据、识别表格内容或导出CSV时使用;触发条件:试验图片表格提取,图片表格提取,试验图片数据提取,提取图片数据,提取试验数据
Generate or edit images with Gemini using the Google GenAI SDK. Use when the user asks to create, transform, render, or save one or more images in an OpenCla...
提供图片生成能力:输入文本描述生成图片,可附加风格提示词。支持 1-4 张图片生成,按图片尺寸对应点数计费。Authorization 使用本系统发放的 Bearer Token。