ifly-pdf-image-ocrifly-pdf&image-ocr skill supporting both image OCR (AI-powered LLM OCR) and PDF document recognition. Use when user asks to OCR images, extract text from ima...
Install via ClawdBot CLI:
clawdbot install qingzhe2020/ifly-pdf-image-ocrGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://example.com/doc.pdfAudited Apr 18, 2026 · audit v1.0
Generated May 10, 2026
Convert large volumes of paper documents (invoices, contracts, reports) into editable Word or Markdown files. Ideal for automating data entry and archival in corporate environments.
Researchers extract text from multi-language PDFs or images (e.g., historical manuscripts, foreign journals) using OCR that supports layout understanding and multi-language recognition.
Automate extraction of key fields from scanned invoices and receipts for accounting and expense management. Output structured JSON for integration with ERP systems.
Convert scanned PDFs into accessible Markdown or Word documents for use with screen readers or content management systems, preserving document structure.
Integrate image OCR into mobile applications to allow users to snap photos of text (menus, signs, notes) and get instant digital text output, supporting multiple output formats.
Offer OCR processing as a cloud service with tiers based on pages processed per month. Customers pay a recurring fee for access to the API and scripts.
Charge per OCR request (per page or per image) for customers with variable demand. Ideal for low-volume or occasional users.
License the OCR technology to other companies to integrate into their own products under their brand. Charge a licensing fee plus royalty per transaction.
💬 Integration Tip
Set up IFLY_APP_ID, IFLY_API_KEY, and IFLY_API_SECRET environment variables. For PDF OCR, ensure the timestamp is within 5 minutes of server time.
Scored May 10, 2026
Generate images using multiple AI models — Midjourney (via Legnext.ai), Flux, Nano Banana Pro (Gemini), Ideogram, Recraft, and more via fal.ai. Intelligently...
Perform image manipulation tasks like background removal, resizing, format conversion, rounding corners, watermarking, and color adjustments using ImageMagic...
A specialized skill for generating high-quality, consistent children's bilingual picture books using 'banana nano'. Supports 18 visual styles across 4 catego...
从图片中提取表格数据并输出为表格和CSV格式;当用户需要从图片提取表格数据、识别表格内容或导出CSV时使用;触发条件:试验图片表格提取,图片表格提取,试验图片数据提取,提取图片数据,提取试验数据
Generate or edit images with Gemini using the Google GenAI SDK. Use when the user asks to create, transform, render, or save one or more images in an OpenCla...
提供图片生成能力:输入文本描述生成图片,可附加风格提示词。支持 1-4 张图片生成,按图片尺寸对应点数计费。Authorization 使用本系统发放的 Bearer Token。