smart-pdf-ocrIntelligent PDF OCR powered by MinerU API. Extract text from scanned PDFs, image-based PDFs, and photographed documents using mineru-open-api CLI with advanc...
Install via ClawdBot CLI:
clawdbot install veeicwgy/smart-pdf-ocrGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Generated Oct 6, 2026
Law firms and courts scan decades of contracts, court filings, and case documents into image PDFs. This skill converts them into searchable text while preserving table structures and complex legal formatting through the VLM model.
Hospitals and clinics digitize handwritten or scanned patient forms, lab reports, and prescriptions. Multilingual OCR support handles diverse patient records, enabling electronic health record systems to index and retrieve information.
Accounting teams use flash-extract to quickly process batches of scanned invoices, expense receipts, and tax documents under 10MB. Extracted text feeds into automated bookkeeping and reconciliation systems.
Universities and libraries digitize historical manuscripts, theses, and academic papers containing formulas and mixed layouts. The VLM model accurately recognizes mathematical notation and multi-column academic formatting.
Agencies convert citizen applications, permits, and archival documents in multiple languages into searchable digital records. This supports compliance requirements and public records request workflows.
Offer tiered monthly subscriptions based on page volume processed, with flash-extract included for small batches and metered pricing for VLM precision extraction. Customers pay for advanced features like table and formula recognition.
Sell an on-premise or cloud platform to large organizations with high-volume archival needs, bundling the OCR skill with workflow management, storage, and integration into existing document management systems.
Provide free flash-extract tier for developers building prototypes, charging credits for advanced extract calls using VLM or pipeline models. Integrate as an npm package with pay-as-you-go billing.
💬 Integration Tip
Install globally via npm and start with flash-extract for files under 10MB/20 pages to validate output before moving to token-based VLM extraction for complex layouts.
Scored Oct 6, 2026
PDF智能处理工具 v2.1 | PDF Smart Tool. 支持PDF转换、OCR识别、合并拆分、数字签名、批量处理、水印添加、加密解密。触发词:PDF、转换、识别。
Generate hand-drawn style diagrams, flowcharts, and architecture diagrams as PNG images from Excalidraw JSON
Convert public web pages into clean Markdown with markdown.new for AI workflows. Use when tasks require URL-to-Markdown conversion for summarization, RAG ing...
PDF扫描件转Word文档。支持中文OCR识别,自动裁掉页眉页脚,保留插图,彩色章节封面页保留为图片。使用百度OCR API(免费额度1000次/月)。当用户要求把扫描PDF转成文字/Word时触发。
Word文档处理工具套件,提供Word文档的创建、读取、内容提取和基本处理功能。
Microsoft SharePoint and OneDrive integration with managed OAuth. Manage sites, lists, libraries, files, folders, permissions, content types, and SharePoint...