paddle-ocr-vlGPU-accelerated document parsing and OCR via PaddleOCR-VL. Detects layout, recognizes Chinese/English text, tables, charts, and seals in images. Use when the...
Install via ClawdBot CLI:
clawdbot install jessy-huang/paddle-ocr-vlGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://github.com/user/paddle-ocr-vl-skillAudited May 28, 2026 · audit v1.0
Generated Oct 2, 2026
Archives, libraries, and family history researchers use PaddleOCR-VL to convert scanned vertical classical Chinese texts such as the Records of the Three Kingdoms into searchable digital data. The skill's layout detection and vertical text recognition handle traditional columns that defeat mainstream OCR engines.
News organizations and media monitoring firms parse historical newspaper pages like People's Daily to build searchable article databases. Table, chart, and seal recognition allows accurate extraction of mixed editorial and advertisement content at scale.
Banks and accounting firms extract figures from scanned invoices, bank statements, and stamped contracts where seals and tables are common. GPU-accelerated batch OCR reduces processing time while keeping sensitive data on-premises inside ephemeral containers.
Hospitals convert scanned patient forms, lab reports, and handwritten Chinese medical records into structured electronic data. Local-only processing satisfies data privacy regulations because no images leave the host machine.
Government agencies and NGOs OCR photographed permits, signage, and forms collected during field inspections. The MCP server exposes the capability directly to Claude-based agents, enabling automated report generation from raw photos.
The core MCP server and Docker integration are freely distributed, monetized through paid deployment assistance, custom model fine-tuning, and SLA-backed support for enterprise document pipelines.
Customers purchase a preconfigured server or VM image bundling the PaddleOCR-VL containers and GPU drivers, avoiding setup complexity. The vendor handles upgrades, image caching, and hardware validation.
A cloud or private-cloud wrapper around PaddleOCR-VL charges per page processed, with volume tiers for newspapers, archives, and enterprises needing burst capacity beyond local GPUs.
💬 Integration Tip
Ensure Docker, nvidia-container-toolkit, and the correct PaddleOCR-VL image for your GPU architecture are installed before registering the MCP server in claude_desktop_config.json. Run check_environment and run_demo first to validate the setup against bundled samples.
Scored Jun 3, 2026
Fetch and read transcripts from YouTube videos. Use when you need to summarize a video, answer questions about its content, or extract information from it.
Monitor RSS and Atom feeds for content research. Track blogs, news sites, newsletters, and any feed source. Use when monitoring competitors, tracking industr...
用 MinerU API 解析 PDF/Word/PPT/图片为 Markdown,支持公式、表格、OCR。适用于论文解析、文档提取。
Provides a personalized morning report with today's reminders, undone Notion tasks, and vault storage summary for daily planning.
Extract text from PDFs with OCR support. Perfect for digitizing documents, processing invoices, or analyzing content. Zero dependencies required.
Fetch scheduled economic events and data releases from the FMP API for specified dates, filtering by impact, country, and type, and output a chronological ma...