mineru-pdfParse PDF documents with MinerU MCP to extract text, tables, and formulas. Supports multiple backends including MLX-accelerated inference on Apple Silicon.
Install via ClawdBot CLI:
clawdbot install etoile04/mineru-pdfGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://github.com/TINKPA/mcp-mineruAudited Apr 16, 2026 · audit v1.0
Generated Mar 1, 2026
Researchers can extract text, tables, and formulas from academic PDFs to quickly gather data for literature reviews or meta-analyses. This accelerates data collection from sources like scientific journals and conference proceedings, enabling faster synthesis of findings.
Finance teams can parse quarterly reports, balance sheets, and regulatory filings to extract tabular data and textual insights for analysis. This automates data entry from PDF documents, reducing manual errors and saving time in financial modeling.
Law firms can use the skill to extract text and structured content from contracts, case files, and legal briefs for faster review and indexing. It helps in identifying key clauses, terms, and data points across large document sets.
Healthcare providers can parse patient records, lab reports, and medical forms stored as PDFs to extract text and tabular data for electronic health record systems. This supports data migration and improves accessibility of patient information.
Engineers and technicians can extract formulas, tables, and text from technical manuals, datasheets, and schematics for reference or integration into design software. This aids in product development and maintenance processes.
Offer a cloud-based PDF parsing service with tiered pricing based on usage volume, such as pages processed per month. This model targets businesses needing regular document processing, with features like API access and priority support.
Sell on-premise or custom licenses to large organizations for integration into their internal workflows, with options for Apple Silicon optimization. This includes dedicated support, training, and customization for specific use cases like legal or finance.
Provide a free tier for basic PDF text extraction with limited pages, while charging for advanced features like table/formula recognition, batch processing, and faster MLX acceleration. This attracts individual users and upsells to teams.
💬 Integration Tip
Use the direct tool for persistent file output in production workflows, and leverage the MLX backend on Apple Silicon for optimal performance with complex documents.
Scored Apr 19, 2026
Fetch and read transcripts from YouTube videos. Use when you need to summarize a video, answer questions about its content, or extract information from it.
Monitor RSS and Atom feeds for content research. Track blogs, news sites, newsletters, and any feed source. Use when monitoring competitors, tracking industr...
用 MinerU API 解析 PDF/Word/PPT/图片为 Markdown,支持公式、表格、OCR。适用于论文解析、文档提取。
Provides a personalized morning report with today's reminders, undone Notion tasks, and vault storage summary for daily planning.
Extract text from PDFs with OCR support. Perfect for digitizing documents, processing invoices, or analyzing content. Zero dependencies required.
Fetch scheduled economic events and data releases from the FMP API for specified dates, filtering by impact, country, and type, and output a chronological ma...