junyi-doc-reader大文档归档与检索管线(v5)。支持本地文件(Word/PDF/TXT/Markdown)和飞书云文档,转换、分块、可选 LLM 增强,输出结构化 Markdown 和索引,存入 Obsidian。触发词:读大文档、归档文档、junyi-doc-reader、doc-reader、文档索引、帮我读这个PDF、把文档...
Install via ClawdBot CLI:
clawdbot install xuanranc/junyi-doc-readerGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Generated Mar 22, 2026
Researchers can process large PDF papers, automatically convert them to structured Markdown with summaries and keywords, and organize them in Obsidian for literature review. This helps in quickly extracting insights and building a searchable knowledge base without manual note-taking.
Law firms can use this skill to convert lengthy legal documents (e.g., contracts, case files) into indexed Markdown, enabling efficient retrieval of specific clauses or sections. Optional LLM insights can highlight key terms and classifications for faster case preparation.
Companies can archive internal documents like policy manuals, reports, and training materials into Obsidian, making them easily searchable for employees. The indexing and chunking features allow for quick access to relevant information across large document sets.
Content creators can process source materials such as e-books or research PDFs, extract structured content with summaries, and use the output to generate blog posts or social media content. This streamlines research and ensures accurate referencing.
Offer a cloud service where users upload documents via a web interface, and the skill processes them with optional LLM enhancements. Charge monthly fees based on document volume or features like advanced insights and priority support.
Sell on-premise or private cloud licenses to large organizations needing secure document processing without external API calls. Include custom integrations, training, and dedicated support for high-value clients in sectors like legal or finance.
Provide a free tier for basic document conversion and indexing, with paid credits for LLM-enhanced insights and bulk processing. Target individual users and small teams, upselling to advanced features as needs grow.
💬 Integration Tip
Ensure pandoc and poppler are installed on the system, and set DOC_READER_ALLOW_EXTERNAL=true for LLM features to work with external APIs.
Scored Jun 19, 2026
Fetch and read transcripts from YouTube videos. Use when you need to summarize a video, answer questions about its content, or extract information from it.
Monitor RSS and Atom feeds for content research. Track blogs, news sites, newsletters, and any feed source. Use when monitoring competitors, tracking industr...
用 MinerU API 解析 PDF/Word/PPT/图片为 Markdown,支持公式、表格、OCR。适用于论文解析、文档提取。
Provides a personalized morning report with today's reminders, undone Notion tasks, and vault storage summary for daily planning.
Extract text from PDFs with OCR support. Perfect for digitizing documents, processing invoices, or analyzing content. Zero dependencies required.
Fetch scheduled economic events and data releases from the FMP API for specified dates, filtering by impact, country, and type, and output a chronological ma...