fundreport-scrape基金月报信息提取。支持文本+OCR 双重提取,自动处理双月对比。从 PDF 月报提取数据并填充 Excel 模板。
Install via ClawdBot CLI:
clawdbot install imkiiki/fundreport-scrapeGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Potentially destructive shell commands in tool definitions
eval(Calls external URL not in known-safe list
https://github.com/UB-Mannheim/tesseract/wikiAI Analysis
The skill processes local PDF and Excel files for data extraction using OCR and text parsing, with no evidence of sending user data to external servers. The only external reference is to a known, legitimate Tesseract OCR installation guide. The primary risk is the presence of `eval()` in shell commands, which could be exploited if user input is not properly sanitized, but this appears within a controlled script context.
Audited Apr 16, 2026 · audit v1.0
Generated Mar 22, 2026
基金公司需要定期从PDF月报中提取核心指标和分布数据,用于内部监控和监管报告。该技能通过文本和OCR双重提取,自动处理双月对比,批量处理多只基金,显著减少人工数据录入时间,提高准确性和效率。
研究机构分析基金表现时,需从大量月报中收集久期、YTM、行业分布等数据。该技能支持自定义Excel模板,保持原有样式和公式,自动填充对比数据,便于研究人员快速生成分析报告和趋势图表。
资产管理公司需定期向客户和监管机构提交基金月报摘要。该技能提取关键指标并生成对比Excel,确保数据一致性,支持批量处理,帮助合规团队高效准备报告,减少手动错误风险。
财务顾问为客户提供基金投资更新时,需从月报中提取数据制作个性化简报。该技能自动识别基金名称和日期,智能分Sheet处理,生成包含上月和本月对比的Excel文件,便于顾问快速定制客户沟通材料。
研究人员进行基金绩效或市场分析时,需从历史PDF月报中提取结构化数据。该技能支持OCR处理图表数据,适应不同表述方式,批量处理多只基金,为学术研究提供可靠的数据基础。
提供基于云的平台,用户上传PDF和模板后自动处理,按月或按年收费。包括基础数据处理和高级OCR功能,适合中小型基金公司或研究机构,降低本地部署成本。
为大型金融机构提供定制化部署,集成到现有系统中,支持批量处理和高级数据分析。包括专属支持、培训和维护服务,确保高可靠性和安全性。
用户按处理的PDF文件数量或数据提取量付费,适合临时或低频需求。提供自助服务平台,无需长期承诺,吸引个体顾问或小型团队。
💬 Integration Tip
确保系统已安装Tesseract OCR和Poppler-utils依赖,并配置中文语言包,以优化PDF处理和OCR识别准确率。
Scored Jun 19, 2026
A fast Rust-based headless browser automation CLI with Node.js fallback that enables AI agents to navigate, click, type, and snapshot pages via structured commands.
Uses a headless browser to navigate web pages, interact with elements, and extract clean, readable text content from URLs.
Headless browser automation CLI optimized for AI agents with accessibility tree snapshots and ref-based element selection
通过已登录的Edge或Chrome浏览器,利用Chrome DevTools Protocol执行JS渲染页面的导航、点击、截图及数据提取等自动化操作。
High-performance browser automation for heavy scraping, multi-tab management, and precise DOM extraction. Use this when you need speed, reliability, or advanced state management (cookies/local storage) beyond standard web fetching.
Ultimate stealth browser automation with anti-detection, Cloudflare bypass, CAPTCHA solving, persistent sessions, and silent operation. Use for any web automation requiring bot detection evasion, login persistence, headless browsing, or bypassing security measures. Triggers on "bypass cloudflare", "solve captcha", "stealth browse", "silent automation", "persistent login", "anti-detection", or any task needing undetectable browser automation. When user asks to "login to X website", automatically use headed mode for login, then save session for future headless reuse.