geminipdfocrExtract text from PDFs using Google Gemini OCR. Use when extracting text from PDFs, performing OCR on scanned documents, or processing image-based PDFs.
Install via ClawdBot CLI:
clawdbot install AshtonIzmev/geminipdfocrGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Generated Mar 21, 2026
Law firms can use this skill to convert scanned legal documents, such as contracts and court filings, into searchable text. This enables efficient document management, keyword searches, and archival without manual transcription, saving time and reducing errors in legal workflows.
Researchers and universities can extract text from scanned academic papers or historical PDFs for analysis and citation. This facilitates literature reviews, data mining, and accessibility for digital libraries, supporting scholarly work across disciplines like humanities and sciences.
Healthcare providers can digitize patient records, such as scanned medical forms and reports, into editable text for electronic health record systems. This improves data accessibility, supports compliance with digital health standards, and enhances patient care coordination.
Banks and financial institutions can process scanned invoices, receipts, and statements to extract financial data for auditing and analysis. This automates data entry, reduces manual effort in accounting, and helps in fraud detection and regulatory reporting.
Libraries and museums can use this skill to convert historical documents, books, and manuscripts from image-based PDFs into searchable digital formats. This preserves cultural heritage, enables public access through online archives, and supports academic research.
Offer a cloud-based OCR service with tiered pricing based on usage, such as pages processed per month. This model provides recurring revenue, scalability for businesses of all sizes, and includes features like API access and batch processing for enterprise clients.
Charge users per page or document processed through an API, ideal for developers and companies with variable OCR needs. This model offers flexibility, low entry costs, and can be integrated into existing applications, generating revenue based on actual usage.
Sell custom licenses to large organizations for on-premise or private cloud deployment, including support and customization. This model targets industries with high security needs, such as legal or healthcare, providing steady revenue through long-term contracts.
💬 Integration Tip
Ensure the GOOGLE_API_KEY is securely stored and passed via environment variables to avoid exposure in logs or code, and test with small PDFs first to validate output quality and API limits.
Scored Apr 19, 2026
Use CodexBar CLI local cost usage to summarize per-model usage for Codex or Claude, including the current (most recent) model or a full model breakdown. Trigger when asked for model-level usage/cost data from codexbar, or when you need a scriptable per-model summary from codexbar cost JSON.
Check Antigravity account quotas for Claude and Gemini models. Shows remaining quota and reset times with ban detection.
使用豆包(火山引擎)语音合成大模型 API 将文本转换为语音音频文件。支持声音复刻音色(S_ 开头的音色ID)和官方预置音色。当用户要求"语音合成"、"文字转语音"、"TTS"、"朗读文本"、"生成语音"、"用我的声音读"、"豆包语音"、"声音复刻合成"等相关请求时,务必使用此 skill。即使用户只是说"帮我把...
Intelligent model routing for sub-agent task delegation. Choose the optimal model based on task complexity, cost, and capability requirements. Reduces costs...
自动生成科技新闻摘要。从多个来源(RSS、Twitter、GitHub、Web Search)抓取科技新闻,整合后生成摘要。
Sync OpenRouter models used by OpenClaw into this installation's config. Fetches the OpenClaw app leaderboard from OpenRouter, verifies model IDs against the...