local-image-ocr-aipcImage OCR, text recognition, extract text from image, scan document, read image text, invoice OCR, receipt OCR, contract recognition, table extraction, busin...
Install via ClawdBot CLI:
clawdbot install violet17/local-image-ocr-aipcGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://github.com/ggml-org/llama.cpp/releases`Audited Apr 16, 2026 · audit v1.0
Generated May 20, 2026
Convert scanned paper documents, such as contracts or invoices, into editable digital text. The skill runs locally on Windows using GLM-OCR, ensuring data privacy without cloud API calls.
Extract text from receipt or invoice images automatically, enabling quick data entry for expense tracking or bookkeeping. Mixed Chinese/English capability supports bilingual documents.
Scan and extract information from ID cards or business cards, useful for customer onboarding or contact management. The local setup eliminates the need for external API keys.
Grab text from screenshots of articles, presentations, or social media posts. Ideal for researchers or content creators who need to quickly capture and reuse text from images.
Parse tabular data from scanned tables in reports or forms, converting them into structured text. Useful for data entry and analysis in various business processes.
Charge an initial fee to install and configure the local OCR environment on the client's Windows machine, then a monthly subscription per user or per 1,000 pages processed. Reduces per-transaction costs for high-volume users.
Offer basic OCR (e.g., 10 pages/day) for free, with paid tiers for higher volume, batch processing, or additional output formats (e.g., searchable PDF). Revenue from subscriptions or one-time purchases.
License the entire skill package to enterprises for internal deployment and use across departments. Includes support and custom integration, priced per seat or per organization.
💬 Integration Tip
Integrate with PowerShell automation scripts by calling the skill via llama-cli.exe with the image path as argument. Use the provided environment setup to ensure models are pre-downloaded for faster processing.
Scored Jun 29, 2026
Generate images using multiple AI models — Midjourney (via Legnext.ai), Flux, Nano Banana Pro (Gemini), Ideogram, Recraft, and more via fal.ai. Intelligently...
Perform image manipulation tasks like background removal, resizing, format conversion, rounding corners, watermarking, and color adjustments using ImageMagic...
Improves the quality of images, especially screenshots, by enhancing resolution, sharpness, and clarity. Perfect for preparing images for presentations, docu...
A specialized skill for generating high-quality, consistent children's bilingual picture books using 'banana nano'. Supports 18 visual styles across 4 catego...
从图片中提取表格数据并输出为表格和CSV格式;当用户需要从图片提取表格数据、识别表格内容或导出CSV时使用;触发条件:试验图片表格提取,图片表格提取,试验图片数据提取,提取图片数据,提取试验数据
AI avatar creation and digital persona design powered by CellCog. Character.ai for agents — create characters, clone voices, generate images, build personalities. Persistent avatars reusable across all chats. Brand mascots, digital twins, creative characters, AI spokespersons, podcast hosts. Describe your character in plain language and CellCog builds it.