image-to-code将图片(含文字、公式、标题)转换为指定代码格式。自动识别标题级别(title1/title2/title3),文字行转为 $word->body("正文=".$F);,公式转为 $word->formula("");,图片标记为 ![image]
Install via ClawdBot CLI:
clawdbot install nidhov01/image-to-codeGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://tesseract-ocr.github.io/Uses known external API (expected, informational)
aip.baidubce.comAudited Apr 16, 2026 · audit v1.0
Generated Mar 22, 2026
Converts scanned academic papers with titles, text, formulas, and images into structured code format for easy integration into LaTeX or markdown editors. Useful for researchers digitizing handwritten notes or printed materials.
Processes screenshots of technical manuals or reports containing hierarchical headings, body text, and mathematical equations into a standardized code output. Helps technical writers automate formatting tasks.
Transforms textbook pages or lecture slides with titles, content, and formulas into code snippets for e-learning platforms or interactive tutorials. Assists educators in quickly repurposing materials.
Converts scanned business reports with sections, bullet points, and charts into a code format for data processing or presentation tools. Streamlines report generation from physical documents.
Extracts text and formulas from screenshots of code documentation or whiteboards, converting them into structured comments or documentation files. Aids developers in maintaining up-to-date docs.
Offer the tool as a cloud-based service with tiered pricing based on usage volume (e.g., pages per month). Includes API access for integration into other platforms, targeting enterprises and institutions.
Provide a free version with basic OCR and limited conversions, and a paid version with advanced features like batch processing, custom templates, and priority support. Targets individual users and small teams.
License the core technology to educational platforms, content management systems, or document processing companies for integration into their products. Includes customization and technical support.
💬 Integration Tip
Integrate with existing OCR pipelines or document management systems by using the provided Python class structure, ensuring proper image preprocessing for optimal accuracy.
Scored Jun 19, 2026
Generate images using multiple AI models — Midjourney (via Legnext.ai), Flux, Nano Banana Pro (Gemini), Ideogram, Recraft, and more via fal.ai. Intelligently...
Perform image manipulation tasks like background removal, resizing, format conversion, rounding corners, watermarking, and color adjustments using ImageMagic...
A specialized skill for generating high-quality, consistent children's bilingual picture books using 'banana nano'. Supports 18 visual styles across 4 catego...
从图片中提取表格数据并输出为表格和CSV格式;当用户需要从图片提取表格数据、识别表格内容或导出CSV时使用;触发条件:试验图片表格提取,图片表格提取,试验图片数据提取,提取图片数据,提取试验数据
Generate or edit images with Gemini using the Google GenAI SDK. Use when the user asks to create, transform, render, or save one or more images in an OpenCla...
提供图片生成能力:输入文本描述生成图片,可附加风格提示词。支持 1-4 张图片生成,按图片尺寸对应点数计费。Authorization 使用本系统发放的 Bearer Token。