image-recognize百度AI识别图片中的物体、场景、文字等内容,需要用户提供本地图片或网络图片,支持Base64编码。支持(题目,文字,图片人脸,植物,动物,表情,素材,商品,玩具,景点,通用识别等)内容识别。用于通用图片内容分类识别,不负责图片生成或编辑
Install via ClawdBot CLI:
clawdbot install ide-rea/image-recognizeGrade Limited — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://example.com/image.jpgAudited Apr 17, 2026 · audit v1.0
Generated May 10, 2026
电商平台上传商品图片后,自动识别商品类别并生成分类标签,如识别出为鞋子、电子产品等,辅助商品上架管理。适用于电商平台的商品管理场景,无需人工标注即可完成分类。
用户上传旅游照片,识别出具体景点名称(如故宫、西湖),并返回景点介绍和百科链接,可用于旅游APP的智能导游功能,提升用户体验。
学生或自然爱好者拍摄动植物照片,识别出物种名称和详细信息,用于生物学习或户外探索,辅助教育应用或科普平台。
社交媒体平台用户上传图片,自动识别图片内容(如是否包含商品、人物、表情等),辅助内容审核或标签推荐,提升内容管理效率。
设计师或内容创作者上传图片,自动识别图片主题(如风景、动物、植物),并生成标签,便于素材库的检索和管理。适用于图库网站或设计平台。
开放API接口给第三方开发者或企业,按调用次数收费。例如每次识别0.01元,量大优惠。适合需要批量图片识别的客户。
提供网页版或集成工具,按月或年收费。用户上传图片即可获取识别结果,适用于中小企业和个人用户。
针对特定行业(如电商、旅游)提供定制化识别服务,包括私有化部署或定制模型,收取项目费或年费。
💬 Integration Tip
集成时需确保环境变量BAIDU_API_KEY已设置,图片大小不超过4MB。通过命令行或API调用即可获得结果,非常适合快速嵌入现有系统。
Scored Apr 19, 2026
Generate images using multiple AI models — Midjourney (via Legnext.ai), Flux, Nano Banana Pro (Gemini), Ideogram, Recraft, and more via fal.ai. Intelligently...
Perform image manipulation tasks like background removal, resizing, format conversion, rounding corners, watermarking, and color adjustments using ImageMagic...
Improves the quality of images, especially screenshots, by enhancing resolution, sharpness, and clarity. Perfect for preparing images for presentations, docu...
A specialized skill for generating high-quality, consistent children's bilingual picture books using 'banana nano'. Supports 18 visual styles across 4 catego...
从图片中提取表格数据并输出为表格和CSV格式;当用户需要从图片提取表格数据、识别表格内容或导出CSV时使用;触发条件:试验图片表格提取,图片表格提取,试验图片数据提取,提取图片数据,提取试验数据
Generate AI videos, images and speech (text-to-video, image-to-video, reference-to-video, speech-to-video, animate, text-to-image, image-to-image, image edit, text-to-speech), bring your own LoRA models, and generate AI product/model photography for e-commerce (Image Studio) via the Phosor AI platform. Use when the user wants to create videos or images from text prompts, animate images, generate lip-synced video from audio, synthesize speech from text, generate images with a custom LoRA, generate product photography or model/clothing photography for e-commerce listings, or manage generation jobs.