glmv-captionGenerate captions (descriptions) for images, videos, and documents using ZhiPu GLM-V multimodal model series. Use this skill whenever the user wants to descr...
Install via ClawdBot CLI:
clawdbot install jaredforreal/glmv-captionGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://github.com/zai-org/GLM-V/tree/main/skills/glmv-captionAudited Apr 17, 2026 · audit v1.0
Generated May 6, 2026
An online retailer can automatically generate descriptive captions for thousands of product images, improving SEO and accessibility. The skill processes multiple images in batch and supports custom prompts to highlight specific features.
Media analysts can upload URLs of news clips or social media videos to get concise text summaries, enabling rapid content review and trend identification without manual viewing.
Law firms or researchers can upload PDFs or DOCX files (via URL) to obtain a summary or interpretation of content, aiding in document review and information extraction.
Web developers or content creators can generate alt text for images automatically, improving website accessibility for visually impaired users and complying with WCAG guidelines.
Data scientists can caption large collections of images and videos to create labeled datasets for training computer vision or multimodal AI models, using the script's batch and output-saving features.
Integrate the GLM-V caption skill into a SaaS product (e.g., a content management system) that charges users per caption or via subscription. The skill relies on Zhipu API, which can be passed through with usage-based pricing.
Offer services to enterprises to automate media description workflows (e.g., for accessibility compliance or content tagging). Charge a project fee for integration, customization (custom prompts, output formats), and maintenance.
Provide the skill as an open-source tool, but offer paid support plans (SLA, custom features, dedicated instance) to large organizations that require high reliability and throughput.
💬 Integration Tip
Ensure the ZHIPU_API_KEY is set in the environment before running the script. For batch processing, leverage the --output flag to save results as JSON for further automation.
Scored Jun 25, 2026
AI creative writing and storytelling powered by CellCog. Novels, short stories, screenplays, fan fiction, poetry. World building, character development, narrative design across fantasy, sci-fi, mystery, romance, horror, and literary fiction.
AI YouTube content creation powered by CellCog. YouTube videos, Shorts, thumbnails, video scripts, tutorials, vlogs, educational videos, product reviews, video essays. From script to finished video with voiceover and music.
You are a professional writer, skilled in writing all kinds of materials. Markdown is the exclusive format for your writing outputs.rrent user query.The othe...
AI长篇网文创作技能包。用于解决长篇网络小说创作中的核心痛点:上下文丢失、文风不一致、设定冲突、节奏失控、多线混乱、质量不稳、读者反馈无法内化。触发场景包括:开始新书、规划大纲、撰写章节、管理伏笔、检测冲突、读者反馈分析、批量创作质量控制。
Deprecated redirect skill that routes legacy 'content creator' requests to the correct specialist. Use when a user invokes 'content creator', asks to write a...
Refine academic writing for computer science research papers targeting top-tier venues (NeurIPS, ICLR, ICML, AAAI, IJCAI, ACL, EMNLP, NAACL, CVPR, WWW, KDD, SIGIR, CIKM, and similar). Use this skill whenever a user asks to improve, polish, refine, edit, or proofread academic or research writing — including paper drafts, abstracts, introductions, related work sections, methodology descriptions, experiment write-ups, or conclusion sections. Also trigger when users paste LaTeX content and ask for writing help, mention "camera-ready", "rebuttal", "paper revision", or reference any academic venue or conference. This skill handles both full paper refinement and section-by-section editing.