human-avatar使用阿里云 DashScope API 与阿里云 LingMou/灵眸生成多种 AI 视频与语音内容。七种能力:① LivePortrait 人像口播(图+音频→说话视频,两步流程)② EMO 人像口播 ③ AA/AnimateAnyone 全身动画(三步流程)④ T2I 文生图(万相2.x,默认 wan2.2-...
Install via ClawdBot CLI:
clawdbot install davideuler/human-avatarGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Sends data to undocumented external endpoint (potential exfiltration)
POST → https://dashscope.aliyuncs.com/api/v1/services/aigc/image2video/video-synthesis`Calls external URL not in known-safe list
https://example.com/portrait.jpgAI Analysis
The skill sends data only to documented Alibaba Cloud APIs (DashScope and LingMou) for its stated purpose of video generation, with no evidence of credential harvesting or hidden instructions. The 'unknown data sink' signal is a false positive as the endpoint is part of the legitimate service, and the example.com URL is just placeholder documentation.
Audited Apr 17, 2026 · audit v1.0
Generated Mar 21, 2026
Online retailers use the EMO mode to generate talking-head videos where a digital avatar presents product features in multiple languages, enhancing engagement and conversion rates. This reduces the need for human actors and allows for rapid content updates across different markets.
Companies leverage the LingMou mode with pre-built templates to create standardized training videos for new employees, ensuring consistent messaging and reducing production costs. The videos can be customized with specific text content for different departments or roles.
Content creators and influencers use the AA mode to animate full-body avatars for dance or fitness tutorials, making content more dynamic and shareable. This helps in building a unique brand identity without requiring complex video editing skills.
Businesses in banking or telecom employ the VideoRetalk mode to replace actors in instructional videos with digital avatars, enabling quick updates for policy changes or new services. This improves scalability and maintains a professional appearance across customer touchpoints.
Educational institutions use the EMO mode to generate lecture videos where instructors' avatars deliver audio content, making learning more accessible and engaging for remote students. It supports multiple languages and can be integrated into online learning platforms.
Offer a cloud-based platform where users pay a monthly fee to access the skill's modes for generating a limited number of videos, with tiered pricing based on usage volume and features like priority processing. This ensures recurring revenue and scalability for small to medium businesses.
Provide direct API access to enterprises for integrating the skill into their existing applications, such as CRM or e-commerce systems, charging per API call or based on video duration. This model targets developers and large organizations needing custom automation solutions.
License the technology to agencies or software vendors who rebrand it as part of their own video production tools, with revenue from upfront licensing fees and ongoing support contracts. This expands market reach through partnerships and reduces direct customer acquisition costs.
💬 Integration Tip
Ensure API keys are correctly configured for the cn-beijing region and use OSS for efficient file handling to avoid public URL limitations.
Scored Jun 17, 2026
Generate images using multiple AI models — Midjourney (via Legnext.ai), Flux, Nano Banana Pro (Gemini), Ideogram, Recraft, and more via fal.ai. Intelligently...
Perform image manipulation tasks like background removal, resizing, format conversion, rounding corners, watermarking, and color adjustments using ImageMagic...
Improves the quality of images, especially screenshots, by enhancing resolution, sharpness, and clarity. Perfect for preparing images for presentations, docu...
A specialized skill for generating high-quality, consistent children's bilingual picture books using 'banana nano'. Supports 18 visual styles across 4 catego...
从图片中提取表格数据并输出为表格和CSV格式;当用户需要从图片提取表格数据、识别表格内容或导出CSV时使用;触发条件:试验图片表格提取,图片表格提取,试验图片数据提取,提取图片数据,提取试验数据
AI avatar creation and digital persona design powered by CellCog. Character.ai for agents — create characters, clone voices, generate images, build personalities. Persistent avatars reusable across all chats. Brand mascots, digital twins, creative characters, AI spokespersons, podcast hosts. Describe your character in plain language and CellCog builds it.