ugc-manualGenerate lip-sync video from image + user's own audio recording. ✅ USE WHEN: - User provides their OWN audio file (voice recording) - Want to sync image to specific audio/voice - User recorded the script themselves - Need exact audio timing preserved ❌ DON'T USE WHEN: - User provides text script (not audio) → use veed-ugc - Need AI to generate the voice → use veed-ugc - Don't have audio file yet → use veed-ugc with script INPUT: Image + audio file (user's recording) OUTPUT: MP4 video with lip-sync to provided audio KEY DIFFERENCE: veed-ugc = script → AI voice → video ugc-manual = user audio → video (no voice generation)
Install via ClawdBot CLI:
clawdbot install PauldeLavallaz/ugc-manualGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://api.comfydeploy.com/api/run/deployment/queue`Audited Apr 17, 2026 · audit v1.0
Generated Mar 21, 2026
Users can create custom video messages by recording their own voice and syncing it to a photo of themselves or a character. This is ideal for sending unique birthday wishes, thank-you notes, or announcements where personal touch is valued.
Teachers or trainers record audio lessons and sync them to an image of themselves or an avatar, making online courses more engaging. It helps preserve exact timing for language pronunciation or step-by-step tutorials.
Businesses use pre-recorded audio ads or voiceovers from influencers and sync them to branded images or character visuals. This allows for consistent messaging without relying on AI-generated voices, enhancing authenticity.
Fans create lip-sync videos by combining audio from songs, podcasts, or movie dialogues with images of celebrities or fictional characters. It's popular for social media challenges and fan art, leveraging user-generated audio.
Offer basic video generation for free with limited features, then charge for premium options like higher resolution, faster processing, or bulk exports. This model attracts individual users and small businesses looking for cost-effective solutions.
License the technology to companies in education, marketing, or social media platforms, allowing them to integrate lip-sync features into their own apps. This generates revenue through licensing fees and custom development contracts.
Charge users based on the number of videos generated or processing time, ideal for developers and enterprises with fluctuating needs. This model scales with usage and can include volume discounts for high-frequency clients.
💬 Integration Tip
Ensure ffmpeg is installed for audio conversion, and use the provided Python script with direct file paths or URLs for seamless automation.
Scored Apr 19, 2026
Sora2 视频平台数据查询助手。覆盖作品详情、用户数据、搜索、评论、Cameo等全功能。
Generate new videos from text prompts, images, or reference inputs using EachLabs AI models. Supports text-to-video, image-to-video, transitions, motion control, talking head, and avatar generation. Use when the user wants to create new video content. For editing existing videos, see eachlabs-video-edit.
Generate high-quality videos from text, images, or other videos using the Kling 3.0 Omni model. Covers text-to-video, image-to-video, video editing, video re...
为 Seedance 2.5 设计、诊断、续接和重写通用动作/打斗视频提示词。覆盖徒手格斗、武器对决、一对多、团队连携、Boss战、追逐战、现代动作、武侠、奇幻、科幻、动画/CG等。严格执行“先补齐必要输入,再生成提示词”;基于成片修改时必须先实际检查视频证据,不得凭历史对话或旧提示词猜测。
Seedance 2.0 二次元打戏视频提示词生成智能体。飞驰影漫社+废话军师联合改造。
Expert prompt engineering for Seedance 2.0. Use when the user wants to generate a video with multimodal assets (images, videos, audio) and needs the best pos...