ai-image-to-video-audioSkip the learning curve of professional editing software. Describe what you want — combine these images with my audio into a video with smooth transitions —...
Install via ClawdBot CLI:
clawdbot install peand-rover/ai-image-to-video-audioGrade Limited — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Sends data to undocumented external endpoint (potential exfiltration)
POST → https://mega-api-prod.nemovideo.ai/api/auth/anonymous-token`Calls external URL not in known-safe list
https://mega-api-prod.nemovideo.ai/api/auth/anonymous-token`AI Analysis
The skill communicates with a documented external API (nemovideo.ai) for its stated purpose of video generation, which is consistent with its functionality. While the API endpoint is not on a pre-approved list, there is no evidence of credential harvesting, hidden instructions, or obfiltration of user data beyond the necessary file uploads for processing.
Audited Apr 17, 2026 · audit v1.0
Generated May 9, 2026
Marketers can combine product images with a voiceover or background music to create promotional videos without specialized editing skills. This enables rapid production of video ads for social media, landing pages, or email campaigns.
E-commerce sellers can upload multiple product photos and audio descriptions to generate a video that highlights features and benefits. This streamlines creating product demos or listings for platforms like Amazon or Shopify.
Educators and trainers can turn instructional images and voiceover audio into engaging video lessons. This reduces the time and cost of producing multimedia educational materials.
Content creators can quickly turn a series of photos and a soundtrack into a video for platforms like Instagram, TikTok, or YouTube. This helps maintain a consistent posting schedule without complex video editing.
Real estate agents can combine property photos with a narrated audio tour to generate a virtual walkthrough video. This provides a dynamic listing presentation for potential buyers.
Users get 100 free credits upon registration, which expire in 7 days. Additional credits can be purchased to continue using the service.
Monthly or yearly subscription plans offer a fixed number of renditions per month, with higher tiers including faster processing, higher resolutions, and priority support.
Businesses can license the API for integration into their own platforms, enabling bulk video generation for customer-facing applications or internal workflows.
💬 Integration Tip
Ensure the NEMO_TOKEN environment variable is set and the config directory exists for seamless authentication. Upload files with accepted formats (JPG, PNG, MP3, WAV, etc.) and keep requests concise for faster processing.
Scored May 9, 2026
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...
Create, manage, and deploy ElevenLabs conversational AI agents. Use when the user wants to work with voice agents, list their agents, create new ones, or manage agent configurations.
Voice note transcription and archival for OpenClaw agents. Powered by Deepgram Nova-3. Transcribes audio messages, saves both audio files and text transcript...
Text-to-Speech (TTS) and Speech-to-Text (ASR) using coze-coding-dev-sdk. Returns results directly to stdout.