ppio-multimodal使用 PPIO 执行多模态任务:文生图、图生图、文生视频、图生视频、TTS、STT。 适用于:生成图片、生成视频、文字转语音、语音识别。
Install via ClawdBot CLI:
clawdbot install ximasadila/ppio-multimodalGrade Limited — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Sends data to undocumented external endpoint (potential exfiltration)
POST → https://api.ppio.com/v3/seedream-5.0-liteCalls external URL not in known-safe list
https://ppio.com/settings/key-managementAudited Apr 17, 2026 · audit v1.0
Generated Mar 21, 2026
Marketers and influencers can use this skill to generate images and videos for posts, ads, and stories based on text prompts or existing images. It streamlines content production, enabling rapid prototyping and customization for platforms like Instagram and TikTok.
Educators and e-learning platforms can create visual aids, explainer videos, and audio narrations from text descriptions. This enhances learning experiences by making complex topics more accessible through multimedia content.
Designers and startups can generate images and videos to visualize product concepts, mockups, and animations from textual or image inputs. This accelerates the ideation phase and facilitates client presentations without extensive manual design work.
Businesses can integrate TTS and STT capabilities to convert text responses to speech for automated phone systems or transcribe customer audio inquiries into text. This improves efficiency in handling support queries and enhances accessibility.
Charge users based on API calls for tasks like image generation, video creation, or speech processing. This model aligns with the skill's pricing structure, allowing scalability for high-volume users while offering low entry costs for occasional use.
Offer tiered subscription plans for businesses integrating the skill into their apps or workflows, providing higher usage limits, priority support, and advanced features. This ensures recurring revenue and caters to enterprise clients needing reliable multimedia generation.
License the skill as a white-label tool for marketing agencies or content studios, allowing them to rebrand and customize it for client projects. This expands market reach through partnerships and generates revenue from licensing fees and usage-based commissions.
💬 Integration Tip
Prioritize configuring the API Key via the recommended file method to ensure seamless authentication and avoid interruptions in task execution.
Scored Jun 19, 2026
Use CodexBar CLI local cost usage to summarize per-model usage for Codex or Claude, including the current (most recent) model or a full model breakdown. Trigger when asked for model-level usage/cost data from codexbar, or when you need a scriptable per-model summary from codexbar cost JSON.
Check Antigravity account quotas for Claude and Gemini models. Shows remaining quota and reset times with ban detection.
使用豆包(火山引擎)语音合成大模型 API 将文本转换为语音音频文件。支持声音复刻音色(S_ 开头的音色ID)和官方预置音色。当用户要求"语音合成"、"文字转语音"、"TTS"、"朗读文本"、"生成语音"、"用我的声音读"、"豆包语音"、"声音复刻合成"等相关请求时,务必使用此 skill。即使用户只是说"帮我把...
Intelligent model routing for sub-agent task delegation. Choose the optimal model based on task complexity, cost, and capability requirements. Reduces costs...
自动生成科技新闻摘要。从多个来源(RSS、Twitter、GitHub、Web Search)抓取科技新闻,整合后生成摘要。
Sync OpenRouter models used by OpenClaw into this installation's config. Fetches the OpenClaw app leaderboard from OpenRouter, verifies model IDs against the...