jiekou-multimodal使用接口AI 执行多模态任务:文生图、图生图、文生视频、图生视频、TTS、STT。 适用于:生成图片、生成视频、文字转语音、语音识别。
Install via ClawdBot CLI:
clawdbot install ximasadila/jiekou-multimodalGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Sends data to undocumented external endpoint (potential exfiltration)
POST → https://api.jiekou.ai/v3/gemini-3.1-flash-image-text-to-imageCalls external URL not in known-safe list
https://jiekou.ai/settings/key-managementAI Analysis
The skill's external API calls (api.jiekou.ai) are directly documented and consistent with its stated multimodal task purpose. The configuration process and key management link (jiekou.ai) are typical for legitimate third-party integrations and do not show evidence of credential harvesting or hidden malicious instructions.
Audited Apr 16, 2026 · audit v1.0
Generated Mar 20, 2026
Marketers and influencers can use this skill to generate images and videos for posts, ads, or stories based on text prompts. It supports quick model options for urgent content needs, enhancing engagement and visual appeal across platforms like Instagram or TikTok.
Educators and e-learning platforms can create visual aids, explainer videos, and audio narrations from text or images. The TTS feature helps generate voiceovers for courses, making content more accessible and engaging for students.
Designers and startups can generate images and videos to visualize product concepts, mockups, or marketing materials. The image editing and video generation capabilities allow for rapid iteration and presentation of ideas without extensive manual work.
Businesses can integrate this skill to convert customer queries into audio responses via TTS or generate visual guides. It helps automate support tasks, providing quick, multimodal assistance and improving user experience in apps or websites.
Game developers and content creators can use it to generate assets like character images, scene videos, or voice lines. The fast model option accelerates production for time-sensitive projects, such as live streams or game updates.
Offer this skill as a paid API service where developers or businesses pay per use based on tasks like image generation or TTS. Revenue comes from usage fees, with tiered pricing for high-volume clients, leveraging the skill's multimodal capabilities.
Build a platform where users get free basic generations with watermarks or limits, and pay for premium features like faster models, higher quality, or commercial use. Monetize through upgrades and ad placements on generated content.
License the skill to large companies for internal use in marketing, training, or product development. Provide custom integrations, support, and bulk pricing, generating steady revenue from long-term contracts and service fees.
💬 Integration Tip
Start by configuring the API key via the recommended file method for reliability, and always send progress prompts before API calls to improve user experience.
Scored Jun 17, 2026
Use CodexBar CLI local cost usage to summarize per-model usage for Codex or Claude, including the current (most recent) model or a full model breakdown. Trigger when asked for model-level usage/cost data from codexbar, or when you need a scriptable per-model summary from codexbar cost JSON.
Check Antigravity account quotas for Claude and Gemini models. Shows remaining quota and reset times with ban detection.
使用豆包(火山引擎)语音合成大模型 API 将文本转换为语音音频文件。支持声音复刻音色(S_ 开头的音色ID)和官方预置音色。当用户要求"语音合成"、"文字转语音"、"TTS"、"朗读文本"、"生成语音"、"用我的声音读"、"豆包语音"、"声音复刻合成"等相关请求时,务必使用此 skill。即使用户只是说"帮我把...
Intelligent model routing for sub-agent task delegation. Choose the optimal model based on task complexity, cost, and capability requirements. Reduces costs...
自动生成科技新闻摘要。从多个来源(RSS、Twitter、GitHub、Web Search)抓取科技新闻,整合后生成摘要。
Sync OpenRouter models used by OpenClaw into this installation's config. Fetches the OpenClaw app leaderboard from OpenRouter, verifies model IDs against the...