azure-ai-voicelive-py面向Azure VoiceLive SDK的实时语音AI开发技能,覆盖WebSocket双向通信、流式音频输入输出、 实时转写、会话管理、VAD端点检测、函数调用工具集成与多语音模型选择。适用于构建语音助手、 客服对话、实时翻译、会议转写等场景。支持DefaultAzureCredential与API Key两种...
Install via ClawdBot CLI:
clawdbot install thegovind/azure-ai-voicelive-pyGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://cognitiveservices.azure.com/.defaultUses known external API (expected, informational)
azure.comAudited Apr 17, 2026 · audit v1.0
Generated Mar 21, 2026
Deploy a voice-enabled chatbot that handles customer inquiries over phone or voice interfaces, providing instant responses with natural speech. It can integrate with CRM systems to access customer data and resolve issues like billing questions or product support, reducing wait times and operational costs.
Enable real-time bidirectional translation during international video conferences or calls, allowing participants to speak in their native languages. This facilitates seamless communication in industries like tourism or remote collaboration, enhancing accessibility and productivity across language barriers.
Create animated avatars that respond with synchronized speech and expressions in gaming, virtual events, or educational apps. This enhances user engagement by providing a lifelike interactive experience, suitable for entertainment platforms or virtual training simulations.
Develop a system where patients can verbally report symptoms or receive medication reminders through a voice interface. It integrates with healthcare databases for real-time transcription and alerts, aiding in remote patient care and reducing manual data entry for medical staff.
Build applications that allow users to control smart devices like lights or thermostats using voice commands via WebSocket streaming. This provides hands-free operation, improving convenience and accessibility in home automation or industrial IoT settings.
Offer the voice AI application as a cloud-based service with tiered pricing based on usage metrics like API calls or concurrent sessions. This generates recurring revenue from businesses needing scalable, real-time voice solutions without upfront infrastructure costs.
Provide tailored development and integration services for enterprises to embed the SDK into existing systems like call centers or educational platforms. Revenue comes from project-based fees and ongoing support contracts, leveraging the skill's advanced features like function calling.
License the technology to original equipment manufacturers for embedding into hardware devices like smart speakers or medical devices. This creates revenue through upfront licensing fees and royalties per unit sold, capitalizing on the real-time audio streaming capabilities.
💬 Integration Tip
Use DefaultAzureCredential for secure authentication in production and handle audio events asynchronously to manage real-time streams efficiently.
Scored Apr 19, 2026
Use CodexBar CLI local cost usage to summarize per-model usage for Codex or Claude, including the current (most recent) model or a full model breakdown. Trigger when asked for model-level usage/cost data from codexbar, or when you need a scriptable per-model summary from codexbar cost JSON.
Check Antigravity account quotas for Claude and Gemini models. Shows remaining quota and reset times with ban detection.
Safely rotate the OpenRouter API key across all config files in an OpenClaw installation. Finds every location where the key is stored, updates them, restart...
Memory graph engine with caller-provided embed and LLM callbacks; core is pure, with real-time correction flow and optional OpenAI integration.
使用豆包(火山引擎)语音合成大模型 API 将文本转换为语音音频文件。支持声音复刻音色(S_ 开头的音色ID)和官方预置音色。当用户要求"语音合成"、"文字转语音"、"TTS"、"朗读文本"、"生成语音"、"用我的声音读"、"豆包语音"、"声音复刻合成"等相关请求时,务必使用此 skill。即使用户只是说"帮我把...
Intelligent model routing for sub-agent task delegation. Choose the optimal model based on task complexity, cost, and capability requirements. Reduces costs...