webchat-voice-full-stackOne-step full-stack installer for OpenClaw WebChat voice input with local speech-to-text. Orchestrates three focused skills in order: local STT backend (fast...
Install via ClawdBot CLI:
clawdbot install neldar/webchat-voice-full-stackGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Generated Mar 20, 2026
Enables customer service agents to use voice input for faster ticket logging and response in web-based support dashboards. It reduces typing effort and supports multilingual agents with localized UI, improving efficiency in high-volume environments.
Allows medical professionals to transcribe patient notes locally via voice in electronic health record systems, ensuring data privacy with no external API calls. The push-to-talk feature helps maintain accuracy during consultations without interrupting workflow.
Integrates voice input for students and instructors in online learning platforms to facilitate interactive discussions and note-taking. The local STT backend avoids recurring costs, making it scalable for institutions with budget constraints.
Enhances team collaboration tools by adding voice transcription for meeting notes and chat inputs in web applications. The HTTPS/WSS proxy ensures secure communication, while the full-stack setup simplifies deployment for distributed teams.
Provides voice input capabilities to make web applications more accessible for users with mobility or typing difficulties. The local deployment ensures low latency and privacy, with i18n support catering to diverse user bases globally.
Offer paid consulting and customization services for businesses integrating this skill into their existing web platforms. Revenue comes from one-time setup fees and ongoing maintenance contracts, leveraging the no-recurring-API-cost advantage.
Develop a managed platform that bundles this skill with other tools for easy deployment in cloud environments. Charge subscription fees for hosting, monitoring, and updates, targeting small to medium enterprises seeking voice-enabled solutions.
Partner with hardware manufacturers to pre-install this skill on devices like kiosks or specialized workstations for industries like healthcare or retail. Revenue is generated through product sales and licensing fees for the integrated software stack.
💬 Integration Tip
Ensure all prerequisites like Python 3.10+ and GStreamer are installed before deployment, and use the integrity verification feature to prevent script tampering during updates.
Scored Apr 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Turn your AI into JARVIS. Voice, wit, and personality — the complete package. Humor cranked to maximum.
Text-to-speech conversion using node-edge-tts npm package for generating audio from text. Supports multiple voices, languages, speed adjustment, pitch contro...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...