voiceclaw-jpVoice conversation interface for OpenClaw using wake word detection, streaming LLM responses, and text-to-speech. Use when a user wants to talk to their Open...
Install via ClawdBot CLI:
clawdbot install kentoku24/voiceclaw-jpGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://voicevox.hiroshiba.jp/Audited Apr 17, 2026 · audit v1.0
Generated Mar 21, 2026
Users can control smart home devices and get information through natural voice conversations with their OpenClaw agent. This enables hands-free operation for tasks like adjusting lights, checking weather, or setting reminders, ideal for busy households or accessibility needs.
Businesses can deploy this skill to provide voice-based customer support on their websites or kiosks. It allows customers to ask questions and receive spoken responses, reducing wait times and improving engagement for inquiries like product details or troubleshooting.
Language learners can practice speaking and listening in Japanese by conversing with the AI agent. The skill supports interactive dialogues, pronunciation feedback via TTS, and customizable wake words, making it useful for self-study or classroom environments.
In healthcare settings, patients can use voice commands to interact with an OpenClaw agent for medication reminders, symptom reporting, or accessing health information. The low-latency streaming ensures timely responses, aiding in patient support and monitoring.
Employees in offices can use this skill for voice-activated tasks like scheduling meetings, retrieving documents, or generating reports through OpenClaw. It integrates with existing workflows to boost efficiency and reduce manual input in fast-paced work environments.
Offer this voice interface as a cloud-based service with tiered pricing based on usage, such as number of voice interactions or custom wake words. Businesses pay monthly fees to integrate it into their customer platforms, generating recurring revenue from enhanced user experiences.
License the skill to smart device makers (e.g., speakers, kiosks) for embedding voice capabilities into their products. Charge upfront fees or royalties per unit sold, leveraging the open-source base for customization and scalability in hardware integrations.
Provide a free version with basic features for individual developers and hobbyists, while offering premium features like advanced wake word customization, priority support, or additional TTS voices for a fee. This attracts a community and monetizes power users.
💬 Integration Tip
Ensure OpenClaw gateway is running locally with chatCompletions enabled, and test VOICEVOX connectivity on port 50021 before deployment to avoid audio playback issues.
Scored Apr 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Search and manage Spotify playlists, tracks, albums, artists, and playback state via the Spotify Web API. Use this skill when users want to search for music,...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Transcribe long audio files safely on 16GB RAM machines using auto-chunking with Whisper’s base model and seamless transcript merging.
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...