faster-whisper-local-serviceOpenClaw local speech-to-text backend using faster-whisper over HTTP on 127.0.0.1:18790. Use when you want voice transcription without external APIs, without...
Install via ClawdBot CLI:
clawdbot install neldar/faster-whisper-local-serviceGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Potentially destructive shell commands in tool definitions
rm -rf ~Calls external URL not in known-safe list
http://127.0.0.1:18790/transcribe`AI Analysis
The skill runs a local HTTP service bound to localhost only, with no evidence of sending data to unauthorized external servers. The primary risks are the inherent attack surface of the GStreamer media parser when processing untrusted audio files and the use of shell commands during deployment, but these are documented and confined to the local system.
Audited Apr 16, 2026 · audit v1.0
Generated Mar 1, 2026
Universities use this skill to transcribe lectures and seminars locally, ensuring student accessibility without recurring API costs. It supports multiple languages for international programs and operates offline after initial model download.
Healthcare providers deploy this service to transcribe patient consultations and medical notes securely on-premises, complying with data privacy regulations like HIPAA. It runs offline to protect sensitive voice data from external exposure.
Media companies integrate this skill to transcribe podcast episodes locally, enabling efficient editing, subtitling, and content repurposing. The offline operation eliminates API fees and supports high-volume audio processing with customizable models.
Businesses use this service to transcribe customer support calls in-house for quality assurance and training, leveraging local processing to reduce costs and enhance data security. It pairs with voice workflows for real-time or batch transcription.
Law firms implement this skill to transcribe legal depositions and meetings locally, ensuring confidentiality and avoiding third-party API dependencies. The service supports fixed languages for accuracy in legal terminology.
Sell a packaged solution with this skill for on-premises deployment, charging a one-time fee for setup and support. Revenue comes from customization, training, and optional updates, targeting industries with strict data sovereignty needs.
Offer the skill as free open-source software, generating revenue through paid support, consulting, and premium features like advanced model tuning. Monetize by assisting organizations with deployment and integration into their workflows.
Bundle this skill with dedicated hardware appliances for offline transcription, selling to sectors like healthcare or government. Revenue includes appliance sales, maintenance, and optional cloud backup services.
💬 Integration Tip
Ensure gst-launch-1.0 is installed from trusted OS packages to mitigate security risks from audio parsing, and pre-download models in air-gapped environments using faster-whisper documentation.
Scored Apr 22, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Turn your AI into JARVIS. Voice, wit, and personality — the complete package. Humor cranked to maximum.
Text-to-speech conversion using node-edge-tts npm package for generating audio from text. Supports multiple voices, languages, speed adjustment, pitch contro...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...