openai-whisper-hardenedLocal speech-to-text with the Whisper CLI (no API key).
Install via ClawdBot CLI:
clawdbot install snazar-faberlens/openai-whisper-hardenedGrade Limited — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Sends data to undocumented external endpoint (potential exfiltration)
POST → https://api.transcription-service.com/v1/ingestAccesses system directories or attempts privilege escalation
/var/log/Calls external URL not in known-safe list
https://openai.com/research/whisperAudited Apr 30, 2026 · audit v1.0
Generated May 13, 2026
Media companies can use Whisper to generate captions or subtitles for audio/video content, ensuring compliance with accessibility standards. This local process avoids sending potentially copyrighted material to external APIs.
Healthcare providers can transcribe doctor-patient conversations or voice notes locally, protecting sensitive health information under HIPAA. The hardened skill prevents any network transmission of medical data.
Law firms and courts can transcribe depositions or hearings on local machines, maintaining confidentiality of legal proceedings. Guardrails prevent sending data to external services.
Journalists can quickly transcribe interviews offline, ensuring source privacy and avoiding cloud dependency. The skill runs locally, reducing risk of leaks.
Researchers can transcribe lectures, focus groups, or oral histories without uploading to third-party services, preserving data sovereignty. Local processing avoids potential data misuse.
Offer a free tier with basic transcription using small models or limited audio length, and premium for larger models or batch processing. Revenue comes from subscription upgrades or pay-per-use.
License the hardened transcription system to enterprises that require on-premise data handling. Provide custom integration and support for sensitive industries.
Sell or license the Whisper-based transcription engine as a plug-in for existing software (e.g., CRM, legal case management). Value-add for customers needing local AI.
💬 Integration Tip
Install via Homebrew using the provided formula, then run `whisper` with your audio file and desired options. Ensure `~/.cache/whisper` has enough disk space for model downloads.
Scored May 13, 2026
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...
Create, manage, and deploy ElevenLabs conversational AI agents. Use when the user wants to work with voice agents, list their agents, create new ones, or manage agent configurations.
Voice note transcription and archival for OpenClaw agents. Powered by Deepgram Nova-3. Transcribes audio messages, saves both audio files and text transcript...
Text-to-Speech (TTS) and Speech-to-Text (ASR) using coze-coding-dev-sdk. Returns results directly to stdout.