any-whisper-apiTranscribe audio via API Whisper with any compatible local servers.
Install via ClawdBot CLI:
clawdbot install nw15d/any-whisper-apiRequires:
Grade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://platform.openai.com/docs/guides/speech-to-textUses known external API (expected, informational)
api.openai.comAudited Apr 18, 2026 · audit v1.0
Generated Mar 21, 2026
Healthcare professionals can use this skill to transcribe patient consultations and medical dictations into text for electronic health records. It supports language specification for multilingual patients and prompt customization for medical terminology, improving documentation accuracy and efficiency.
Podcast creators can automate the transcription of audio episodes to generate show notes, subtitles, and searchable text content. The JSON output option facilitates integration with content management systems, enhancing accessibility and SEO for podcast platforms.
Law firms can transcribe audio recordings from depositions and client meetings into text for case files and legal briefs. The skill allows for prompt usage to specify speaker names, ensuring accurate attribution and streamlining legal documentation processes.
Businesses can transcribe customer support calls to analyze sentiment, identify common issues, and train agents. The ability to use local servers ensures data privacy for sensitive customer interactions, supporting compliance with regulations like GDPR.
Educational institutions can transcribe lectures and seminars to create accessible captions for students, including those with hearing impairments. The skill supports various audio formats and language settings, making it versatile for multilingual academic environments.
Offer a cloud-based transcription service using this skill to process audio files for clients on a subscription or pay-per-use basis. Revenue is generated through tiered pricing plans based on usage volume and additional features like JSON output or custom prompts.
Sell customized deployments of this skill integrated with local Whisper servers to enterprises requiring high data security, such as in healthcare or finance. Revenue comes from licensing fees, setup costs, and ongoing maintenance and support contracts.
Provide consulting services to help businesses integrate this skill into their existing workflows, such as CRM or content management systems. Revenue is generated through project-based fees, training sessions, and ongoing technical support for optimization.
💬 Integration Tip
Ensure WHISPER_API_KEY and WHISPER_API_HOST are set in environment variables or the clawdbot.json config file for seamless authentication with local or cloud Whisper servers.
Scored Apr 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Search and manage Spotify playlists, tracks, albums, artists, and playback state via the Spotify Web API. Use this skill when users want to search for music,...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Transcribe long audio files safely on 16GB RAM machines using auto-chunking with Whisper’s base model and seamless transcript merging.
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...