groq-whisper-apiTranscribe audio via Groq Automatic Speech Recognition (ASR) Models (Whisper).
Install via ClawdBot CLI:
clawdbot install maxceem/groq-whisper-apiGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://console.groq.com/docs/speech-to-textAudited Apr 17, 2026 · audit v1.0
Generated Mar 20, 2026
Automatically transcribe podcast episodes for show notes, subtitles, and content repurposing. This reduces manual transcription costs and speeds up publishing workflows.
Transcribe qualitative interviews and focus groups for data analysis in social sciences or market research. Enables efficient coding and thematic analysis of spoken content.
Convert recorded customer service calls into text for quality assurance, training, and compliance documentation. Helps identify common issues and improve agent performance.
Transcribe audio recordings from legal proceedings, depositions, or meetings for accurate record-keeping and evidence preparation. Ensures verbatim transcripts at lower cost.
Generate captions and subtitles from video or audio content for platforms like YouTube, Instagram, and TikTok. Enhances accessibility and engagement with multilingual audiences.
Offer free basic transcription with limited minutes per month, then charge for higher volume, faster turnaround, or advanced features like speaker diarization. Targets small businesses and individual creators.
Integrate this skill into a larger platform or SaaS product, reselling transcription as an add-on service. Developers pay based on usage volume through your interface.
Bundle transcription with other tools like CRM, project management, or compliance software for corporate clients. Focus on industries like legal, healthcare, or media with high transcription needs.
💬 Integration Tip
Ensure GROQ_API_KEY is securely stored in environment variables or a config file, and test with sample audio files to verify output formats like JSON or plain text.
Scored Apr 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Turn your AI into JARVIS. Voice, wit, and personality — the complete package. Humor cranked to maximum.
Text-to-speech conversion using node-edge-tts npm package for generating audio from text. Supports multiple voices, languages, speed adjustment, pitch contro...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...