jx-transcribeSpeech-to-text via SkillBoss API Hub (STT, powered by Whisper and more).
Install via ClawdBot CLI:
clawdbot install kirkraman/jx-transcribeGrade Limited — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://api.skillbossai.comAudited Apr 21, 2026 · audit v1.0
Generated Aug 10, 2026
Record team meetings, conferences, or interviews and automatically transcribe them for documentation and review. This saves time, improves record-keeping, and enables searchable archives of discussions.
Generate accurate captions and subtitles for videos to make content accessible to deaf or hard-of-hearing audiences and to comply with accessibility regulations. This expands audience reach and improves user engagement.
Transcribe customer service calls, support tickets, and voice feedback to analyze sentiment and identify common issues. This enables better customer insights and improves service quality.
Transcribe doctors' voice notes, patient consultations, and dictations to streamline clinical documentation. This reduces administrative burden and helps maintain accurate electronic health records.
Transcribe audio in various languages and use the translation feature to convert content into English. This assists learners in studying foreign language materials and aids in translation workflows.
Offer a cloud-based transcription service where users pay a monthly or annual fee for a certain number of transcription hours. This provides recurring revenue and scales with usage.
Expose the transcription API to developers and charge per minute or per character of audio processed. This allows businesses to integrate STT into their own applications while generating revenue based on volume.
Offer a limited free tier (e.g., 10 minutes of transcription per month) to attract users, then upsell paid plans with higher limits, advanced features like speaker diarization, or faster processing. This drives user acquisition and converts to paid.
💬 Integration Tip
For a quick start, wrap the pilot() function in a separate module and use environment variables for the API key. Consider error handling for network issues and API rate limits.
Scored Apr 21, 2026
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...
Create, manage, and deploy ElevenLabs conversational AI agents. Use when the user wants to work with voice agents, list their agents, create new ones, or manage agent configurations.
Voice note transcription and archival for OpenClaw agents. Powered by Deepgram Nova-3. Transcribes audio messages, saves both audio files and text transcript...
Text-to-Speech (TTS) and Speech-to-Text (ASR) using coze-coding-dev-sdk. Returns results directly to stdout.