toby-transcribeSpeech-to-text via SkillBoss API Hub (STT, powered by Whisper and more).
Install via ClawdBot CLI:
clawdbot install kirkraman/toby-transcribeGrade Limited — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://api.skillbossai.comAudited Apr 21, 2026 · audit v1.0
Generated May 8, 2026
Automatically transcribe business meetings, webinars, or conference calls to text for record-keeping and searchability. This reduces manual note-taking and ensures accurate documentation.
Transcribe customer support calls to analyze sentiment, identify common issues, and improve service quality. Enables automated quality assurance and training.
Generate written transcripts of podcast episodes or video content to improve SEO, create show notes, and make content accessible. Facilitates repurposing audio into blog posts or social media.
Translate audio from any language into English to aid language learners or cross-cultural communication. Supports education and global collaboration.
Charge users per audio minute or per transcription request. Ideal for low-volume, sporadic use. Revenue scales with usage.
Offer monthly or yearly plans with caps on transcription minutes. Encourage higher tiers with features like batch processing or priority support.
Integrate the API into existing SaaS platforms (e.g., CRM, project management) and charge a markup or bundle cost into the platform fee.
💬 Integration Tip
Simply set the SKILLBOSS_API_KEY environment variable and make a POST request to /v1/pilot with the audio encoded in base64. The response text is at result['result']['text'].
Scored May 8, 2026
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...
Create, manage, and deploy ElevenLabs conversational AI agents. Use when the user wants to work with voice agents, list their agents, create new ones, or manage agent configurations.
Voice note transcription and archival for OpenClaw agents. Powered by Deepgram Nova-3. Transcribes audio messages, saves both audio files and text transcript...
Text-to-Speech (TTS) and Speech-to-Text (ASR) using coze-coding-dev-sdk. Returns results directly to stdout.