byted-las-asr-proASR / STT / speech recognition / voice recognition engine powered by Volcengine LAS. Transcribes and converts speech to text from audio and video files — ext...
Install via ClawdBot CLI:
clawdbot install volcengine-skills/byted-las-asr-proGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Sends data to undocumented external endpoint (potential exfiltration)
POST → https://operator.las.cn-beijing.volces.com/api/v1/submit`Calls external URL not in known-safe list
https://operator.las.cn-beijing.volces.com/api/v1/submit`AI Analysis
The skill's external API calls are explicitly documented and directly serve its stated purpose of audio transcription, with no evidence of credential harvesting, obfuscation, or hidden instructions. The primary risk is the standard data privacy consideration of sending user audio to a third-party service, which is disclosed.
Audited Apr 17, 2026 · audit v1.0
Generated Jul 11, 2026
Transcribes meeting recordings, identifies each speaker, and generates meeting notes or minutes. Useful for teams that need accurate records of discussions and action items.
Converts audio from podcasts or lectures into text for subtitles, show notes, or accessibility. Supports multiple languages and can handle long-form content.
Transcribes phone calls and customer service interactions, then performs sentiment analysis and emotion recognition to gauge customer satisfaction and agent performance.
Generates accurate subtitles for videos in multiple languages. Ideal for content creators, filmmakers, and online education platforms seeking to improve accessibility.
Transcribes interviews, dictation, or any recorded speech into text quickly. Useful for journalists, researchers, and professionals who need written records of spoken content.
Charge per second of audio processed, with different rates based on model complexity (e.g., basic vs. big model). Customers submit jobs via API and pay for actual usage.
Offer the skill as an integrated feature within existing platforms (e.g., video conferencing, e-learning). License fee per user or per API call.
Provide bulk transcription services for enterprises with large volumes of audio data (e.g., call centers, media archives). Custom pricing based on monthly quotas.
💬 Integration Tip
Set up LAS_API_KEY environment variable and ensure your input audio format is a container format (wav, mp3, m4a). Use the async submit-poll workflow for long audio files.
Scored Jul 11, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Secure, offline, OpenAI-compatible local Whisper ASR endpoint for OpenClaw. Features faster-whisper (large-v3-turbo), built-in privacy with no cloud telemetr...
Turn your AI into JARVIS. Voice, wit, and personality — the complete package. Humor cranked to maximum.
Text-to-speech conversion using node-edge-tts npm package for generating audio from text. Supports multiple voices, languages, speed adjustment, pitch contro...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。