alicloud-ai-audio-asrTranscribe non-realtime speech with Alibaba Cloud Model Studio Qwen ASR models (`qwen3-asr-flash`, `qwen-audio-asr`, `qwen3-asr-flash-filetrans`). Use when c...
Install via ClawdBot CLI:
clawdbot install cinience/alicloud-ai-audio-asrGrade Limited — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Sends data to undocumented external endpoint (potential exfiltration)
POST → https://dashscope.aliyuncs.com/compatible-mode/v1/chat/completions`Calls external URL not in known-safe list
https://dashscope.aliyuncs.com/compatible-mode/v1/chat/completionsAI Analysis
The skill's external API call (dashscope.aliyuncs.com) is explicitly documented as the official Alibaba Cloud Model Studio endpoint for the stated transcription purpose, with no hidden instructions or credential harvesting patterns. The data transmission is required for the core functionality and is consistent with the skill's description.
Audited Apr 18, 2026 · audit v1.0
Generated Mar 21, 2026
Transcribe recorded university lectures or seminars into text for creating accessible study materials and notes. This supports students and educators by enabling easy review and searchability of spoken content.
Convert recorded customer service calls into transcripts for quality assurance, training, and compliance documentation. Helps businesses analyze interactions and improve service standards.
Generate timed transcripts for videos or podcasts to create accurate subtitles or captions. Facilitates content localization and accessibility for diverse audiences.
Transcribe audio recordings of legal depositions or meetings into text for official records and case preparation. Ensures precise documentation and easy reference for legal teams.
Convert audio from patient consultations into structured text notes for electronic health records. Aids healthcare providers in maintaining accurate and efficient documentation.
Offer a subscription-based platform where users upload audio files via API for automated transcription. Revenue is generated through tiered pricing based on usage volume and features like timestamping.
License the ASR skill to large organizations for embedding into their internal systems, such as call centers or content management tools. Revenue comes from one-time licensing fees and annual support contracts.
Provide API access where customers pay per minute of audio transcribed, with additional charges for advanced features like async long-file processing. Targets developers and businesses with variable transcription needs.
💬 Integration Tip
Ensure the DASHSCOPE_API_KEY is securely set in environment variables and use the bundled Python script for quick testing with local or URL audio inputs.
Scored Apr 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Secure, offline, OpenAI-compatible local Whisper ASR endpoint for OpenClaw. Features faster-whisper (large-v3-turbo), built-in privacy with no cloud telemetr...
Turn your AI into JARVIS. Voice, wit, and personality — the complete package. Humor cranked to maximum.
Text-to-speech conversion using node-edge-tts npm package for generating audio from text. Supports multiple voices, languages, speed adjustment, pitch contro...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。