luis-audio-translatorConvert, compress, merge, split, clip, inspect, and extract audio locally with FFmpeg, plus decode supported music-cache formats including pure-Python Ximala...
Install via ClawdBot CLI:
clawdbot install luis1213899/luis-audio-translatorGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://clawhub.ai/user/luis1213899Audited Jun 5, 2026 · audit v1.0
Generated Oct 7, 2026
Podcast producers use the skill to batch-convert raw interview recordings from various formats into standardized MP3 or WAV files, extract audio from video calls, and split long episodes into segments for editing. They also inspect metadata to ensure consistent loudness and codec settings before publishing.
Audiophiles and DJs use the skill to compress large FLAC collections to portable MP3s, merge tracks for continuous mixes, and decode encrypted cache files from streaming services like Ximalaya and Kugou. The tool runs locally, preserving privacy and avoiding cloud uploads.
Language teachers extract audio from video lessons, clip specific dialogues, and convert them to mobile-friendly formats like M4A for students. They also split long recordings into short segments for vocabulary drills and listening exercises.
Law firms use the skill to convert and clip audio from interviews, body cams, or surveillance footage into standard formats for court submissions. Metadata inspection ensures compliance with evidentiary standards, and local processing maintains chain of custody.
Game developers convert sound effects and voice lines between formats like WAV and OGG, merge ambient tracks, and split long recordings into individual cues. The skill's batch capabilities streamline integration into game engines.
A desktop application offers basic audio conversion and merging for free, while advanced features like Ximalaya/Kugou decryption, batch processing, and priority support require a paid license. The skill serves as the core engine, with a GUI wrapper for non-technical users.
Businesses license the skill as a local API service that integrates with their existing workflows, such as podcast production platforms or legal case management systems. Pricing is based on usage volume or number of seats, with enterprise support contracts.
Streaming services and social media platforms embed the skill as an SDK to allow users to process audio locally before uploading, reducing server costs. Revenue comes from per-integration fees and revenue sharing on premium features.
💬 Integration Tip
Resolve the script path relative to SKILL.md and use environment variables for FFmpeg/FFprobe paths to ensure portability. For advanced decryption, run diagnose first to verify optional helper binaries are available.
Scored Oct 7, 2026
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...
Create, manage, and deploy ElevenLabs conversational AI agents. Use when the user wants to work with voice agents, list their agents, create new ones, or manage agent configurations.
Voice note transcription and archival for OpenClaw agents. Powered by Deepgram Nova-3. Transcribes audio messages, saves both audio files and text transcript...
Text-to-Speech (TTS) and Speech-to-Text (ASR) using coze-coding-dev-sdk. Returns results directly to stdout.