Logo
ClawHub Skills Lib
HomeCategoriesUse CasesTrendingStatisticsBlog
HomeCategoriesUse CasesTrendingStatisticsBlog
ClawHub Skills Lib
ClawHub Skills Lib

Browse 50.000+ community-built AI agent skills for OpenClaw. Updated daily from clawhub.ai.

Explore

  • Home
  • Categories
  • Use Cases
  • Trending
  • Blog

Categories

  • Development
  • AI & Agents
  • Productivity
  • Communication
  • Data & Research
  • Business
  • Platforms
  • Lifestyle
  • Education
  • Design

Use Cases

  • AI Code Generation
  • Code Review & Testing
  • DevOps & Cloud
  • Security & Compliance
  • Build an AI Agent
  • Agent Memory & RAG
  • Multi-Agent Orchestration
  • Browser & Web Automation
  • Financial & Market Data
  • Crypto & Web3
  • Real-Time Web Search
  • News & Media Monitoring
  • Academic Research
  • Data & Analytics
  • AI Image Generation
  • Voice & Audio AI
  • AI Video Creation
  • Content Writing
  • Task & Project Management
  • Knowledge Management
  • Email & Messaging
  • SEO & Content Marketing
  • Sales & CRM
  • Workflow Automation
  • Social Media
  • Chinese Platforms
  • E-Commerce
  • Education & Tutoring
  • HR & Recruiting
  • Legal & Compliance
  • AI Code Generation
  • Code Review & Testing
  • DevOps & Cloud
  • Security & Compliance
  • Build an AI Agent
  • Agent Memory & RAG
  • Multi-Agent Orchestration
  • Browser & Web Automation
  • Financial & Market Data
  • Crypto & Web3
  • Real-Time Web Search
  • News & Media Monitoring
  • Academic Research
  • Data & Analytics
  • AI Image Generation
  • Voice & Audio AI
  • AI Video Creation
  • Content Writing
  • Task & Project Management
  • See all use cases →
  • AI Code Generation
  • Code Review & Testing
  • DevOps & Cloud
  • Security & Compliance
  • Build an AI Agent
  • Agent Memory & RAG
  • Multi-Agent Orchestration
  • Browser & Web Automation
  • Financial & Market Data
  • See all use cases →
© 2026 ClawHub Skills Lib. All rights reserved.Built with Next.js · Neon · Prisma
Home/Use Cases/🎙️ Voice & Audio AI/🎚️ Audio Processing

🎚️ Audio Processing AI Skills

Clean audio, remove noise, separate vocals, and process audio files at scale.

60 skillsPart of 🎙️ Voice & Audio AI
Lang:

60 skills found

Page 1 of 3

🎤Speech & Audio

audio-segmenter

audio-segmenter
wangminrui2022
v1.1.9
View Details

当用户想要**把长音频切成小段**、**音频切片**、**音频分割**、**把音频分成固定时长片段**、**制作语音数据集**、**准备Karaoke素材**、**翻唱音频切片**时自动触发。 支持单个音频文件或整个文件夹(支持递归),自动用 ffmpeg 把音频按指定秒数切成小片段,完美保留原始文件夹结构,并智能选择输出路径。 常见触发口语: - “帮我把这个音频切成60秒一段” - “把这个长音频分割成小段” - “音频切片,这个文件夹” - “把语音文件切成每段30秒” - “制作数据集,把音频切片” - “Karaoke素材切片” - “翻唱音频分割” - “把MP3切成小段” -

0
788
1mo ago
🎤Speech & Audio

purevocals-uvr-automator

purevocals-uvr-automator
wangminrui2022
v1.0.5
View Details

当用户想要**一键批量从音频文件中提取超干净纯人声(干声 / Vocals Only)**、去除伴奏/背景音乐时,自动调用此技能。 一键音频人声分离工具。专门从音频文件(.mp3/.wav/.flac等)中提取超干净干声(Acapella)或去除背景音制作伴奏。 核心用途:支持单个音频文件或整个文件夹批量处理(.mp3/.wav/.flac 等格式),输出高质量无杂音干声,自动在输入同级创建 [输入文件夹]_vocals 文件夹,完美保留原目录结构。 高频触发场景包括: - 翻唱练习、翻唱视频制作、B站/抖音/小红书演唱素材清洗 - 卡拉OK 伴奏制作(只保留人声) - 音乐制作中的人声分

0
783
1mo ago
🎤Speech & Audio

Pet Vocal Emotion Analysis Skill | 宠物叫声情绪解析技能

smyx-pet-vocal-emotion-analysis
smyx-sunjinhui
v1.0.9
View Details

Recognizes cat and dog barks through pet voiceprint AI, translates and outputs emotions and behavioral intentions such as happiness, excitement, anger, anxiety, pain, vigilance, and attention-seeking, enabling human-pet smart interaction. | 宠物叫声情绪解析技能,通过宠物声纹AI识别猫狗叫声,翻译输出开心、兴奋、愤怒、焦虑、痛苦、警惕、求关注等情绪与行为意图,实现人宠智能交互

0
1.1k
4
2d ago
🎤Speech & Audio

Whisper Transcribe

whisper-transcribe
JosunLP
v1.0.0
View Details

Transcribe audio files to text using OpenAI Whisper. Supports speech-to-text with auto language detection, multiple output formats (txt, srt, vtt, json), batch processing, and model selection (tiny to large). Use when transcribing audio recordings, podcasts, voice messages, lectures, meetings, or any audio/video file to text. Handles mp3, wav, m4a, ogg, flac, webm, opus, aac formats.

0
2.2k
3
5mo ago
🎤Speech & Audio

AudioPod

audiopod
Rakesh1002
v1.2.3
View Details

Use AudioPod AI's API for audio processing tasks including AI music generation (text-to-music, text-to-rap, instrumentals, samples, vocals), stem separation, text-to-speech, noise reduction, speech-to-text transcription, speaker separation, and media extraction. Use when the user needs to generate music/songs/rap from text, split a song into stems/vocals/instruments, generate speech from text, clean up noisy audio, transcribe audio/video, or extract audio from YouTube/URLs. Requires AUDIOPOD_API_KEY env var or pass api_key directly.

0
3.9k
3
5mo ago
🎤Speech & Audio

Luis Audio Translator

luis-audio-translator
luis1213899
v1.0.0
View Details

Convert, compress, merge, split, clip, inspect, and extract audio locally with FFmpeg, plus decode supported music-cache formats including pure-Python Ximala...

0
271
2mo ago
🎤Speech & Audio

AI Stem Splitter

ai-stem-splitter
codeugar
v1.0.0
View Details

Use when the user wants to split a song or audio file into vocals, drums, bass, guitar, piano, and other stems; remove vocals for karaoke; extract acapellas...

+2
0
368
2mo ago
🎤Speech & Audio

Telegram Voice To Voice Macos

telegram-voice-to-voice-macos
fiberian1981
v0.1.3
View Details

Telegram voice-to-voice for macOS Apple Silicon: transcribe inbound .ogg voice notes with yap (Speech.framework) and reply with Telegram voice notes via say+ffmpeg. Not compatible with Linux/Windows.

+1
0
2.5k
2mo ago
🎤Speech & Audio

audiobook

pg-essay-to-audiobook-audiobook
lnj22
v0.1.0
View Details

Create audiobooks from web content or text files. Handles content fetching, text processing, and TTS conversion with automatic fallback between ElevenLabs, O...

0
434
2mo ago
🎤Speech & Audio

智能音频分离工具

sound-split
nabian1990amber-cmd
v1.0.0
View Details

智能音频分离工具,一键将任意音频或视频分离出人声、伴奏、鼓、贝斯、钢琴等独立音轨。适用于音乐人翻唱伴奏提取、歌曲 Remix 制作、播客人声降噪、视频配乐替换、音乐教学素材准备等场景。当用户需要「分离人声和伴奏」「提取伴奏」「去除人声」「拆分音轨」「vocal split」「stem splitter」等操作时触...

0
543
2mo ago
🎤Speech & Audio

Music Cog

music-generation-cellcog
nitishgargiitd
v1.0.10
View Details

AI music generation powered by CellCog. Original instrumental and vocal tracks, 5 seconds to 10 minutes. Cinematic scores, background tracks, podcast intros,...

0
4.9k
4
20d ago
🎤Speech & Audio

Virtual voice builder

virtual-voice-ai
suhas12345685-pro
v1.0.0
View Details

Wires a real microphone through an AI brain (STT → LLM → TTS) and routes the output to a virtual audio cable so apps like Google Meet hear the processed voic...

0
714
2mo ago
🎤Speech & Audio

ElevenLabs AI Music Generation — Pro Pack on RunComfy

elevenlabs-music-generation
kalvinrv
v1.0.0
View Details

Generate full songs and instrumental tracks with ElevenLabs Music on RunComfy via the `runcomfy` CLI. ElevenLabs Music turns a style description plus structured lyrics into studio-quality 44.1 kHz stereo audio — 5 seconds to 5 minutes — with section-level control (Intro / Verse / Chorus / Bridge), multilingual vocals, and commercial-friendly output. Generate a backing track, a full vocal song, a jingle, a podcast intro, a game loop, or an instrumental bed. Calls `runcomfy run elevenlabs/elevenlabs/music-generation` through the local RunComfy CLI. Triggers on "generate music", "make a song", "AI music", "background music", "instrumental track", "ElevenLabs Music", "soundtrack", "jingle", "theme music", "royalty-free music", "compose", or any explicit ask to generate music or a song from a text description.

0
34
9d ago
🎤Speech & Audio

Voice Transcriber Toolkit

voice-transcriber-toolkit
kaiyuelv
v1.0.0
View Details

Voice-to-Text Transcription Toolkit - 语音识别转文字,支持Whisper/Vosk引擎,批量处理,字幕导出 | Speech recognition & transcription with Whisper/Vosk engines, batch processing, su...

0
471
2mo ago
🎤Speech & Audio

Multimodal Base

yuyonghao-multimodal-base
yuyonghao-123
v0.1.0
View Details

Supports image understanding, OCR, speech-to-text, and text-to-speech synthesis with multi-voice and multimodal unified processing using OpenAI and Edge TTS.

0
475
2mo ago
🎤Speech & Audio

Kie Audio Generator

kie-audio
benhuebner01
v1.0.0
View Details

Generate music and audio via Kie.ai's Suno gateway (V3.5 through V5.5). Use for background tracks, instrumental beds, full songs with vocals, or extending ex...

0
456
2mo ago
🎤Speech & Audio

mmVoiceMaker

mm-voice-maker
blue-coconut
v1.0.1
View Details

Enables voice synthesis, voice cloning, voice design, and audio post-processing using MiniMax Voice API and FFmpeg. Use when converting text to speech, creat...

0
1.3k
3
3mo ago
🎤Speech & Audio

飞书TTS Pro

feishu-tts-pro
minybear
v1.0.0
View Details

将文字转换为语音并以飞书语音气泡消息发送。使用 Edge-TTS(微软,免费,无限次)生成中文语音,FFmpeg 转码为 ogg/opus,上传到飞书作为 audio 类型消息。支持自定义音色、接收者和 Python 环境。当用户要求用语音回复、发语音消息、TTS 朗读内容时触发。

0
429
2mo ago
🎤Speech & Audio

macOS Local Voice

macos-local-voice
strrl
v1.0.0
View Details

Local STT and TTS on macOS using native Apple capabilities. Speech-to-text via yap (Apple Speech.framework), text-to-speech via say + ffmpeg. Fully offline, no API keys required. Includes voice quality detection and smart voice selection.

0
5.1k
1
2mo ago
🎤Speech & Audio

Evolink Music — AI Music Generation (Suno v4/v4.5/v5)

evolink-music
EvoLinkAI
v2.0.0
View Details

AI music generation with Suno v4, v4.5, v5. Text-to-music, custom lyrics, instrumental, vocal control. 5 models, one API key.

0
1.4k
2
3mo ago
🎤Speech & Audio

Audio

audio
ivangdavila
v1.0.1
View Details

Process, enhance, and convert audio files with noise removal, normalization, format conversion, transcription, and podcast workflows.

0
5k
2
2mo ago
🎤Speech & Audio

虾转音频

xia-zhuan-audio
luis1213899
v1.3.1
View Details

🎵 音视频格式转换与处理工具箱。基于 FFmpeg + Whisper AI,支持:格式转换、视频提取音频、合并、分割、压缩、查看信息、音频转文字。

0
659
1
2mo ago
🎤Speech & Audio

Audio Command Executor

audio-command-executor
sirkovz
v1.0.1
View Details

Processes inbound audio files, transcribes them, and answers to resulting texts. Converts non-WAV inputs to WAV before transcription.

0
563
2mo ago
🎤Speech & Audio

unisound-treatment-process

unisound-treatment-process
unisound-llm
v1.0.2
View Details

根据病程记录生成诊疗经过。输入病程记录文本,调用内部医疗大模型,输出结构化诊疗经过文本。

0
450
20d ago

Other 🎙️ Voice & Audio AI Phases

🔊
Text-to-Speech
Convert text to natural-sounding speech with AI voice models and custom voice styles.
📝
Transcription & STT
Transcribe audio and video files to text with speaker labels and timestamps.
🌍
Audio Translation
Translate spoken content across languages — transcribe, translate, and re-synthesize.