Logo
ClawHub Skills Lib
HomeCategoriesUse CasesTrendingStatisticsBlog
HomeCategoriesUse CasesTrendingStatisticsBlog
ClawHub Skills Lib
ClawHub Skills Lib

Browse 50.000+ community-built AI agent skills for OpenClaw. Updated daily from clawhub.ai.

Explore

  • Home
  • Categories
  • Use Cases
  • Trending
  • Blog

Categories

  • Development
  • AI & Agents
  • Productivity
  • Communication
  • Data & Research
  • Business
  • Platforms
  • Lifestyle
  • Education
  • Design

Use Cases

  • AI Code Generation
  • Code Review & Testing
  • DevOps & Cloud
  • Security & Compliance
  • Build an AI Agent
  • Agent Memory & RAG
  • Multi-Agent Orchestration
  • Browser & Web Automation
  • Financial & Market Data
  • Crypto & Web3
  • Real-Time Web Search
  • News & Media Monitoring
  • Academic Research
  • Data & Analytics
  • AI Image Generation
  • Voice & Audio AI
  • AI Video Creation
  • Content Writing
  • Task & Project Management
  • Knowledge Management
  • Email & Messaging
  • SEO & Content Marketing
  • Sales & CRM
  • Workflow Automation
  • Social Media
  • Chinese Platforms
  • E-Commerce
  • Education & Tutoring
  • HR & Recruiting
  • Legal & Compliance
  • AI Code Generation
  • Code Review & Testing
  • DevOps & Cloud
  • Security & Compliance
  • Build an AI Agent
  • Agent Memory & RAG
  • Multi-Agent Orchestration
  • Browser & Web Automation
  • Financial & Market Data
  • Crypto & Web3
  • Real-Time Web Search
  • News & Media Monitoring
  • Academic Research
  • Data & Analytics
  • AI Image Generation
  • Voice & Audio AI
  • AI Video Creation
  • Content Writing
  • Task & Project Management
  • See all use cases →
  • AI Code Generation
  • Code Review & Testing
  • DevOps & Cloud
  • Security & Compliance
  • Build an AI Agent
  • Agent Memory & RAG
  • Multi-Agent Orchestration
  • Browser & Web Automation
  • Financial & Market Data
  • See all use cases →
© 2026 ClawHub Skills Lib. All rights reserved.Built with Next.js · Neon · Prisma
Home/Use Cases/🎙️ Voice & Audio AI/🔊 Text-to-Speech

🔊 Text-to-Speech AI Skills

Convert text to natural-sounding speech with AI voice models and custom voice styles.

713 skillsPart of 🎙️ Voice & Audio AI
Lang:

713 skills found

Page 1 of 30

🎤Speech & Audio

AssemblyAI Transcriber

assemblyai-transcriber
xenofex7
v1.1.0
View Details

Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.

71
1.9k
4mo ago
🎤Speech & Audio

poocr vatinvoice2excel

poocr-vatinvoice2excel
coderwanfeng
v1.0.0
View Details

使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。

24
643
5mo ago
🎤Speech & Audio

ElevenLabs Agents

elevenlabs-agents
pennyroyaltea
v1.0.0
View Details

Create, manage, and deploy ElevenLabs conversational AI agents. Use when the user wants to work with voice agents, list their agents, create new ones, or manage agent configurations.

10
4k
2
4mo ago
🎤Speech & Audio

yaps-transcription

yaps-transcription
yaps
v1.0.2
View Details

Turn an audio or video file into text with Yaps. Good for interviews, podcasts, and voice memos. New users: install Yaps and sign in.

+1
1
167
4d ago
🎤Speech & Audio

yaps-dictation

yaps-dictation
yaps
v1.0.2
View Details

Type with your voice using Yaps on desktop. Get set up, fix a problem, or recover a recent dictation. New users: install Yaps and sign in.

+1
1
161
4d ago
🎤Speech & Audio

yaps-text-to-speech

yaps-text-to-speech
yaps
v1.0.2
View Details

Turn text into spoken audio with Yaps. Choose a supported voice and language, then save the file. New users: install Yaps and sign in.

+1
1
173
4d ago
🎤Speech & Audio

ifly-voiceclone-tts

ifly-voiceclone-tts
qingzhe2020
v1.0.0
View Details

iFlytek Voice Clone tts(声音复刻) — train a custom voice model from audio samples and synthesize speech with the cloned voice. Supports the full workflow: get tr...

0
892
4mo ago
🎤Speech & Audio

Whisper Transcribe

whisper-transcribe
JosunLP
v1.0.0
View Details

Transcribe audio files to text using OpenAI Whisper. Supports speech-to-text with auto language detection, multiple output formats (txt, srt, vtt, json), batch processing, and model selection (tiny to large). Use when transcribing audio recordings, podcasts, voice messages, lectures, meetings, or any audio/video file to text. Handles mp3, wav, m4a, ogg, flac, webm, opus, aac formats.

0
2.5k
3
7mo ago
🎤Speech & Audio

Openai Tts.Bak 2026 01 28T18:01:23+10:30

openai-tts-bak-2026-01-28t18-01-23-10-30
nicoataiza
v1.0.0
View Details

Text-to-speech via OpenAI Audio Speech API.

0
2.2k
1
7mo ago
🎤Speech & Audio

飞书语音

feishu-voice-lobster
godzff
v1.0.0
View Details

实现飞书语音消息的上传下载、语音转文字及文字转语音,支持与 ElevenLabs 语音服务集成。

0
1.5k
4mo ago
🎤Speech & Audio

Speech is Cheap Transcribe

asr
ilyakam
v1.2.0
View Details

Fast, affordable automatic speech-to-text transcription supporting 100 languages, speaker diarization, word timestamps, and customizable output formats.

0
3.9k
5
7mo ago
🎤Speech & Audio

🗣️ Edge-TTS Skill using uvx

edge-tts-uvx
al-one
v1.0.0
View Details

Text-to-speech conversion using `uvx edge-tts` for generating audio from text. Use when: (1) User requests audio/voice output with the "tts" trigger or keyword. (2) Content needs to be spoken rather than read (multitasking, accessibility, driving, cooking). (3) User wants a specific voice, speed, pitch, or format for TTS output.

0
3.7k
3
7mo ago
🎤Speech & Audio

Elevenlabs AI

elevenlabs-ai
codedao12
v1.0.0
View Details

Access ElevenLabs APIs for text-to-speech, speech-to-speech, realtime speech-to-text, voice/model management, and dialogue workflows with direct HTTP calls.

0
2.3k
7mo ago
🎤Speech & Audio

ifly-hyper-tts

ifly-hyper-tts
qingzhe2020
v1.0.0
View Details

讯飞超拟人语音合成 - 支持文本转语音、语音合成(发音人/语速/语调/音量/输出格式)。大模型语音合成技能。语音合成, 文字转语音, 超拟人, TTS. 用户指令如"把这段文案读出来"时使用此Skill。

0
761
5mo ago
🎤Speech & Audio

Sound FX

sound-fx
javicasper
v0.1.1
View Details

Generate short sound effects via ElevenLabs SFX (text-to-sound). Use when you need SFX clips like applause, canned laughter, whooshes, ambience, or short stingers, and optionally convert to WhatsApp-friendly .ogg/opus.

0
3k
1
7mo ago
🎤Speech & Audio

ElevenLabs

elevenlabs-api
byungkyu
v1.2.5
View Details

ElevenLabs API integration with managed authentication. AI-powered text-to-speech, voice cloning, sound effects, and audio processing. Use this skill when users want to generate speech from text, clone voices, create sound effects, or process audio. For other third party apps, use the api-gateway skill (https://clawhub.ai/byungkyu/api-gateway). Calls run through the `maton` CLI with OAuth login, or over raw HTTP with a Maton API key where the CLI cannot be installed. Every call is authenticated as the user's connection and reaches only what that connection's authorization allows, which the provider enforces on every request; the endpoints documented here are the ones this skill uses, and any other endpoint of this app needs the user to ask for it by name. Default to read and list calls, and confirm every write or new connection with the user. This file also documents the three constructs that turn an ElevenLabs connection into automation, in the order they are used: the connection (the first step), a hosted function that runs an ElevenLabs action through the Maton SDK, and a trigger that calls that function on a schedule or on an event. Those sections are the platform's own reference text, shared with the api-gateway skill, with ElevenLabs examples; they add no ElevenLabs capability - ElevenLabs is not an event source, a trigger cannot read ElevenLabs data, and the files under `references/<source>/triggers.md` are the platform's event catalogues for the sources Maton offers (time, Calendly, GitHub, Gmail, HubSpot, Linear, Notion, Slack, Stripe).

0
3.7k
3
5d ago
🎤Speech & Audio

Podcast Transcript Mining Authority Positioning

podcast-transcript-mining-authority-positioning
ncreighton
v1.0.0
View Details

Extract guest appearances, speaking topics, and soundbites from podcast transcripts to build authority portfolios and generate podcast pitch templates. Use w...

0
784
4mo ago
🎤Speech & Audio

Multimodal Base

yuyonghao-multimodal-base
yuyonghao-123
v0.1.0
View Details

Supports image understanding, OCR, speech-to-text, and text-to-speech synthesis with multi-voice and multimodal unified processing using OpenAI and Edge TTS.

0
682
4mo ago
🎤Speech & Audio

baiyin-voice-generate-skill

baiyin-voice-generate-skill
jiuping520
v1.0.3
View Details

使用百音开放平台创建 AI 语音任务,支持文本转语音、音色克隆,并在同一 skill 内继续查询任务状态和结果链接。用于用户要生成语音、克隆音色、查询语音任务进度或下载结果时。

0
778
4mo ago
🎤Speech & Audio

九马免费声音克隆

jiuma-free-voice-clone
dddcn1
v1.0.14
View Details

九马AI语音克隆技能,TTS。使用九马AI API进行语音克隆和合成,支持在线音色选择或自定义音频参考。当用户需要语音克隆、语音合成或选择不同音色时使用此技能。Jiuma AI voice cloning skill, TTS. Utilize Jiuma AI API for voice cloning and...

0
951
4mo ago
🎤Speech & Audio

Zyt TTS

zyt-tts-test
zuoyuting214
v0.6.0
View Details

Use Chanjing TTS API to convert text to speech by listing voices, creating synthesis tasks, and polling task status. This skill reads app_id and secret_key f...

0
793
4mo ago
🎤Speech & Audio

Voice Ai Integration

voice-ai-integration
HugoChaan
v0.1.5
View Details

Integrate Shengwang products: ConvoAI voice agents, RTC audio/video, RTM messaging, Cloud Recording, and token generation. Use when the user mentions Shengwa...

0
985
5mo ago
🎤Speech & Audio

Voice Recognition

smart-voice-recognition
08jacky04
v1.1.0
View Details

Intelligent speech-to-text using local OpenAI Whisper (no API key needed, fully private). Use when you need to transcribe audio files, convert voice messages...

0
595
4mo ago
🎤Speech & Audio

音频生成工具-专业版

dlazy-audio-tool-pro
v1.0.0
View Details

支持15+专业音频模型,包含TTS、语音克隆、多角色对话、原创音乐生成及管道自动化处理,满足内容团队需求。

0
170
7d ago
…

Other 🎙️ Voice & Audio AI Phases

📝
Transcription & STT
Transcribe audio and video files to text with speaker labels and timestamps.
🌍
Audio Translation
Translate spoken content across languages — transcribe, translate, and re-synthesize.
🎚️
Audio Processing
Clean audio, remove noise, separate vocals, and process audio files at scale.