Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio and video files to text with speaker labels and timestamps.
366 skills found
Page 1 of 16
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Create, manage, and deploy ElevenLabs conversational AI agents. Use when the user wants to work with voice agents, list their agents, create new ones, or manage agent configurations.
Extract audio from video URLs and transcribe using STT (Speech-to-Text). Supports local Whisper or cloud APIs. Use when: user provides a video URL and wants...
Simple local Speech-To-Text using Whisper. One-command install with auto model download. Supports 99+ languages.
Convert text, scripts, and captions into natural voiceovers for videos, explainers, product demos, and social posts.
可以拨打中国电话号码的机器人外呼, 专为openclaw(龙虾)用户打造的专业ai呼叫能力,只要一个prompt就可以帮你打电话干活了,支持查看电话对话记录,查看电话状态等。
Speech-to-text via SkillBoss API Hub (STT, powered by Whisper and more).
groq-whisper-apiTranscribe audio via Groq Automatic Speech Recognition (ASR) Models (Whisper).
Local speech-to-text with the Whisper CLI (no API key).
All-in-one voice identity toolkit: speaker identification, voice library management, voice cloning, and speech-to-text. The only OpenClaw skill with speaker...
Push text notifications to Windows Azure TTS service for audio broadcast via Bluetooth speakers. Perfect for family reminders, alarms, and announcements.
Text-to-Speech via Zvukogram API with SSML support. Use when you need to generate speech from text, create podcasts, voice notifications, or work with audio....
语音回复技能 - 每次回复自动生成语音并保存到桌面,支持 Noiz AI TTS
Local speech-to-text using Qwen3-ASR (CPU-only, no API key, no cloud). Use when: (1) a voice message or audio file needs transcription, (2) user asks to tran...
Speech-to-text via SkillBoss API Hub (STT, powered by Whisper and more).
个性化资讯电台生成服务。使用场景:(1) 生成特定主题的电台,(2) 设置每日定时推送,(3) 配置TTS音色,(4) 收听历史电台。不适用:音乐播放、实时广播、视频内容。
Speech-to-text via SkillBoss API Hub (STT, powered by Whisper and more).
Receive and engage with transcribed voice memos from Yapp, a voice journaling app, capturing raw, unedited speech-to-text recordings with metadata.
Text-To-Speech with MLX (Apple Silicon) and opensource models (default QWen3-TTS) locally.
Create chapters, highlights, and show notes from podcast audio or transcripts. Use when a user wants chapter markers, highlight clips, or show-note drafts without publishing or distribution actions.
QQ Bot 语音消息自动识别 v2.0。自动解码 QQ Silk V3 格式,Whisper medium 模型识别,Gateway 集成,用户确认流程。
One-step full-stack installer for OpenClaw WebChat voice input with local speech-to-text. Orchestrates three focused skills in order: local STT backend (fast...
lessac-offline-voice-systemComplete offline voice system with high-quality Lessac TTS and faster-whisper speech recognition. Provides natural voice conversations without internet. Use...