speech-notes将录音/语音转写为结构化演讲纪要。适用于:会议讲话、内部分享、演讲录音的转写整理。 触发条件:用户发送音频文件并要求整理/转写/纪要,或要求将已有转写文本整理成结构化纪要。
Install via ClawdBot CLI:
clawdbot install guoqunabc/speech-notesGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Generated Mar 20, 2026
Transcribes executive speeches or team discussions into structured notes for internal distribution. Ensures key decisions and action items are clearly documented without AI artifacts, suitable for companies needing formal records.
Converts recorded academic talks or training sessions into organized summaries. Preserves the speaker's first-person perspective and key insights, aiding students or professionals in review and study materials.
Processes audio from industry conferences or public speeches into polished notes. Highlights core themes and impactful quotes, useful for attendees or organizers to share highlights with stakeholders.
Transcribes media content into structured formats for publication or analysis. Maintains conversational tone while removing filler words, ideal for content creators needing readable summaries.
Converts recorded legal discussions or compliance updates into formal documents. Ensures accuracy and clarity while adhering to structured formatting, supporting regulatory documentation needs.
Offers tiered plans based on usage volume (e.g., hours of audio processed per month). Includes features like priority support and advanced formatting options, targeting businesses with regular transcription needs.
Charges per audio minute processed, with discounts for bulk usage. Appeals to occasional users or small teams, providing flexibility without long-term commitments.
Provides custom integrations with corporate tools like Feishu, along with dedicated support and security compliance. Targets large organizations needing scalable, secure transcription solutions.
💬 Integration Tip
Integrate with Feishu APIs for seamless document creation and updates, ensuring real-time collaboration and formatting consistency.
Scored Apr 19, 2026
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...
Create, manage, and deploy ElevenLabs conversational AI agents. Use when the user wants to work with voice agents, list their agents, create new ones, or manage agent configurations.
Voice note transcription and archival for OpenClaw agents. Powered by Deepgram Nova-3. Transcribes audio messages, saves both audio files and text transcript...
Text-to-Speech (TTS) and Speech-to-Text (ASR) using coze-coding-dev-sdk. Returns results directly to stdout.