voice-memo-syncSync, transcribe, and intelligently organize voice memos, audio/video files, and URLs. 同步、转录、智能整理语音备忘录、音视频文件和视频链接。
Install via ClawdBot CLI:
clawdbot install ying-wen/voice-memo-syncGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://github.com/ying-wen/voice-memo-syncAudited Apr 17, 2026 · audit v1.0
Generated Mar 22, 2026
Professionals and students can record meetings or lectures, upload audio files, and receive transcribed, organized notes with key points and action items. This saves time on manual note-taking and ensures accurate records for review and sharing.
Content creators can input YouTube or Bilibili URLs to automatically transcribe videos, extract insights, and generate structured summaries. This aids in repurposing content for blogs, social media, or subtitles, enhancing workflow efficiency.
Individuals use the skill to sync voice memos from devices, transcribe them into text, and organize with tags and reminders. It helps capture ideas, to-dos, and reflections, integrating with Apple Notes for easy access and management.
Professionals in law or healthcare can upload audio recordings of client consultations or patient notes for transcription and analysis. The skill ensures secure, local processing and creates searchable records for compliance and reference.
Researchers can process interviews, focus groups, or field recordings by transcribing audio/video files and using LLM analysis to identify themes and insights. This streamlines qualitative data analysis and report generation.
Offer a free basic version with limited transcriptions and storage, then charge for premium features like advanced LLM processing, higher accuracy, and integration with cloud services. Revenue comes from monthly subscriptions and enterprise plans.
License the skill to businesses for internal use, providing custom integrations with existing tools like CRM or project management software. Revenue is generated through one-time setup fees and annual maintenance contracts.
Expose transcription and processing capabilities as an API for developers to build into their applications. Monetize through API usage tiers, with revenue based on the number of requests or processing minutes consumed.
💬 Integration Tip
Ensure all required binaries like ffmpeg and whisper-cpp are installed for optimal performance, and configure the skill to auto-detect GPU acceleration on Apple Silicon for faster processing.
Scored Jun 19, 2026
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...
Create, manage, and deploy ElevenLabs conversational AI agents. Use when the user wants to work with voice agents, list their agents, create new ones, or manage agent configurations.
Voice note transcription and archival for OpenClaw agents. Powered by Deepgram Nova-3. Transcribes audio messages, saves both audio files and text transcript...
Text-to-Speech (TTS) and Speech-to-Text (ASR) using coze-coding-dev-sdk. Returns results directly to stdout.