voice-memo-syncSync, transcribe, and intelligently organize voice memos, audio/video files, and URLs. 同步、转录、智能整理语音备忘录、音视频文件和视频链接。
Install via ClawdBot CLI:
clawdbot install ying-wen/voice-memo-syncGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://github.com/ying-wen/voice-memo-syncAudited Apr 17, 2026 · audit v1.0
Generated Mar 22, 2026
Professionals and students can record meetings or lectures, upload audio files, and receive transcribed, organized notes with key points and action items. This saves time on manual note-taking and ensures accurate records for review and sharing.
Content creators can input YouTube or Bilibili URLs to automatically transcribe videos, extract insights, and generate structured summaries. This aids in repurposing content for blogs, social media, or subtitles, enhancing workflow efficiency.
Individuals use the skill to sync voice memos from devices, transcribe them into text, and organize with tags and reminders. It helps capture ideas, to-dos, and reflections, integrating with Apple Notes for easy access and management.
Professionals in law or healthcare can upload audio recordings of client consultations or patient notes for transcription and analysis. The skill ensures secure, local processing and creates searchable records for compliance and reference.
Researchers can process interviews, focus groups, or field recordings by transcribing audio/video files and using LLM analysis to identify themes and insights. This streamlines qualitative data analysis and report generation.
Offer a free basic version with limited transcriptions and storage, then charge for premium features like advanced LLM processing, higher accuracy, and integration with cloud services. Revenue comes from monthly subscriptions and enterprise plans.
License the skill to businesses for internal use, providing custom integrations with existing tools like CRM or project management software. Revenue is generated through one-time setup fees and annual maintenance contracts.
Expose transcription and processing capabilities as an API for developers to build into their applications. Monetize through API usage tiers, with revenue based on the number of requests or processing minutes consumed.
💬 Integration Tip
Ensure all required binaries like ffmpeg and whisper-cpp are installed for optimal performance, and configure the skill to auto-detect GPU acceleration on Apple Silicon for faster processing.
Scored Jun 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Search and manage Spotify playlists, tracks, albums, artists, and playback state via the Spotify Web API. Use this skill when users want to search for music,...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Transcribe long audio files safely on 16GB RAM machines using auto-chunking with Whisper’s base model and seamless transcript merging.
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...