browser-audio-captureCapture audio from any browser tab — meetings, YouTube, podcasts, courses, webinars — and stream to any AI agent. Zero API keys, works with any framework.
Install via ClawdBot CLI:
clawdbot install jarvis563/browser-audio-captureGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
http://127.0.0.1:8900/audio/browser`Audited Apr 16, 2026 · audit v1.0
Generated Mar 21, 2026
Automatically capture audio from virtual meetings on platforms like Google Meet or Zoom to generate real-time summaries and action items. This reduces manual note-taking and ensures key points are documented for team follow-up.
Stream audio from online courses or webinars to AI agents for transcription and note extraction. This helps students and professionals create study guides or review materials without manual effort.
Capture earnings calls or public webinars from company websites to analyze financial discussions and strategic insights. AI agents can process the audio for trends and competitor monitoring.
Use audio from customer calls on browser-based platforms like Discord or Cal.com to transcribe and analyze interactions. This enables real-time feedback and quality assurance for support teams.
Stream audio from YouTube videos or podcasts to AI agents for automatic transcription and content summarization. This aids content creators in generating show notes or repurposing media efficiently.
Offer a cloud-based service with enhanced features like multi-user support, advanced analytics, and integration with popular AI platforms. Charge monthly or annual fees based on usage tiers and storage limits.
Provide customized solutions for large organizations with needs such as on-premise deployment, security compliance, and dedicated support. License fees can be structured per user or as a flat enterprise rate.
Release a free basic version of the skill with limited features, then monetize through paid upgrades like higher-quality audio streaming, priority support, or integration with premium transcription services.
💬 Integration Tip
Ensure Chrome is running with remote debugging enabled and verify the audio endpoint is correctly configured to avoid connectivity issues.
Scored Jun 19, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Turn your AI into JARVIS. Voice, wit, and personality — the complete package. Humor cranked to maximum.
Text-to-speech conversion using node-edge-tts npm package for generating audio from text. Supports multiple voices, languages, speed adjustment, pitch contro...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...