gladia-youtube-transcribeTranscribe speech from YouTube videos or audio URLs into text using Gladia API with up to 10 free hours of monthly transcription. Use when: you need to summa...
Install via ClawdBot CLI:
clawdbot install kanfred/gladia-youtube-transcribeGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Sends data to undocumented external endpoint (potential exfiltration)
POST → https://api.gladia.io/v2/pre-recordedCalls external URL not in known-safe list
https://gladia.ioAI Analysis
The skill's external API calls (gladia.io) are directly documented and required for its stated transcription functionality; there is no evidence of credential harvesting, hidden instructions, or obfuscation. The 'unknown data sink' signal is a false positive, as the endpoint is the legitimate Gladia transcription API.
Audited Apr 16, 2026 · audit v1.0
Generated Mar 20, 2026
Transcribe educational YouTube videos, especially those in Cantonese or Chinese without captions, to generate text summaries for study materials or lecture notes. This enables platforms to offer accessible content summaries for students and researchers, enhancing learning efficiency.
Convert podcast episodes from audio files or public video URLs into text transcripts for SEO optimization, content repurposing, and accessibility compliance. Media agencies can use this to create blog posts, social media snippets, or closed captions from spoken content.
Transcribe publicly available video interviews, webinars, or product reviews to extract insights for competitive analysis and trend monitoring. Businesses can analyze customer feedback or industry discussions without manual note-taking, speeding up research processes.
Provide text transcripts for video content to support accessibility initiatives, such as aiding hearing-impaired users or non-native speakers. Nonprofits can use this to make their outreach videos more inclusive and compliant with accessibility standards.
Offer 10 free hours of transcription monthly to attract users, then charge $0.61/hour for async or $0.75/hour for real-time beyond the quota. This model encourages adoption while generating revenue from heavy users or businesses with higher transcription needs.
Partner with educational or media platforms to integrate transcription as a value-added service, charging a subscription fee per user or volume-based pricing. This provides steady revenue by embedding the skill into larger ecosystems for automated content processing.
Offer tailored solutions for specific industries, such as adding speaker diarization or multi-language translation, with one-time setup fees and ongoing support charges. This targets businesses needing advanced features beyond basic transcription, leveraging the skill's API capabilities.
💬 Integration Tip
Set the GLADIA_API_KEY as a session environment variable to avoid security risks from storing in config files, and test with a public YouTube URL first to ensure proper setup.
Scored Apr 21, 2026
Local speech-to-text with the Whisper CLI (no API key).
Transcribe audio files with speaker diarization (who speaks when). Supports 100+ languages, automatic language detection, and timestamps. Use for meetings, interviews, podcasts, or voice messages. Requires AssemblyAI API key.
Search and manage Spotify playlists, tracks, albums, artists, and playback state via the Spotify Web API. Use this skill when users want to search for music,...
使用 poocr 库识别发票并导出 Excel。当用户需要识别增值税发票、批量处理发票文件或提取发票信息到 Excel 时调用此技能。
Transcribe long audio files safely on 16GB RAM machines using auto-chunking with Whisper’s base model and seamless transcript merging.
Use when the user has audio or video and wants a timestamped transcript (SRT) in the source language. Routes by source language — Chinese defaults to Volcano...