midscene-computer-browserVision-driven browser automation using Midscene. Operates from screenshots — no DOM or accessibility labels needed. Runs in headless Puppeteer — does NOT tak...
Install via ClawdBot CLI:
clawdbot install quanru/midscene-computer-browserGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://midscenejs.comUses known external API (expected, informational)
googleapis.comAudited Apr 16, 2026 · audit v1.0
Generated Mar 21, 2026
Automate daily checks on competitor websites to scrape product prices and availability. The skill can navigate to product pages, extract pricing data, and take screenshots for verification, enabling dynamic pricing strategies.
Automate filling out contact forms on multiple websites to collect leads or submit inquiries. The skill interacts with form fields, clicks submit buttons, and validates submissions, streamlining marketing campaigns.
Verify frontend functionality by automating user interactions like clicking buttons, navigating menus, and checking for visual elements. It captures screenshots to document UI behavior and identify bugs without manual testing.
Scrape articles, news, or data from various web sources by navigating pages, extracting text, and saving screenshots. This supports market research or academic studies by automating data collection from diverse sites.
Execute complex sequences such as logging into a web portal, downloading reports, and processing data. The skill handles each step via screenshots, enabling automation of repetitive business processes.
Offer browser automation services to clients for tasks like data scraping or form filling, charging per task or subscription. This leverages the skill's ability to handle diverse websites without coding.
Develop a software tool that integrates this skill to provide automated UI testing for websites. Sell licenses to development teams to reduce manual testing efforts and improve efficiency.
Use the skill to scrape and aggregate public web data (e.g., prices, reviews), then sell curated datasets to businesses for analytics. This model capitalizes on the skill's visual data extraction capabilities.
💬 Integration Tip
Ensure environment variables for AI models are properly configured before use, and run commands synchronously to maintain the screenshot-analyze-act loop.
Scored Jun 19, 2026
A fast Rust-based headless browser automation CLI with Node.js fallback that enables AI agents to navigate, click, type, and snapshot pages via structured commands.
Uses a headless browser to navigate web pages, interact with elements, and extract clean, readable text content from URLs.
Headless browser automation CLI optimized for AI agents with accessibility tree snapshots and ref-based element selection
通过已登录的Edge或Chrome浏览器,利用Chrome DevTools Protocol执行JS渲染页面的导航、点击、截图及数据提取等自动化操作。
High-performance browser automation for heavy scraping, multi-tab management, and precise DOM extraction. Use this when you need speed, reliability, or advanced state management (cookies/local storage) beyond standard web fetching.
Ultimate stealth browser automation with anti-detection, Cloudflare bypass, CAPTCHA solving, persistent sessions, and silent operation. Use for any web automation requiring bot detection evasion, login persistence, headless browsing, or bypassing security measures. Triggers on "bypass cloudflare", "solve captcha", "stealth browse", "silent automation", "persistent login", "anti-detection", or any task needing undetectable browser automation. When user asks to "login to X website", automatically use headed mode for login, then save session for future headless reuse.