panscrapling-web-scraper强大的网页抓取技能。基于 Scrapling,自动绕过 Cloudflare/反爬系统。 触发词:抓取网页、爬取、scrape、fetch、抓取内容、提取网页、获取页面。 使用场景: (1) 抓取被 Cloudflare 保护的网页 (2) 提取页面内容 (3) 网页数据采集 (4) 动态渲染页面抓取 自动安装:...
Install via ClawdBot CLI:
clawdbot install dashiming/panscrapling-web-scraperGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://github.com/D4Vinci/ScraplingUses known external API (expected, informational)
raw.githubusercontent.comAudited Apr 18, 2026 · audit v1.0
Generated Aug 11, 2026
Automatically extract product prices and availability from competitor websites, even if protected by Cloudflare. This enables real-time price tracking and market analysis to adjust pricing strategies dynamically.
Scrape articles and blog posts from news sites that employ anti-bot measures. This allows media monitoring services to aggregate content for analysis, sentiment tracking, or trend reporting without manual copy-pasting.
Extract contact details, company information, and job postings from business directories and corporate websites. This provides sales teams with structured data to build targeted prospect lists and personalize outreach.
Gather property listings, prices, and descriptions from real estate portals that use advanced bot protection. This supports market analysis, property valuation, and investment decisions by providing comprehensive data sets.
Collect publicly available data from government sites, research databases, and social media for academic studies. The scraper handles sites with JavaScript rendering and anti-bot measures, ensuring reliable data for research projects.
Offer a subscription-based service that scrapes specified websites on a schedule and delivers structured data via API. This is ideal for clients needing continuous data streams for analytics or monitoring.
Provide bespoke scraping solutions for businesses that need tailored data extraction. This includes setting up scraping scripts, handling edge cases, and delivering cleaned data in the desired format.
Compile and sell pre-collected, processed datasets (e.g., product catalogs, real estate listings) to companies that require ready-to-use data for market research, competitive intelligence, or training AI models.
💬 Integration Tip
Integrate with existing Python workflows by using the CLI within scripts or cron jobs, and connect output to databases or BI tools.
Scored Aug 11, 2026
A fast Rust-based headless browser automation CLI with Node.js fallback that enables AI agents to navigate, click, type, and snapshot pages via structured commands.
Uses a headless browser to navigate web pages, interact with elements, and extract clean, readable text content from URLs.
Headless browser automation CLI optimized for AI agents with accessibility tree snapshots and ref-based element selection
High-performance browser automation for heavy scraping, multi-tab management, and precise DOM extraction. Use this when you need speed, reliability, or advanced state management (cookies/local storage) beyond standard web fetching.
Ultimate stealth browser automation with anti-detection, Cloudflare bypass, CAPTCHA solving, persistent sessions, and silent operation. Use for any web automation requiring bot detection evasion, login persistence, headless browsing, or bypassing security measures. Triggers on "bypass cloudflare", "solve captcha", "stealth browse", "silent automation", "persistent login", "anti-detection", or any task needing undetectable browser automation. When user asks to "login to X website", automatically use headed mode for login, then save session for future headless reuse.
Browser automation CLI (browser-act) for AI agents. MUST trigger when: (1) user mentions 'browser-act' in any form, or user needs to: (2) open/visit/browse/c...