alvis-web-scrapeLegal web scraping with robots.txt compliance, rate limiting, and GDPR/CCPA-aware data handling. Supports both direct HTTP scraping and managed scraping via...
Install via ClawdBot CLI:
clawdbot install alvisdunlop/alvis-web-scrapeGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://SkillBoss.co/skill.mdAudited Apr 17, 2026 · audit v1.0
Generated Sep 3, 2026
Track competitor pricing for products on e-commerce websites. The skill handles robots.txt compliance, rate limiting, and data privacy, ensuring the scraping process is legal and ethical. Retailers can adjust their prices dynamically to stay competitive.
Scrape public product reviews, forum discussions, and social media mentions (where allowed by ToS) to gauge consumer sentiment. The skill's emphasis on data minimization and PII stripping helps maintain GDPR compliance while gathering valuable insights for product development and marketing strategies.
Aggregate property listings from various real estate websites to build a comprehensive database for market analysis or a property search platform. The skill's legal checks ensure proper authorization for public data, while rate limiting prevents server overload.
Collect article metadata (headlines, publication dates, authors) from news sites to analyze trends, coverage, or for media monitoring services. By respecting robots.txt and terms of service, the tool supports legal news aggregation and journalistic research.
Provide scraped data (e.g., competitor prices, market trends) to clients on a subscription basis. The skill ensures data is legally sourced, making it attractive to businesses concerned about compliance.
Offer custom web scraping and data extraction services to clients who need specific data but lack in-house expertise. The agency charges for project setup and delivery, leveraging the skill's compliance framework as a value-add.
Build a platform that delivers insights (e.g., brand monitoring, pricing intelligence) derived from scraped data. Monetize through premium analytics dashboards and reports targeting enterprises that need business intelligence.
💬 Integration Tip
Start by setting the SkillBoss_API_KEY environment variable and test with a small, low-risk target site to ensure you understand the compliance checks and rate limiting before scaling up.
Scored Sep 3, 2026
A fast Rust-based headless browser automation CLI with Node.js fallback that enables AI agents to navigate, click, type, and snapshot pages via structured commands.
Uses a headless browser to navigate web pages, interact with elements, and extract clean, readable text content from URLs.
Headless browser automation CLI optimized for AI agents with accessibility tree snapshots and ref-based element selection
High-performance browser automation for heavy scraping, multi-tab management, and precise DOM extraction. Use this when you need speed, reliability, or advanced state management (cookies/local storage) beyond standard web fetching.
Ultimate stealth browser automation with anti-detection, Cloudflare bypass, CAPTCHA solving, persistent sessions, and silent operation. Use for any web automation requiring bot detection evasion, login persistence, headless browsing, or bypassing security measures. Triggers on "bypass cloudflare", "solve captcha", "stealth browse", "silent automation", "persistent login", "anti-detection", or any task needing undetectable browser automation. When user asks to "login to X website", automatically use headed mode for login, then save session for future headless reuse.
Browser automation CLI (browser-act) for AI agents. MUST trigger when: (1) user mentions 'browser-act' in any form, or user needs to: (2) open/visit/browse/c...