webscraper-pulpminerConvert any webpage into structured JSON data using AI. Scrape websites, extract data into custom JSON schemas, and call saved APIs programmatically. Useful for web scraping, data extraction, content monitoring, lead generation, price tracking, and building data pipelines.
Install via ClawdBot CLI:
clawdbot install melvin2016/webscraper-pulpminerGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Sends data to undocumented external endpoint (potential exfiltration)
POST → https://api.pulpminer.com/external/<apiIdCalls external URL not in known-safe list
https://pulpminer.comAI Analysis
The skill's external API calls are consistent with its stated purpose of web scraping and data extraction, requiring explicit user configuration of a target URL and API key. While it sends data to an external service not on a pre-approved list, this is a documented and necessary function of the skill, not a hidden exfiltration channel.
Audited Apr 17, 2026 · audit v1.0
Generated Mar 21, 2026
Automatically track product prices from competitor websites by setting up saved APIs for key product pages. Use dynamic URLs to handle different product IDs and schedule regular calls to monitor price changes, enabling dynamic pricing strategies.
Scrape property listings from real estate portals to extract details like price, location, and features into structured JSON. Use CSS selectors to focus on listing sections and extra instructions to filter for specific criteria, such as properties under a certain price.
Collect articles from news websites by configuring APIs with JSON templates to standardize output fields like headline, author, and publication date. Enable JS rendering for modern news sites and use caching to reduce credit usage while monitoring for updates.
Extract job postings from career sites to analyze trends in skills, salaries, and locations. Set up dynamic APIs for search results with variables like job title and location, and use extra instructions to format data for analytics pipelines.
Scrape financial reports or stock data from investment websites by using saved APIs with JSON templates for structured output. Configure CSS selectors to target specific tables and enable caching to efficiently track daily changes without excessive API calls.
Offer tiered subscription plans based on API call volume and features like JS rendering or priority support. Charge monthly fees with higher tiers providing more credits and advanced configurations, targeting businesses needing regular data extraction.
Sell credits that users purchase upfront to pay for individual API calls, with costs varying by endpoint and options like JS rendering. This model appeals to occasional users or those with fluctuating scraping needs, with bulk discounts to encourage larger purchases.
Provide tailored integration services, custom API configurations, and dedicated support for large organizations. Charge premium fees for setup, training, and ongoing maintenance, focusing on clients with complex data pipelines or high-volume requirements.
💬 Integration Tip
Start with static APIs to test basic scraping, then use dynamic URLs and JSON templates for more complex workflows; enable caching to save credits for monitoring tasks.
Scored Apr 19, 2026
A fast Rust-based headless browser automation CLI with Node.js fallback that enables AI agents to navigate, click, type, and snapshot pages via structured commands.
Uses a headless browser to navigate web pages, interact with elements, and extract clean, readable text content from URLs.
Headless browser automation CLI optimized for AI agents with accessibility tree snapshots and ref-based element selection
通过已登录的Edge或Chrome浏览器,利用Chrome DevTools Protocol执行JS渲染页面的导航、点击、截图及数据提取等自动化操作。
Send and receive SMS/RCS via Google Messages web interface (messages.google.com). Use when asked to "send a text", "check texts", "SMS", "text message", "Google Messages", or forward incoming texts to other channels.
High-performance browser automation for heavy scraping, multi-tab management, and precise DOM extraction. Use this when you need speed, reliability, or advanced state management (cookies/local storage) beyond standard web fetching.