firecrawl-scrape-cn从任意 URL 提取干净的 Markdown 内容,包括 JS 渲染的 SPA。当用户提供 URL 并想要其内容、说"抓取"、"抓网页"、"获取页面"、"从 URL 提取"或"读取网页"时使用此 Skill。支持 JS 渲染页面、多个并发 URL,返回 LLM 优化的 Markdown。
Install via ClawdBot CLI:
clawdbot install yang1002378395-cmyk/firecrawl-scrape-cnGrade Fair — based on market validation, documentation quality, package completeness, maintenance status, and authenticity signals.
Calls external URL not in known-safe list
https://example.comAudited Apr 18, 2026 · audit v1.0
Generated Apr 4, 2026
Marketing teams can scrape competitor websites to extract pricing pages, product descriptions, and blog content in clean Markdown format. This enables easy comparison of features, pricing structures, and content strategies without manual copy-pasting.
Researchers can extract content from multiple academic papers, news articles, or documentation sites published online. The ability to handle JS-rendered SPAs ensures comprehensive data gathering from modern publishing platforms.
Organizations migrating websites or creating backups can scrape existing content into Markdown format. The concurrent URL processing allows efficient batch operations while maintaining clean, structured output for storage or transfer.
Law firms can extract terms of service, privacy policies, and regulatory documents from client or competitor websites. The --only-main-content option helps isolate legal text from navigation elements for focused analysis.
SEO specialists can scrape client websites to analyze content structure, identify duplicate material, and extract text for optimization. The Markdown output facilitates easy processing by content analysis tools.
Offer tiered monthly subscriptions based on scrape volume, concurrent URL limits, and advanced features like --query functionality. Enterprise plans could include priority processing and custom integration support.
Sell credit packages where users pay per scrape operation, with premium features like JS rendering and --query commands costing additional credits. This accommodates both casual users and high-volume clients.
License the scraping technology to large corporations for internal use, with custom branding and integration into existing data pipelines. Include dedicated support and compliance features for regulated industries.
💬 Integration Tip
Start with single URL scraping using basic Markdown output, then gradually incorporate concurrent processing and format options as workflows mature.
Scored Apr 19, 2026
A fast Rust-based headless browser automation CLI with Node.js fallback that enables AI agents to navigate, click, type, and snapshot pages via structured commands.
Uses a headless browser to navigate web pages, interact with elements, and extract clean, readable text content from URLs.
Headless browser automation CLI optimized for AI agents with accessibility tree snapshots and ref-based element selection
通过已登录的Edge或Chrome浏览器,利用Chrome DevTools Protocol执行JS渲染页面的导航、点击、截图及数据提取等自动化操作。
High-performance browser automation for heavy scraping, multi-tab management, and precise DOM extraction. Use this when you need speed, reliability, or advanced state management (cookies/local storage) beyond standard web fetching.
Ultimate stealth browser automation with anti-detection, Cloudflare bypass, CAPTCHA solving, persistent sessions, and silent operation. Use for any web automation requiring bot detection evasion, login persistence, headless browsing, or bypassing security measures. Triggers on "bypass cloudflare", "solve captcha", "stealth browse", "silent automation", "persistent login", "anti-detection", or any task needing undetectable browser automation. When user asks to "login to X website", automatically use headed mode for login, then save session for future headless reuse.