Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Scrape web pages via Scrapling — stealth fetching, anti-bot bypass, CSS selectors, no API key. Use when the user says 'scrape', 'pull data from this URL', 'extract from this site'. Not for meaning-based search of the user's own vault; use `enable-semantic-search`.
.claude/skills/davekilleen-scrape/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-02 | ✗→✓ | ▲ Improved | 329% | 0% |
| case-04 | ✗→✓ | ▲ Improved | 39% | 0% |
| case-05 | ✗→✓ | ▲ Improved | 19% | 0% |
| case-06 | ✗→✓ | ▲ Improved | 231% | 0% |
| case-07 | ✗→✓ | ▲ Improved | 73% | 0% |
Extract data from any website using Scrapling's MCP tools. Bypasses Cloudflare, handles dynamic JS-rendered pages, supports CSS selectors to pre-filter content (saves tokens).
/scrape <url>
/scrape <url> with selector .article-content
/scrape stealth <url>
/scrape bulk <url1> <url2> <url3>| Scenario | Tool | Why | |----------|------|-----| | Simple page, no anti-bot | scrapling_get | Fastest. HTTP with browser TLS fingerprint | | JS-rendered / SPA content | scrapling_fetch | Uses real Chromium browser | | Cloudflare / anti-bot protected | scrapling_stealthy_fetch | Stealth mode, solves captchas | | Multiple pages, same pattern | scrapling_bulk_get / scrapling_bulk_fetch | Parallel processing |
Determine from the user's request:
.main-content, #article, table.data)Default path (try in order, escalate on failure):
scrapling_get — fast HTTP, handles most sitesscrapling_fetch (real browser)scrapling_stealthy_fetch (stealth mode)User explicitly asks for stealth: Go straight to scrapling_stealthy_fetch
Multiple URLs: Use the bulk_ variants for parallel processing
Use the Scrapling MCP server tools. All tools accept:
url (required): The URL to scrapecss_selector (optional): CSS selector to extract specific elements — always use this when possible to reduce token consumptionSingle page:
Call scrapling MCP tool: get
Arguments: { "url": "<url>", "css_selector": "<selector if provided>" }Stealth:
Call scrapling MCP tool: stealthy_fetch
Arguments: { "url": "<url>", "css_selector": "<selector if provided>" }Bulk:
Call scrapling MCP tool: bulk_get
Arguments: { "urls": ["<url1>", "<url2>"], "css_selector": "<selector>" }The MCP returns extracted content (HTML or text depending on selector).
If user wants raw data: Present it formatted If user wants summary: Summarize the extracted content If user wants vault storage: Save to 00-Inbox/Scrape - [Title].md:
markdown# [Page Title] **Source:** [URL] **Scraped:** YYYY-MM-DD **Selector:** [CSS selector used, if any] ## Content [Extracted content]
| Error | Action | |-------|--------| | Empty content | Escalate to next fetcher tier | | Connection refused | Check URL is valid | | Cloudflare challenge | Auto-escalate to stealthy_fetch | | Timeout | Retry with longer timeout, suggest fetch for slow JS sites |
article, .post-content, .entry-content selectors.product, .price, .description selectors table selector, offer to convert to markdown tableul, ol selectorsUser: /scrape https://example.com/blog/ai-trends
→ Use scrapling_get, auto-detect article content
User: /scrape stealth https://protected-site.com/data
→ Use scrapling_stealthy_fetch with Cloudflare bypass
User: scrape this page and grab just the pricing table: https://saas.com/pricing
→ Use scrapling_get with css_selector="table" or ".pricing"
User: scrape these 5 competitor pages and compare their features
→ Use scrapling_bulk_get, extract feature lists, present comparisonScrapling must be installed with MCP support:
bashpip install "scrapling[ai]" scrapling install
The scrapling MCP server must be configured in .mcp.json.
Other measured skills in the registry, with their headline benchmark lift.