CLI-first web scraping & content extraction with optional MCP server.

Apache-2.0Auto-check passedData & Analytics

Install Scrapling

skills CLI
$ npx skills add foryourhealth111-pixel/Vibe-Skills --skill scrapling -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install foryourhealth111-pixel/Vibe-Skills scrapling --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/foryourhealth111-pixel/Vibe-Skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/bundled/skills/scrapling .claude/skills/scrapling && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
scrapling
GitHub stars
3.6k
Token cost
~1.1k tokens
SKILL.md length
423 words
Files
2 (incl. scripts)
Skills in repo
81
Repo updated
First seen
Licence
Apache-2.0

At a glance

CLI-first web scraping & content extraction with optional MCP server.

  • Works in 4 steps: Extract full page body (to Markdown) → Extract a specific element (CSS… → Extract HTML for downstream parsing → …
  • You have target URLs and need clean
  • SKILL.md covers When to use, Boundaries (vs Playwright /…, Prerequisite check (required) and Installation (recommended), plus 4 more sections
  • Runs PowerShell scripts from its folder; calls python

What it does

Scrapling is an agent skill from foryourhealth111-pixel/Vibe-Skills. CLI-first web scraping & content extraction with optional MCP server. Use when you have target URLs and need clean, selector-based outputs (html/md/txt).

Its SKILL.md is about 1.1k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts.

It sits in Data & Analytics, covering Web scraping and MCP servers. It works with Model Context Protocol, Playwright and Python. The repository describes itself as: Intelligent Skill routing and workflow orchestration for AI agents — +21.12 pp reward, −29.6% tokens on SkillsBench with DeepSeekV4Flash-VE. The licence is Apache-2.0.

When your agent uses it

  • You have target URLs and need clean
  • Selector-based outputs (html/md/txt)

Example prompts

  • “/scrapling”

Requirements

  • Python 3
  • PowerShell

Workflow steps

4 steps, taken from the step headings in SKILL.md.

  1. Extract full page body (to Markdown)
  2. Extract a specific element (CSS selector) to text
  3. Extract HTML for downstream parsing
  4. Use browser-backed fetcher mode (when simple GET is blocked / dynamic)

What it can do on your machine

Read from SKILL.md and the folder at commit ddcaa2a. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (PowerShell), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Scrapling loads about 1.1k tokens when it runs. Until then it costs about 41 tokens; SKILL.md has 423 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~41
When it runs · the whole SKILL.md, loaded when a task matches
~1.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from foryourhealth111-pixel/Vibe-Skills at commit ddcaa2a, republished under its Apache-2.0 licence (© foryourhealth111-pixel). 423 words, ~1,082 tokens.

Download SKILL.mdSave it as .claude/skills/scrapling/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
scrapling
description
CLI-first web scraping & content extraction with optional MCP server. Use when you have target URLs and need clean, selector-based outputs (html/md/txt).

Scrapling Skill (VCO)

Scrapling is a Python-based web scraping / extraction toolkit that exposes:

  • a CLI (scrapling ...) for fetching + extracting content into files
  • an optional MCP server (scrapling mcp) so an agent can call structured scraping tools

This skill is CLI-first. Prefer it when you already have URLs and need reliable, repeatable extraction (CSS selector → file).

When to use

Use scrapling when you need:

  • Extract specific parts of a web page (CSS selector / XPath) into .txt / .md / .html
  • Run repeatable scraping jobs (batch URLs with a small wrapper script)
  • Reduce token usage by extracting only the relevant DOM region before passing to the LLM
  • Provide a local MCP endpoint for scraping tools (agent → MCP → scrapling)
vs playwright
  • scrapling: best for “get URL → extract selector → write file” workflows; simpler, faster iteration
  • playwright: best for interactive UI flows (login, multi-step navigation, downloads, complex JS actions, stateful sessions)

If you must navigate or click through a UI, use playwright. If you can directly fetch the target page and just need extraction, use scrapling.

vs search tools
  • Search tools are for discovering sources/URLs (query → result list → choose URLs).
  • scrapling is for acquisition + extraction once you already know the URL(s).

A common pipeline:

  1. Search → find candidate URLs
  2. Scrapling → extract focused content from chosen URLs
  3. LLM → summarize / transform / analyze extracted outputs

Prerequisite check (required)

  1. Python version (Scrapling requires Python >= 3.10):
powershell
python --version
  1. Scrapling CLI availability:
powershell
scrapling --help

Scrapling’s CLI and MCP features are enabled via extras.

Recommended (CLI + MCP + fetchers):

powershell
python -m pip install "scrapling[ai]"

If you only want CLI fetch/extract without MCP:

powershell
python -m pip install "scrapling[fetchers]"

If you use browser-based fetchers, you may need browser binaries:

powershell
# Option A: via Scrapling helper (after install)
scrapling install

# Option B: directly via Playwright
python -m playwright install
Show full SKILL.md (155 more words)Show less

Wrapper script (Windows convenience)

This skill ships a thin PowerShell wrapper:

  • C:/Users/羽裳/.codex/skills/scrapling/scripts/scrapling.ps1

It checks whether scrapling exists and prints install hints if missing.

Common CLI patterns

1) Extract full page body (to Markdown)
powershell
scrapling extract get "https://example.com" out.md
2) Extract a specific element (CSS selector) to text
powershell
scrapling extract get "https://example.com" out.txt --css-selector "main article"
3) Extract HTML for downstream parsing
powershell
scrapling extract get "https://example.com" out.html --css-selector "#content"
4) Use browser-backed fetcher mode (when simple GET is blocked / dynamic)
powershell
scrapling extract fetch "https://example.com" out.md --css-selector "main"

Tip: keep outputs in files and only feed the smallest relevant snippet to the LLM.

MCP server relationship (optional)

Scrapling can run as an MCP server. This is useful when:

  • the agent needs tool-style scraping calls
  • you want scraping results to be structured and deterministic

Start MCP server (stdio transport by default):

powershell
scrapling mcp

Optional: run MCP server with HTTP transport:

powershell
scrapling mcp --http --host 127.0.0.1 --port 8765
Example MCP server config snippet
json
{
  "servers": {
    "scrapling": {
      "mode": "stdio",
      "command": "scrapling",
      "args": ["mcp"],
      "required": false,
      "note": "Requires: python -m pip install \"scrapling[ai]\""
    }
  }
}

Safety & ops notes

  • Prefer selector-based extraction to minimize data volume.
  • Treat scraping as an external dependency: handle timeouts, retries, and failures explicitly.
  • For aggressive bot protection, consider switching fetchers or using playwright.

© foryourhealth111-pixel, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (scripts) in bundled/skills/scrapling of foryourhealth111-pixel/Vibe-Skills.

  • SKILL.md
  • scripts/scrapling.ps1

Open the folder on GitHubat commit ddcaa2a

Compare with similar skills

Scrapling next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Scrapling compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Scrapling this skillforyourhealth111-pixel/Vibe-Skills3.6k—~1.1kAutomated safety check: PassApache-2.0
Verified Researchsweetcornna/free-search-mcp126—~1.8kAutomated safety check: PassMIT
Openbb Data Fetchermonarchjuno/vibe-investing299—~2.9kAutomated safety check: NotesMIT
Diagnosekbanc85/claudia296—~1.8kAutomated safety check: PassCustom licence
Anti Detect Browserantibrow/anti-detect-browser-skills17—~9.8kAutomated safety check: WarnMIT
Cloudflare Browser Renderingeinverne/dotfiles121—~4.9kAutomated safety check: PassGPL-3.0

Similar skills

  • Verified Research

    sweetcornna/free-search-mcp

    Use with the free-search MCP tools whenever a web lookup must yield facts someone will rely on: dates, deadlines, prices, prizes, fees, rules, eligibility, schedules, versions, statistics, news, or…

    126 GitHub stars~1.8k tokensUpdated 6 days ago
    Agent WorkflowsAuto-check passed
  • Openbb Data Fetcher

    monarchjuno/vibe-investing

    Fetch financial, market, economic, fundamental, news, options, crypto, ETF, index, and macro data through the OpenBB Python interface instead of the OpenBB MCP server.

    299 GitHub stars~2.9k tokensUpdated 5 mo ago
    Data & AnalyticsAuto-check: notes
  • Diagnose

    kbanc85/claudia

    Check memory system health and troubleshoot connectivity issues.

    296 GitHub stars~1.8k tokensUpdated 1 mo ago
    Data & AnalyticsAuto-check passed
  • Anti Detect Browser

    antibrow/anti-detect-browser-skills

    Drive Chromium from standard Playwright APIs with a real-device fingerprint applied in the kernel, one persistent isolated profile per identity, and a per-profile proxy whose exit IP sets timezone…

    17 GitHub stars~9.8k tokensUpdated 1 mo ago
    Testing & QAAuto-check: warnings
  • Guide for implementing Cloudflare Browser Rendering - a headless browser automation API for screenshots, PDFs, web scraping, and testing.

    121 GitHub stars~4.9k tokensUpdated 1 mo ago
    Productivity & AutomationAuto-check passed
  • Browser MCP Agent

    antibrow/anti-detect-browser-skills

    Give an AI agent its own real browser over MCP tool calls - launch, navigate, click, fill, screenshot, extract text, run JS - with a kernel-level real-device fingerprint and a persistent profile, so…

    17 GitHub starsUsed in 1 repo~4.2k tokens
    Productivity & AutomationAuto-check: warnings

More from foryourhealth111-pixel/Vibe-Skills

All 81 skills in this repo
  • Market Research Reports

    foryourhealth111-pixel/Vibe-Skills

    Produces long consulting-style market research and industry reports covering market sizing, competitive landscape, market entry and investment theses.

    3.6k GitHub stars~2.5k tokensUpdated 1 mo ago
    Auto-check: notes
  • Academic Venue Templates

    foryourhealth111-pixel/Vibe-Skills

    Supplies venue-specific LaTeX templates and formatting rules for journals, conferences and posters, and checks a manuscript against page limits and submission requirements.

    3.6k GitHub stars~3.9k tokensUpdated 1 mo ago
    Auto-check: notes
  • Digital Brain

    foryourhealth111-pixel/Vibe-Skills

    This skill should be used when the user asks to "write a post", "check my voice", "look up contact", "prepare for meeting", "weekly review", "track goals", or mentions personal brand, content…

    3.6k GitHub stars~1.7k tokensUpdated 1 mo ago
    Auto-check passed
  • Smart File Writer

    foryourhealth111-pixel/Vibe-Skills

    Diagnoses why a file write failed (permissions, disk space, path length, locks, read-only mounts) before retrying, instead of repeating the same call blindly.

    3.6k GitHub stars~2.6k tokensUpdated 1 mo ago
    Auto-check passed
  • Automated Video Studio

    foryourhealth111-pixel/Vibe-Skills

    Turns footage, audio and a storyboard plan into a finished short video with FFmpeg jump-cuts, subtitle burn-in and a final polish pass.

    3.6k GitHub stars~838 tokensUpdated 1 mo ago
    Auto-check passed
  • Citation Management

    foryourhealth111-pixel/Vibe-Skills

    Turns DOIs, PMIDs and arXiv IDs into clean BibTeX, searches Google Scholar and PubMed, and checks and deduplicates a reference list.

    3.6k GitHub stars~7.6k tokensUpdated 1 mo ago
    Auto-check: notes

Questions about Scrapling

What does Scrapling do?

CLI-first web scraping & content extraction with optional MCP server. Scrapling is an agent skill from foryourhealth111-pixel/Vibe-Skills. CLI-first web scraping & content extraction with optional MCP server.

When should I use Scrapling?

Scrapling fits situations like: you have target URLs and need clean; selector-based outputs (html/md/txt).

How do I install Scrapling in Claude Code?

Run `npx skills add foryourhealth111-pixel/Vibe-Skills --skill scrapling -a claude-code`. Or copy the skill folder (bundled/skills/scrapling in foryourhealth111-pixel/Vibe-Skills) into .claude/skills/scrapling in your project. Claude Code loads it when a task matches its description.

How do I install Scrapling in Codex?

Run `npx skills add foryourhealth111-pixel/Vibe-Skills --skill scrapling -a codex`. Or copy the skill folder (bundled/skills/scrapling in foryourhealth111-pixel/Vibe-Skills) into .agents/skills/scrapling in your project. Codex loads it when a task matches its description.

Can I use Scrapling in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add foryourhealth111-pixel/Vibe-Skills --skill scrapling -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/scrapling, .gemini/skills/scrapling, .github/skills/scrapling and .opencode/skills/scrapling in your project.

What does Scrapling need to run?

Going by SKILL.md and its folder, Scrapling needs PowerShell for the scripts in its folder and the command-line tools its instructions call (python). Our summary lists: Python 3; PowerShell.

Does Scrapling access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Scrapling safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Scrapling use?

Scrapling is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Scrapling use?

About 1.1k tokens (SKILL.md is roughly 4.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Scrapling?

Skills that share tags, products or a category with Scrapling: Verified Research (sweetcornna/free-search-mcp, 126 stars), Openbb Data Fetcher (monarchjuno/vibe-investing, 299 stars), Diagnose (kbanc85/claudia, 296 stars) and Anti Detect Browser (antibrow/anti-detect-browser-skills, 17 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Scrapling?

foryourhealth111-pixel (a GitHub user) maintains it in foryourhealth111-pixel/Vibe-Skills, which has 3,627 GitHub stars. The repository holds 81 skills in this directory. The repository was last updated on August 31, 2026.

Source: foryourhealth111-pixel/Vibe-Skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.