Agent skill

Scrapling Skill

by daymade in daymade/claude-code-skills

Install, troubleshoot, and use Scrapling CLI to extract HTML, Markdown, or text from webpages.

MITAuto-check: warningsProductivity & Automation

Install Scrapling Skill

The automated check flagged lines worth reading first. See the safety section below.

skills CLI
$ npx skills add daymade/claude-code-skills --skill scrapling-skill -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install daymade/claude-code-skills scrapling-skill --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/daymade/claude-code-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/scrapling-skill .claude/skills/scrapling-skill && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
scrapling-skill
GitHub stars
1.4k
Token cost
~1.4k tokens
SKILL.md length
584 words
Files
3 (incl. scripts, references)
Skills in repo
102
Repo updated
First seen
Licence
MIT

At a glance

Install, troubleshoot, and use Scrapling CLI to extract HTML, Markdown, or text from webpages.

  • Works in 6 steps: Diagnose the Install → Fix the Install → Choose the Fetcher → …
  • The user mentions Scrapling
  • SKILL.md covers Overview, Default Workflow, Step 1: Diagnose the Install and Step 2: Fix the Install, plus 7 more sections
  • Runs Python scripts from its folder; calls python3, uv and rg; reaches mp.weixin.qq.com

What it does

Scrapling Skill is an agent skill from daymade/claude-code-skills. Install, troubleshoot, and use Scrapling CLI to extract HTML, Markdown, or text from webpages. Use this skill whenever the user mentions Scrapling, uv tool install scrapling, scrapling extract, WeChat/mp.weixin articles, browser-backed page fetching, or needs help deciding between static and dynamic extraction.

Its SKILL.md is about 1.4k tokens, which your agent loads only when the skill is triggered. The skill folder holds 4 other files, including scripts and reference files (for example `references/troubleshooting.md` and `scripts/diagnose_scrapling.py`).

It sits in Productivity & Automation, covering Messaging and chat bots. It works with WeChat. The repository describes itself as: Professional Claude Code skills marketplace featuring production-ready skills for enhanced development workflows. The licence is MIT.

When your agent uses it

  • The user mentions Scrapling
  • Uv tool install scrapling
  • Scrapling extract
  • WeChat/mp.weixin articles

Example prompts

  • “/scrapling-skill”

Requirements

  • Python 3

Workflow steps

6 steps, taken from the step headings in SKILL.md.

  1. Diagnose the Install
  2. Fix the Install
  3. Choose the Fetcher
  4. Run the Smallest Useful Command
  5. Validate the Output
  6. Handle Known Failure Modes

What it can do on your machine

Read from SKILL.md and the folder at commit 3c268d6. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python3
    • uv
    • rg

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • mp.weixin.qq.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Scrapling Skill loads about 1.4k tokens when it runs, and up to ~2.3k if it reads all its reference files. Until then it costs about 83 tokens; SKILL.md has 584 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~83
When it runs · the whole SKILL.md, loaded when a task matches
~1.4k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~2.3k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: warnings

The automated check found patterns that need a careful read before installing.

  • WarningContains instruction-override wording (e.g. “without asking the user”)SKILL.md:174
    - Do not tell the user to reinstall blindly. Verify first.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from daymade/claude-code-skills at commit 3c268d6, republished under its MIT licence (© daymade). 584 words, ~1,416 tokens.

Download SKILL.mdSave it as .claude/skills/scrapling-skill/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.
name
scrapling-skill
description
Install, troubleshoot, and use Scrapling CLI to extract HTML, Markdown, or text from webpages. Use this skill whenever the user mentions Scrapling, `uv tool install scrapling`, `scrapling extract`, WeChat/mp.weixin articles, browser-backed page fetching, or needs help deciding between static and dynamic extraction.

Scrapling Skill

Overview

Use Scrapling through its CLI as the default path. Start with the smallest working command, validate the saved output, and only escalate to browser-backed fetching when the static fetch does not contain the real page content.

Do not assume the user's Scrapling install is healthy. Verify it first.

Default Workflow

Copy this checklist and keep it updated while working:

text
Scrapling Progress:
- [ ] Step 1: Diagnose the local Scrapling install
- [ ] Step 2: Fix CLI extras or browser runtime if needed
- [ ] Step 3: Choose static or dynamic fetch
- [ ] Step 4: Save output to a file
- [ ] Step 5: Validate file size and extracted content
- [ ] Step 6: Escalate only if the previous path failed

Step 1: Diagnose the Install

Run the bundled diagnostic script first:

bash
python3 scripts/diagnose_scrapling.py

Use the result as the source of truth for the next step.

Step 2: Fix the Install

If the CLI was installed without extras

If scrapling --help fails with missing click or a message about installing Scrapling with extras, reinstall it with the CLI extra:

bash
uv tool uninstall scrapling
uv tool install 'scrapling[shell]'

Do not default to scrapling[all] unless the user explicitly needs the broader feature set.

If browser-backed fetchers are needed

Install the Playwright runtime:

bash
scrapling install

If the install looks slow or opaque, read references/troubleshooting.md before guessing. Do not claim success until either:

  • scrapling install reports that dependencies are already installed, or
  • the diagnostic script confirms both Chromium and Chrome Headless Shell are present.

Step 3: Choose the Fetcher

Use this decision rule:

  • Start with extract get for normal pages, article pages, and most WeChat public articles.
  • Use extract fetch when the static HTML does not contain the real content or the page depends on JavaScript rendering.
  • Use extract stealthy-fetch only after fetch still fails because of anti-bot or challenge behavior. Do not make it the default.

Step 4: Run the Smallest Useful Command

Always quote URLs in shell commands. This is mandatory in zsh when the URL contains ?, &, or other special characters.

Full page to HTML
bash
scrapling extract get 'https://example.com' page.html
Main content to Markdown
bash
scrapling extract get 'https://example.com' article.md -s 'main'
JS-rendered page with browser automation
bash
scrapling extract fetch 'https://example.com' page.html --timeout 20000
WeChat public article body

Use #js_content first. This is the default selector for article body extraction on mp.weixin.qq.com pages.

bash
scrapling extract get 'https://mp.weixin.qq.com/s/ARTICLE_ID?scene=1' article.md -s '#js_content'

Step 5: Validate the Output

After every extraction, verify the file instead of assuming success:

bash
wc -c article.md
sed -n '1,40p' article.md

For HTML output, check that the expected title, container, or selector target is actually present:

bash
rg -n '<title>|js_content|rich_media_title|main' page.html

If the file is tiny, empty, or missing the expected container, the extraction did not succeed. Go back to Step 3 and switch fetchers or selectors.

Show full SKILL.md (228 more words)Show less

Step 6: Handle Known Failure Modes

Local TLS trust store problem

If extract get fails with curl: (60) SSL certificate problem, treat it as a local trust-store problem first, not a Scrapling content failure.

Retry the same command with:

bash
--no-verify

Only do this after confirming the failure matches the local certificate verification error pattern. Do not silently disable verification by default.

WeChat article pages

For mp.weixin.qq.com:

  • Try extract get before extract fetch
  • Use -s '#js_content' for the article body
  • Validate the saved Markdown or HTML immediately
Browser-backed fetch failures

If extract fetch fails:

  1. Re-check the install with python3 scripts/diagnose_scrapling.py
  2. Confirm Chromium and Chrome Headless Shell are present
  3. Retry with a slightly longer timeout
  4. Escalate to stealthy-fetch only if the site behavior justifies it

Command Patterns

Diagnose and smoke test a URL
bash
python3 scripts/diagnose_scrapling.py --url 'https://example.com'
Diagnose and smoke test a WeChat article body
bash
python3 scripts/diagnose_scrapling.py \
  --url 'https://mp.weixin.qq.com/s/ARTICLE_ID?scene=1' \
  --selector '#js_content' \
  --no-verify
Diagnose and smoke test a browser-backed fetch
bash
python3 scripts/diagnose_scrapling.py \
  --url 'https://example.com' \
  --dynamic

Guardrails

  • Do not tell the user to reinstall blindly. Verify first.
  • Do not default to the Python library API when the user is clearly asking about the CLI.
  • Do not jump to browser-backed fetching unless the static result is missing the real content.
  • Do not claim success from exit code alone. Inspect the saved file.
  • Do not hardcode user-specific absolute paths into outputs or docs.

Resources

  • Installation and smoke test helper: scripts/diagnose_scrapling.py
  • Verified failure modes and recovery paths: references/troubleshooting.md

© daymade, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 2 other files (scripts, references) in scrapling-skill of daymade/claude-code-skills.

  • SKILL.md
  • references/troubleshooting.md
  • scripts/diagnose_scrapling.py

Open the folder on GitHubat commit 3c268d6

Compare with similar skills

Scrapling Skill next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Scrapling Skill compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Scrapling Skill this skilldaymade/claude-code-skills1.4k—~1.4kAutomated safety check: WarnMIT
Huashu Wechat Imagealchaincyf/huashu-skills1.7k1 repos~2.4kAutomated safety check: NotesMIT
Wechat Article Browserusedracohu2025-cloud/draco-skills-collection227—~513Automated safety check: PassMIT
She Love Me863401402/she-love-me9141 repos~1.3kAutomated safety check: PassMIT
Claude To Imop7418/Claude-to-IM-skill2.9k—~3.4kAutomated safety check: NotesMIT
Yichen Mac Wechat Dual Openmcncarl/yichen-skills4.3k—~1.5kAutomated safety check: PassCustom licence

Similar skills

  • Huashu Wechat Image

    alchaincyf/huashu-skills

    为微信公众号文章生成高质量配图。支持封面图(2.35:1)、正文插图(16:9/4:3)、信息图。提供两条路径:AI生成(视觉创意型)和HTML渲染(文字精确型)。当用户提到"公众号配图"、"公众号封面"、"文章配图"、"正文插图"、"公众号图片"时使用此技能。

    1.7k GitHub starsUsed in 1 repo~2.4k tokens
    Productivity & AutomationAuto-check: notes
  • Wechat Article Browseruse

    dracohu2025-cloud/draco-skills-collection

    使用 BrowserUse 云浏览器 + Playwright CDP 抓取微信公众号文章,并可直接发布到飞书原生文档。

    227 GitHub stars~513 tokensUpdated 20 days ago
    Productivity & AutomationAuto-check passed
  • She Love Me

    863401402/she-love-me

    Acquire, import, and analyze WeChat or QQ chat histories, including installing supported exporters, guiding required login or contact selection, converting exports, assessing relationship dynamics…

    914 GitHub starsUsed in 1 repo~1.3k tokens
    Productivity & AutomationAuto-check passed
  • Claude To Im

    op7418/Claude-to-IM-skill

    Bridge THIS Claude Code or Codex session to Telegram, Discord, Feishu/Lark, QQ, or WeChat so the user can chat with Claude from their phone.

    2.9k GitHub stars~3.4k tokensUpdated 6 mo ago
    Productivity & AutomationAuto-check: notes
  • Yichen Mac Wechat Dual Open

    mcncarl/yichen-skills

    Create, inspect, repair, and polish a second WeChat app on macOS by copying WeChat, changing the bundle identifier, ad-hoc re-signing, preparing it for the user to open manually, setting Chinese…

    4.3k GitHub stars~1.5k tokensUpdated 3 days ago
    Productivity & AutomationAuto-check passed
  • Create Ex

    perkfly/ex-skill

    Distill an ex-girlfriend into an AI Skill. An agent skill from perkfly/ex-skill.

    2.5k GitHub stars~3.7k tokensUpdated 6 mo ago
    Productivity & AutomationAuto-check: notes

More from daymade/claude-code-skills

All 102 skills in this repo
  • Video Comparer

    daymade/claude-code-skills

    This skill should be used when comparing two videos to analyze compression results or quality differences.

    1.4k GitHub starsUsed in 1 repo~1.4k tokens
    Auto-check: notes
  • CLI Demo Generator

    daymade/claude-code-skills

    Generates professional animated CLI demos as GIFs using VHS terminal recordings.

    1.4k GitHub stars~1.7k tokensUpdated today
    Auto-check passed
  • Doc To Markdown

    daymade/claude-code-skills

    Converts DOCX/PDF/PPTX and saved HTML/HTM to high-quality Markdown with automatic post-processing.

    1.4k GitHub stars~2.5k tokensUpdated today
    Auto-check passed
  • Interaction Design Board

    daymade/claude-code-skills

    Generates several distinct, clickable HTML interaction prototypes for one product surface into a Design Board and collects selection/remix feedback before implementation.

    1.4k GitHub stars~2.7k tokensUpdated today
    Auto-check passed
  • Auto Repo Setup

    daymade/claude-code-skills

    Diagnoses and repairs repository setup and guarded Git workflows for Claude Code or Codex — environment repair, startup sync, hook auditing, collaborator handoff.

    1.4k GitHub stars~2.6k tokensUpdated today
    Auto-check: notes
  • Bigdata Skill

    daymade/claude-code-skills

    Pulls Bigdata.com (RavenPack) financial and news data via the official bigdata-client SDK and /v1/ REST endpoints — structured financials, prices, analyst estimates, entity-sentiment series…

    1.4k GitHub stars~3.7k tokensUpdated today
    Auto-check passed

Works with

Questions about Scrapling Skill

What does Scrapling Skill do?

Install, troubleshoot, and use Scrapling CLI to extract HTML, Markdown, or text from webpages. Scrapling Skill is an agent skill from daymade/claude-code-skills. Install, troubleshoot, and use Scrapling CLI to extract HTML, Markdown, or text from webpages.

When should I use Scrapling Skill?

Scrapling Skill fits situations like: the user mentions Scrapling; uv tool install scrapling; scrapling extract; weChat/mp.weixin articles.

How do I install Scrapling Skill in Claude Code?

Run `npx skills add daymade/claude-code-skills --skill scrapling-skill -a claude-code`. Or copy the skill folder (scrapling-skill in daymade/claude-code-skills) into .claude/skills/scrapling-skill in your project. Claude Code loads it when a task matches its description.

How do I install Scrapling Skill in Codex?

Run `npx skills add daymade/claude-code-skills --skill scrapling-skill -a codex`. Or copy the skill folder (scrapling-skill in daymade/claude-code-skills) into .agents/skills/scrapling-skill in your project. Codex loads it when a task matches its description.

Can I use Scrapling Skill in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add daymade/claude-code-skills --skill scrapling-skill -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/scrapling-skill, .gemini/skills/scrapling-skill, .github/skills/scrapling-skill and .opencode/skills/scrapling-skill in your project.

What does Scrapling Skill need to run?

Going by SKILL.md and its folder, Scrapling Skill needs Python for the scripts in its folder and the command-line tools its instructions call (python3, uv and rg). Our summary lists: Python 3.

Does Scrapling Skill access the network?

SKILL.md names 1 domain. In commands or code: mp.weixin.qq.com; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.

Is Scrapling Skill safe to install?

Our automated static check of SKILL.md flagged 1 warning(s): contains instruction-override wording (e.g. “without asking the user”). Read the flagged lines before installing; the check is not a guarantee either way. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Scrapling Skill use?

Scrapling Skill is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Scrapling Skill use?

About 1.4k tokens (SKILL.md is roughly 5.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 843 tokens, read only when the agent opens those files.

What are the alternatives to Scrapling Skill?

Skills that share tags, products or a category with Scrapling Skill: Huashu Wechat Image (alchaincyf/huashu-skills, 1.7k stars), Wechat Article Browseruse (dracohu2025-cloud/draco-skills-collection, 227 stars), She Love Me (863401402/she-love-me, 914 stars) and Claude To Im (op7418/Claude-to-IM-skill, 2.9k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Scrapling Skill?

daymade (a GitHub user) maintains it in daymade/claude-code-skills, which has 1,443 GitHub stars. The repository holds 102 skills in this directory. The repository was last updated on October 7, 2026.

Source: daymade/claude-code-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.