Agent skill

Scrape

by davekilleen in davekilleen/Dex

Scrape web pages via Scrapling — stealth fetching, anti-bot bypass, CSS selectors, no API key.

Custom licenceAuto-check passedData & Analytics

Install Scrape

skills CLI
$ npx skills add davekilleen/Dex --skill scrape -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install davekilleen/Dex scrape --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/davekilleen/Dex.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/scrape .claude/skills/scrape && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
scrape
GitHub stars
493
Token cost
~1.2k tokens
SKILL.md length
387 words
Files
1
Skills in repo
58
Repo updated
First seen
Licence
Custom licence

At a glance

Scrape web pages via Scrapling — stealth fetching, anti-bot bypass, CSS selectors, no API key.

  • Works in 5 steps: Parse User Intent → Choose the Right Scrapling MCP Tool → Call the MCP Tool → …
  • The user says scrape
  • SKILL.md covers Usage, When to Use (Tool Selection), Implementation and Smart Defaults, plus 3 more sections
  • Calls pip; reaches protected-site.com and saas.com

What it does

Scrape is an agent skill from davekilleen/Dex. Scrape web pages via Scrapling — stealth fetching, anti-bot bypass, CSS selectors, no API key. Use when the user says 'scrape', 'pull data from this URL', 'extract from this site'. Not for meaning-based search of the user's own vault; use enable-semantic-search.

Its SKILL.md is about 1.2k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Data & Analytics, covering Web scraping. It works with Cloudflare. The repository describes itself as: Your AI Chief of Staff — a personal operating system starter kit that adapts to your role. No coding required.

When your agent uses it

  • The user says scrape
  • Pull data from this URL
  • Extract from this site

Example prompts

  • “scrape”
  • “pull data from this URL”
  • “extract from this site”
  • “/scrape”

Requirements

  • Python 3

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. Parse User Intent
  2. Choose the Right Scrapling MCP Tool
  3. Call the MCP Tool
  4. Process Results
  5. Handle Failures

What it can do on your machine

Read from SKILL.md and the folder at commit ab5d6ad. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • pip

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • protected-site.com
    • saas.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Scrape loads about 1.2k tokens when it runs. Until then it costs about 68 tokens; SKILL.md has 387 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~68
When it runs · the whole SKILL.md, loaded when a task matches
~1.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 387 words (~1,207 tokens).

“Extract data from any website using Scrapling's MCP tools. Bypasses Cloudflare, handles dynamic JS-rendered pages, supports CSS selectors to pre-filter content (saves tokens).”

— opening of SKILL.md by davekilleen, Custom licence
name
scrape

Read the full SKILL.md on GitHub

Files

Just SKILL.md in .agents/skills/scrape of davekilleen/Dex.

Open the folder on GitHubat commit ab5d6ad

Compare with similar skills

Scrape next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Scrape compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Scrape this skilldavekilleen/Dex493—~1.2kAutomated safety check: PassCustom licence
ScraplingCedriccmh/claude-code-skill-scrapling443—~1.1kAutomated safety check: PassMIT
Anti Bot Analyzerrevfactory/harness-1001.3k—~1.1kAutomated safety check: PassApache-2.0
News9600dev/mmr131—~6.1kAutomated safety check: PassCustom licence
ScraplingTommy-yw/RunbookHermes5464 repos~2.3kAutomated safety check: PassMIT
Scraplingarchibate/dotfiles-opencode1081 repos~4.9kAutomated safety check: WarnBSD-3-Clause

Similar skills

  • Scrapling

    Cedriccmh/claude-code-skill-scrapling

    使用 scrapling 进行网页抓取和数据提取。根据目标网站特征自动选择最佳 Fetcher, 生成并执行 Python 脚本完成任务。Use when: (1) 抓取/爬取网页内容或数据(scrape, crawl, fetch page, extract data) (2) 需要绕过 Cloudflare/WAF 等反爬保护 (3) 登录后抓取受保护页面 (4) 解析已有 HTML…

    443 GitHub stars~1.1k tokensUpdated 3 mo ago
    Data & AnalyticsAuto-check passed
  • Anti Bot Analyzer

    revfactory/harness-100

    A skill for analyzing website anti-bot defense mechanisms and developing legitimate evasion strategies.

    1.3k GitHub stars~1.1k tokensUpdated 6 mo ago
    Data & AnalyticsAuto-check passed
  • News

    9600dev/mmr

    Fetch, search, and summarize financial news articles via the local scraper service (~/dev/scraper) at http://127.0.0.1:8089.

    131 GitHub stars~6.1k tokensUpdated 24 days ago
    Data & AnalyticsAuto-check passed
  • Scrapling

    Tommy-yw/RunbookHermes

    Web scraping with Scrapling - HTTP fetching, stealth browser automation, Cloudflare bypass, and spider crawling via CLI and Python.

    546 GitHub starsUsed in 4 repos~2.3k tokens
    Data & AnalyticsAuto-check passed
  • Scrapling

    archibate/dotfiles-opencode

    Scrape web pages using Scrapling with anti-bot bypass (like Cloudflare Turnstile), stealth headless browsing, spiders framework, adaptive scraping, and JavaScript rendering.

    108 GitHub starsUsed in 1 repo~4.9k tokens
    Data & AnalyticsAuto-check: warnings
  • Scrapedo Web Scraper

    artwist-polyakov/polyakov-claude-skills

    Веб-скрапинг через Scrape.do. An agent skill from artwist-polyakov/polyakov-claude-skills.

    206 GitHub stars~251 tokensUpdated 15 days ago
    Data & AnalyticsAuto-check passed

More from davekilleen/Dex

All 58 skills in this repo
  • Diff Adopt Profile

    davekilleen/Dex

    Adopt a full published Heydex profile by handle ('set me up like @davekilleen').

    493 GitHub stars~2.1k tokensUpdated 6 days ago
    Auto-check passed
  • Diff Generate

    davekilleen/Dex

    Package one workflow — how you use Dex for a specific job — into a shareable DexDiff methodology doc.

    493 GitHub stars~1.4k tokensUpdated 6 days ago
    Auto-check passed
  • Creating Agent Skills

    davekilleen/Dex

    Expert guidance for creating, writing, and refining Claude Code Skills.

    493 GitHub starsUsed in 1 repo~1.7k tokens
    Auto-check passed
  • Feedback

    davekilleen/Dex

    Report a Dex bug to the Dex team with zero homework — Dex investigates locally, builds a privacy-safe report, shows it to you (or auto-sends if you've chosen that), and tracks the ticket until it's…

    493 GitHub stars~2.4k tokensUpdated 6 days ago
    Auto-check passed
  • Dhh Rails Style

    davekilleen/Dex

    This skill should be used when writing Ruby and Rails code in DHH's distinctive 37signals style.

    493 GitHub starsUsed in 1 repo~1.7k tokens
    Auto-check passed
  • Skill Score

    davekilleen/Dex

    Grade a Dex skill against the shape-aware quality rubric and report a ship/revise/no verdict with the exact fixes.

    493 GitHub stars~2.6k tokensUpdated 6 days ago
    Auto-check passed

Works with

Questions about Scrape

What does Scrape do?

Scrape web pages via Scrapling — stealth fetching, anti-bot bypass, CSS selectors, no API key. Scrape is an agent skill from davekilleen/Dex. Scrape web pages via Scrapling — stealth fetching, anti-bot bypass, CSS selectors, no API key.

When should I use Scrape?

Scrape fits situations like: the user says scrape; pull data from this URL; extract from this site.

How do I install Scrape in Claude Code?

Run `npx skills add davekilleen/Dex --skill scrape -a claude-code`. Or copy the skill folder (.agents/skills/scrape in davekilleen/Dex) into .claude/skills/scrape in your project. Claude Code loads it when a task matches its description.

How do I install Scrape in Codex?

Run `npx skills add davekilleen/Dex --skill scrape -a codex`. Or copy the skill folder (.agents/skills/scrape in davekilleen/Dex) into .agents/skills/scrape in your project. Codex loads it when a task matches its description.

Can I use Scrape in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add davekilleen/Dex --skill scrape -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/scrape, .gemini/skills/scrape, .github/skills/scrape and .opencode/skills/scrape in your project.

What does Scrape need to run?

Going by SKILL.md and its folder, Scrape needs the command-line tools its instructions call (pip). Our summary lists: Python 3.

Does Scrape access the network?

SKILL.md names 2 domains. In commands or code: protected-site.com and saas.com; the agent is likely to contact these when it follows the instructions. This is read from the text; nothing was executed.

Is Scrape safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Scrape use?

Scrape has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Scrape use?

About 1.2k tokens (SKILL.md is roughly 4.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Scrape?

Skills that share tags, products or a category with Scrape: Scrapling (Cedriccmh/claude-code-skill-scrapling, 443 stars), Anti Bot Analyzer (revfactory/harness-100, 1.3k stars), News (9600dev/mmr, 131 stars) and Scrapling (Tommy-yw/RunbookHermes, 546 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Scrape?

davekilleen (a GitHub user) maintains it in davekilleen/Dex, which has 493 GitHub stars. The repository holds 58 skills in this directory. The repository was last updated on October 2, 2026.

Source: davekilleen/Dex on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.