Agent skill

Web Scraper

by skrun-dev in skrun-dev/skrun

Browse websites and extract structured information. An agent skill from skrun-dev/skrun.

MITAuto-check passedData & Analytics

Install Web Scraper

skills CLI
$ npx skills add skrun-dev/skrun --skill web-scraper -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install skrun-dev/skrun web-scraper --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/skrun-dev/skrun.git skills-src && mkdir -p .claude/skills && cp -r skills-src/agents/web-scraper .claude/skills/web-scraper && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
web-scraper
GitHub stars
210
Token cost
~310 tokens
SKILL.md length
161 words
Files
2
Skills in repo
14
Repo updated
First seen
Licence
MIT

At a glance

Browse websites and extract structured information. An agent skill from skrun-dev/skrun.

  • Works in 4 steps: When given a URL and a question, use… → Use browser_snapshot to get the page… → Analyze the content to answer the user's… → …
  • The user needs data from a web page
  • SKILL.md covers Instructions, Available Browser Tools, Rules and Output Format
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Web Scraper is an agent skill from skrun-dev/skrun. Browse websites and extract structured information. Use when the user needs data from a web page.

Its SKILL.md is about 310 tokens, which your agent loads only when the skill is triggered. The skill folder holds 1 other file (for example `agent.yaml`).

It sits in Data & Analytics, covering Web scraping. The repository describes itself as: Deploy any Agent Skill as an API via POST /run. The open-source multi-model alternative to Claude Managed Agents, Microsoft Foundry & Mistral/Koyeb — works with any LLM. The licence is MIT.

When your agent uses it

  • The user needs data from a web page
  • Tasks that involve Web scraping

Example prompts

  • “/web-scraper”

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. When given a URL and a question, use browser_navigate to visit the page
  2. Use browser_snapshot to get the page content (accessibility tree)
  3. Analyze the content to answer the user's question
  4. Extract the relevant information and return a structured response

What it can do on your machine

Read from SKILL.md and the folder at commit b1d963b. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Web Scraper loads about 310 tokens when it runs. Until then it costs about 27 tokens; SKILL.md has 161 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~27
When it runs · the whole SKILL.md, loaded when a task matches
~310

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from skrun-dev/skrun at commit b1d963b, republished under its MIT licence (© skrun-dev). 161 words, ~310 tokens.

Download SKILL.mdSave it as .claude/skills/web-scraper/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
web-scraper
description
Browse websites and extract structured information. Use when the user needs data from a web page.

Web Scraper

You are a web scraping agent with access to a headless browser via Playwright MCP tools.

Instructions

  1. When given a URL and a question, use browser_navigate to visit the page
  2. Use browser_snapshot to get the page content (accessibility tree)
  3. Analyze the content to answer the user's question
  4. Extract the relevant information and return a structured response

Available Browser Tools

  • browser_navigate — go to a URL
  • browser_snapshot — get page content as accessibility snapshot
  • browser_click — click an element
  • browser_type — type text into an input

Rules

  • Always use the browser tools to access web content — never make up content
  • Start with browser_navigate to the URL, then browser_snapshot to read the page
  • If the page has multiple sections, focus on what's relevant to the question
  • Keep responses factual — only report what's actually on the page

Output Format

Return a JSON object with:

  • title: the page title
  • answer: direct answer to the user's question
  • extracted_data: relevant data points from the page

© skrun-dev, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in agents/web-scraper of skrun-dev/skrun.

  • SKILL.md
  • agent.yaml

Open the folder on GitHubat commit b1d963b

Compare with similar skills

Web Scraper next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Web Scraper compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Web Scraper this skillskrun-dev/skrun210—~310Automated safety check: PassMIT
Tmuxtrpc-group/trpc-agent-go1.8k23 repos~868Automated safety check: PassApache-2.0
Ketch1broseidon/ketch6961 repos~3.9kAutomated safety check: PassMIT
Crawl4AI Web Scrapingsmallnest/goclaw5981 repos~2.5kAutomated safety check: PassMIT
Boss Zhipin Scrapereatmoreduck/boss-zhipin-scraper1.5k—~2.6kAutomated safety check: PassMIT
Axyusukebe/ax7191 repos~918Automated safety check: PassMIT

Similar skills

  • Tmux

    trpc-group/trpc-agent-go

    Remote-control tmux sessions for interactive CLIs by sending keystrokes and scraping pane output.

    1.8k GitHub starsUsed in 23 repos~868 tokens
    Data & AnalyticsAuto-check passed
  • Ketch

    1broseidon/ketch

    Research skill for ketch — a fast stateless CLI for web search, OSS code search, curated library docs, page scraping, and site crawling; an optional MCP server exists for operators who want it, but…

    696 GitHub starsUsed in 1 repo~3.9k tokens
    Data & AnalyticsAuto-check passed
  • Crawl4AI Web Scraping

    smallnest/goclaw

    Scrapes sites, handles JavaScript-heavy pages and extracts structured data with Crawl4AI, through its crwl CLI or Python SDK, including schema-based extraction without an LLM.

    598 GitHub starsUsed in 1 repo~2.5k tokens
    Data & AnalyticsAuto-check passed
  • Boss Zhipin Scraper

    eatmoreduck/boss-zhipin-scraper

    Scrape BOSS直聘 (job listing site) via Chrome CDP. An agent skill from eatmoreduck/boss-zhipin-scraper.

    1.5k GitHub stars~2.6k tokensUpdated 8 days ago
    Data & AnalyticsAuto-check passed
  • Ax

    yusukebe/ax

    Use the ax CLI instead of curl + throwaway parsing scripts whenever you fetch a URL, explore an unknown web page, or extract structured data from HTML.

    719 GitHub starsUsed in 1 repo~918 tokens
    Data & AnalyticsAuto-check passed
  • Anakinscraper

    Anakin-Inc/anakin

    Scrape any website into clean markdown or structured JSON. An agent skill from Anakin-Inc/anakin.

    4.5k GitHub stars~859 tokensUpdated 1 mo ago
    Data & AnalyticsAuto-check passed

More from skrun-dev/skrun

All 14 skills in this repo
  • Adr Writer

    skrun-dev/skrun

    Generate a numbered Architecture Decision Record (ADR) following the standard nygard/MADR convention.

    210 GitHub stars~861 tokensUpdated 15 days ago
    Auto-check passed
  • Changelog Generator

    skrun-dev/skrun

    Generate a polished CHANGELOG.md and release-notes.md from a local git repository (or a captured .git-log.txt dump).

    210 GitHub stars~770 tokensUpdated 15 days ago
    Auto-check passed
  • Turn a CSV of operational data (sales, usage, signups, support tickets) into a multi-page styled PDF executive report with narrative + matplotlib charts.

    210 GitHub stars~1.1k tokensUpdated 15 days ago
    Auto-check passed
  • Turn a folder of Markdown notes (Obsidian vault, Notion export, plain repo docs) into a navigable static HTML knowledge base bundled as a single .zip file.

    210 GitHub stars~1.4k tokensUpdated 15 days ago
    Auto-check passed
  • Listen to a meeting recording and extract structured action items, decisions, and open questions.

    210 GitHub stars~1.3k tokensUpdated 15 days ago
    Auto-check passed
  • Semgrep Rule Creator

    skrun-dev/skrun

    Generate a complete Semgrep rule bundle (rule.yml + tests.md + README.md) from a CVE description and a bad-code example.

    210 GitHub stars~1.3k tokensUpdated 15 days ago
    Auto-check passed

Questions about Web Scraper

What does Web Scraper do?

Browse websites and extract structured information. An agent skill from skrun-dev/skrun. Web Scraper is an agent skill from skrun-dev/skrun. Browse websites and extract structured information.

When should I use Web Scraper?

Web Scraper fits situations like: the user needs data from a web page; tasks that involve Web scraping.

How do I install Web Scraper in Claude Code?

Run `npx skills add skrun-dev/skrun --skill web-scraper -a claude-code`. Or copy the skill folder (agents/web-scraper in skrun-dev/skrun) into .claude/skills/web-scraper in your project. Claude Code loads it when a task matches its description.

How do I install Web Scraper in Codex?

Run `npx skills add skrun-dev/skrun --skill web-scraper -a codex`. Or copy the skill folder (agents/web-scraper in skrun-dev/skrun) into .agents/skills/web-scraper in your project. Codex loads it when a task matches its description.

Can I use Web Scraper in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add skrun-dev/skrun --skill web-scraper -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/web-scraper, .gemini/skills/web-scraper, .github/skills/web-scraper and .opencode/skills/web-scraper in your project.

What does Web Scraper need to run?

SKILL.md names no scripts, command-line tools or credentials: Web Scraper is instructions for the agent only.

Does Web Scraper access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Web Scraper safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Web Scraper use?

Web Scraper is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Web Scraper use?

About 310 tokens (SKILL.md is roughly 1.2k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Web Scraper?

Skills that share tags, products or a category with Web Scraper: Tmux (trpc-group/trpc-agent-go, 1.8k stars), Ketch (1broseidon/ketch, 696 stars), Crawl4AI Web Scraping (smallnest/goclaw, 598 stars) and Boss Zhipin Scraper (eatmoreduck/boss-zhipin-scraper, 1.5k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Web Scraper?

skrun-dev (a GitHub organization) maintains it in skrun-dev/skrun, which has 210 GitHub stars. The repository holds 14 skills in this directory. The repository was last updated on September 22, 2026.

Source: skrun-dev/skrun on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.