Agent skill

Web Scraping

by nicepkg in nicepkg/auto-company

Web scraping with anti-bot bypass, content extraction, undocumented APIs and poison pill detection.

No licenceAuto-check passedData & Analytics

Install Web Scraping

skills CLI
$ npx skills add nicepkg/auto-company --skill web-scraping -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install nicepkg/auto-company web-scraping --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/nicepkg/auto-company.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/web-scraping .claude/skills/web-scraping && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
web-scraping
GitHub stars
192
Used in
1 other repo
Token cost
~5k tokens
SKILL.md length
242 words
Files
1
Skills in repo
23
Repo updated
First seen
Licence
None found

At a glance

Web scraping with anti-bot bypass, content extraction, undocumented APIs and poison pill detection.

  • Works in 7 steps: Open developer tools (right-click →… → Go to the Network tab to monitor all… → Filter by Fetch/XHR to show only API calls → …
  • Extracting content from websites
  • SKILL.md covers Scraping cascade architecture, Undocumented APIs, Poison pill detection and Social media scraping, plus 2 more sections
  • Reaches completion.amazon.com and instagram.com

What it does

Web Scraping is an agent skill from nicepkg/auto-company. Web scraping with anti-bot bypass, content extraction, undocumented APIs and poison pill detection. Use when extracting content from websites, handling paywalls, implementing scraping cascades or processing social media. Covers requests, trafilatura, Playwright with stealth mode, yt-dlp and instaloader patterns.

Its SKILL.md is about 5k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Data & Analytics, covering Web scraping. It works with Playwright. The repository describes itself as: 🤖 A fully autonomous AI company that runs 24/7. 14 AI agents (Bezos, Munger, DHH...) brainstorm ideas, write code, deploy products & make money — no human in the loop. Powered…

When your agent uses it

  • Extracting content from websites
  • Handling paywalls
  • Implementing scraping cascades
  • Processing social media

Example prompts

  • “/web-scraping”

Requirements

  • Python 3

Workflow steps

7 steps, taken from the first numbered list in SKILL.md.

  1. Open developer tools (right-click → Inspect, or F12)
  2. Go to the Network tab to monitor all requests
  3. Filter by Fetch/XHR to show only API calls
  4. Trigger the action you want to capture (search, scroll, click)
  5. Analyze the response — usually JSON with key-value pairs
  6. Copy as cURL (right-click the request)
  7. Convert to code using curlconverter.com

What it can do on your machine

Read from SKILL.md and the folder at commit 1252920. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are python).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • completion.amazon.com
    • instagram.com
    • tiktok.com

    Also links to:

    • curlconverter.com
    • inspectelement.org

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Web Scraping loads about 5k tokens when it runs. Until then it costs about 82 tokens; SKILL.md has 242 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~82
When it runs · the whole SKILL.md, loaded when a task matches
~5k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 242 words (~4,961 tokens).

“Patterns for reliable, ethical web scraping with fallback strategies and anti-bot handling.”

— opening of SKILL.md by nicepkg
name
web-scraping

Read the full SKILL.md on GitHub

Files

Just SKILL.md in .claude/skills/web-scraping of nicepkg/auto-company.

Open the folder on GitHubat commit 1252920

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in nicepkg/auto-company, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Web Scraping next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Web Scraping compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Web Scraping this skillnicepkg/auto-company1921 repos~5kAutomated safety check: PassNone
AnakinscraperAnakin-Inc/anakin4.5k—~859Automated safety check: PassAGPL-3.0
Python Executorcortega26/chile-hub1132 repos~1.5kAutomated safety check: PassMIT
Web Crawlerbyungjunjang/web-crawler162—~7.2kAutomated safety check: PassMIT
Scraplingforyourhealth111-pixel/Vibe-Skills3.6k—~1.1kAutomated safety check: PassApache-2.0
Scraper Builderjwynia/agent-skills166—~4kAutomated safety check: PassMIT

Similar skills

  • Anakinscraper

    Anakin-Inc/anakin

    Scrape any website into clean markdown or structured JSON. An agent skill from Anakin-Inc/anakin.

    4.5k GitHub stars~859 tokensUpdated 1 mo ago
    Data & AnalyticsAuto-check passed
  • Python Executor

    cortega26/chile-hub

    Execute Python code in a safe sandboxed environment via [inference.sh](https://inference.sh).

    113 GitHub starsUsed in 2 repos~1.5k tokens
    Data & AnalyticsAuto-check passed
  • Web Crawler

    byungjunjang/web-crawler

    URL과 수집 항목을 받아 사이트를 정찰하고 데이터를 수집하여 엑셀로 출력하는 범용 웹 크롤링 에이전트. An agent skill from byungjunjang/web-crawler.

    162 GitHub stars~7.2k tokensUpdated 2 days ago
    Data & AnalyticsAuto-check passed
  • Scrapling

    foryourhealth111-pixel/Vibe-Skills

    CLI-first web scraping & content extraction with optional MCP server.

    3.6k GitHub stars~1.1k tokensUpdated 1 mo ago
    Data & AnalyticsAuto-check passed
  • Scraper Builder

    jwynia/agent-skills

    Guide AI agents to generate complete PageObject pattern web scraper projects using Playwright and TypeScript with Docker deployment.

    166 GitHub stars~4k tokensUpdated 7 mo ago
    Data & AnalyticsAuto-check passed
  • Twitter X Scraping

    swyxio/skills

    A skill your agent uses when scraping public Twitter/X timelines or lists through Nitter-style mirrors with axios and Playwright fallback, including anti-bot challenge handling and pagination limits.

    175 GitHub stars~3.5k tokensUpdated 3 days ago
    Data & AnalyticsAuto-check passed

More from nicepkg/auto-company

All 23 skills in this repo
  • Senior QA

    nicepkg/auto-company

    Comprehensive QA and testing skill for quality assurance, test automation, and testing strategies for ReactJS, NextJS, NodeJS applications.

    192 GitHub starsUsed in 3 repos~1.1k tokens
    Auto-check: notes
  • Market Sizing Analysis

    nicepkg/auto-company

    This skill should be used when the user asks to "calculate TAM", "determine SAM", "estimate SOM", "size the market", "calculate market opportunity", "what's the total addressable market", or…

    192 GitHub starsUsed in 12 repos~3.1k tokens
    Auto-check passed
  • Devops

    nicepkg/auto-company

    Deploy to Cloudflare (Workers, R2, D1), Docker, GCP (Cloud Run, GKE), Kubernetes (kubectl, Helm).

    192 GitHub starsUsed in 2 repos~814 tokens
    Auto-check passed
  • Startup Financial Modeling

    nicepkg/auto-company

    This skill should be used when the user asks to "create financial projections", "build a financial model", "forecast revenue", "calculate burn rate", "estimate runway", "model cash flow", or…

    192 GitHub starsUsed in 12 repos~2.8k tokens
    Auto-check passed
  • Deep Research

    nicepkg/auto-company

    Conduct enterprise-grade research with multi-source synthesis, citation tracking, and verification.

    192 GitHub starsUsed in 4 repos~8.3k tokens
    Auto-check passed
  • Micro SaaS Launcher

    nicepkg/auto-company

    Expert in launching small, focused SaaS products fast - the indie hacker approach to building profitable software.

    192 GitHub starsUsed in 9 repos~1.3k tokens
    Auto-check passed

Works with

Questions about Web Scraping

What does Web Scraping do?

Web scraping with anti-bot bypass, content extraction, undocumented APIs and poison pill detection. Web Scraping is an agent skill from nicepkg/auto-company. Web scraping with anti-bot bypass, content extraction, undocumented APIs and poison pill detection.

When should I use Web Scraping?

Web Scraping fits situations like: extracting content from websites; handling paywalls; implementing scraping cascades; processing social media.

How do I install Web Scraping in Claude Code?

Run `npx skills add nicepkg/auto-company --skill web-scraping -a claude-code`. Or copy the skill folder (.claude/skills/web-scraping in nicepkg/auto-company) into .claude/skills/web-scraping in your project. Claude Code loads it when a task matches its description.

How do I install Web Scraping in Codex?

Run `npx skills add nicepkg/auto-company --skill web-scraping -a codex`. Or copy the skill folder (.claude/skills/web-scraping in nicepkg/auto-company) into .agents/skills/web-scraping in your project. Codex loads it when a task matches its description.

Can I use Web Scraping in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add nicepkg/auto-company --skill web-scraping -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/web-scraping, .gemini/skills/web-scraping, .github/skills/web-scraping and .opencode/skills/web-scraping in your project.

What does Web Scraping need to run?

SKILL.md names no scripts, command-line tools or credentials: Web Scraping is instructions for the agent only. Our summary lists: Python 3.

Does Web Scraping access the network?

SKILL.md names 5 domains. In commands or code: completion.amazon.com, instagram.com and tiktok.com; the agent is likely to contact these when it follows the instructions. As links in the text: curlconverter.com and inspectelement.org. This is read from the text; nothing was executed.

Is Web Scraping safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Web Scraping use?

No licence was found for Web Scraping or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Web Scraping use?

About 5k tokens (SKILL.md is roughly 20k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Web Scraping?

Skills that share tags, products or a category with Web Scraping: Anakinscraper (Anakin-Inc/anakin, 4.5k stars), Python Executor (cortega26/chile-hub, 113 stars), Web Crawler (byungjunjang/web-crawler, 162 stars) and Scrapling (foryourhealth111-pixel/Vibe-Skills, 3.6k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Web Scraping?

nicepkg (a GitHub organization) maintains it in nicepkg/auto-company, which has 192 GitHub stars. The repository holds 23 skills in this directory. The repository was last updated on February 12, 2026.

Source: nicepkg/auto-company on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.