Agent skill

Scrape Webpage

by NeverSight in NeverSight/learn-skills.dev

Scrape webpage content, extract metadata, download images, and prepare for import/migration to AEM Edge Delivery Services.

No licenceAuto-check passedData & Analytics

Install Scrape Webpage

skills CLI
$ npx skills add NeverSight/learn-skills.dev --skill scrape-webpage -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install NeverSight/learn-skills.dev scrape-webpage --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/NeverSight/learn-skills.dev.git skills-src && mkdir -p .claude/skills && cp -r skills-src/data/skills-md/adobe/helix-website/scrape-webpage .claude/skills/scrape-webpage && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
scrape-webpage
GitHub stars
216
Used in
1 other repo
Token cost
~1.2k tokens
SKILL.md length
383 words
Files
13
Skills in repo
43
Repo updated
First seen
Licence
None found

At a glance

Scrape webpage content, extract metadata, download images, and prepare for import/migration to AEM Edge Delivery Services.

  • Works in 3 steps: Run Analysis Script → Verify Output → Review Metadata JSON
  • Tasks that involve Web scraping
  • SKILL.md covers When to Use This Skill, Prerequisites, Related Skills and Scraping Workflow, plus 2 more sections
  • Calls npm, npx and node

What it does

Scrape Webpage is an agent skill from NeverSight/learn-skills.dev. Scrape webpage content, extract metadata, download images, and prepare for import/migration to AEM Edge Delivery Services. Returns analysis JSON with paths, metadata, cleaned HTML, and local images.

Its SKILL.md is about 1.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 12 other files (for example `stats.json`).

It sits in Data & Analytics, covering Web scraping. It works with Adobe Experience Manager, npm and Playwright. The repository describes itself as: Curated high-quality AI Agent Skills. Search, install, copy and share. Works with Claude Code, Cursor, OpenClaw, and other AI coding tools.

When your agent uses it

  • Tasks that involve Web scraping

Example prompts

  • “/scrape-webpage”

Requirements

  • Node.js

Workflow steps

3 steps, taken from the step headings in SKILL.md.

  1. Run Analysis Script
  2. Verify Output
  3. Review Metadata JSON

What it can do on your machine

Read from SKILL.md and the folder at commit 08f9d22. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • npm
    • npx
    • node

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use npm and npx, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Scrape Webpage loads about 1.2k tokens when it runs. Until then it costs about 53 tokens; SKILL.md has 383 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~53
When it runs · the whole SKILL.md, loaded when a task matches
~1.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 383 words (~1,175 tokens).

“Extract content, metadata, and images from a webpage for import/migration.”

— opening of SKILL.md by NeverSight
name
scrape-webpage

Read the full SKILL.md on GitHub

Files

SKILL.md and 12 other files in data/skills-md/adobe/helix-website/scrape-webpage of NeverSight/learn-skills.dev.

  • SKILL.md
  • description_ar.txt
  • description_cn.txt
  • description_de.txt
  • description_en.txt
  • description_es.txt
  • description_fr.txt
  • description_it.txt
  • description_ja.txt
  • description_ko.txt
  • description_ru.txt
  • description_tw.txt
  • stats.json

Open the folder on GitHubat commit 08f9d22

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in NeverSight/learn-skills.dev, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Scrape Webpage next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Scrape Webpage compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Scrape Webpage this skillNeverSight/learn-skills.dev2161 repos~1.2kAutomated safety check: PassNone
Anti Detect Browserantibrow/anti-detect-browser-skills914—~9.8kAutomated safety check: WarnMIT
AnakinscraperAnakin-Inc/anakin4.5k—~859Automated safety check: PassAGPL-3.0
Python Executorcortega26/chile-hub1132 repos~1.5kAutomated safety check: PassMIT
Chrome Devtoolseinverne/dotfiles1211 repos~1.6kAutomated safety check: NotesApache-2.0
Web Crawlerbyungjunjang/web-crawler162—~7.2kAutomated safety check: PassMIT

Similar skills

  • Anti Detect Browser

    antibrow/anti-detect-browser-skills

    Drive Chromium from standard Playwright APIs with a real-device fingerprint applied in the kernel, one persistent isolated profile per identity, and a per-profile proxy whose exit IP sets timezone…

    914 GitHub stars~9.8k tokensUpdated 1 mo ago
    Testing & QAAuto-check: warnings
  • Anakinscraper

    Anakin-Inc/anakin

    Scrape any website into clean markdown or structured JSON. An agent skill from Anakin-Inc/anakin.

    4.5k GitHub stars~859 tokensUpdated 1 mo ago
    Data & AnalyticsAuto-check passed
  • Python Executor

    cortega26/chile-hub

    Execute Python code in a safe sandboxed environment via [inference.sh](https://inference.sh).

    113 GitHub starsUsed in 2 repos~1.5k tokens
    Data & AnalyticsAuto-check passed
  • Chrome Devtools

    einverne/dotfiles

    Browser automation, debugging, and performance analysis using Puppeteer CLI scripts.

    121 GitHub starsUsed in 1 repo~1.6k tokens
    Data & AnalyticsAuto-check: notes
  • Web Crawler

    byungjunjang/web-crawler

    URL과 수집 항목을 받아 사이트를 정찰하고 데이터를 수집하여 엑셀로 출력하는 범용 웹 크롤링 에이전트. An agent skill from byungjunjang/web-crawler.

    162 GitHub stars~7.2k tokensUpdated yesterday
    Data & AnalyticsAuto-check passed
  • Multi Account Scraping

    antibrow/anti-detect-browser-skills

    Run the same scrape or task across many accounts at once - each in its own browser profile with its own fingerprint, cookies and exit IP - and read data from sites that need a session or that answer…

    914 GitHub stars~3.7k tokensUpdated 1 mo ago
    Data & AnalyticsAuto-check: warnings

More from NeverSight/learn-skills.dev

All 43 skills in this repo
  • AI Marketing Videos

    NeverSight/learn-skills.dev

    Create AI marketing videos for ads, promos, product launches, and brand content.

    216 GitHub starsUsed in 3 repos~2.1k tokens
    Auto-check passed
  • Google Calendar

    NeverSight/learn-skills.dev

    Interact with Google Calendar via the Google Calendar API – list upcoming events, create new events, update or delete them.

    216 GitHub starsUsed in 2 repos~826 tokens
    Auto-check passed
  • Agent Orchestrator

    NeverSight/learn-skills.dev

    Meta-agent skill for orchestrating complex tasks through autonomous sub-agents.

    216 GitHub starsUsed in 1 repo~1.4k tokens
    Auto-check passed
  • AI Automation Workflows

    NeverSight/learn-skills.dev

    Build automated AI workflows combining multiple models and services.

    216 GitHub starsUsed in 1 repo~2.6k tokens
    Auto-check passed
  • AI Content Pipeline

    NeverSight/learn-skills.dev

    Build multi-step AI content creation pipelines combining image, video, audio, and text.

    216 GitHub starsUsed in 1 repo~1.8k tokens
    Auto-check passed
  • AI Podcast Creation

    NeverSight/learn-skills.dev

    Create AI-powered podcasts with text-to-speech, music, and audio editing.

    216 GitHub starsUsed in 1 repo~2k tokens
    Auto-check passed

Questions about Scrape Webpage

What does Scrape Webpage do?

Scrape webpage content, extract metadata, download images, and prepare for import/migration to AEM Edge Delivery Services. dev. Scrape webpage content, extract metadata, download images, and prepare for import/migration to AEM Edge Delivery Services.

When should I use Scrape Webpage?

Scrape Webpage fits situations like: tasks that involve Web scraping.

How do I install Scrape Webpage in Claude Code?

Run `npx skills add NeverSight/learn-skills.dev --skill scrape-webpage -a claude-code`. Or copy the skill folder (data/skills-md/adobe/helix-website/scrape-webpage in NeverSight/learn-skills.dev) into .claude/skills/scrape-webpage in your project. Claude Code loads it when a task matches its description.

How do I install Scrape Webpage in Codex?

Run `npx skills add NeverSight/learn-skills.dev --skill scrape-webpage -a codex`. Or copy the skill folder (data/skills-md/adobe/helix-website/scrape-webpage in NeverSight/learn-skills.dev) into .agents/skills/scrape-webpage in your project. Codex loads it when a task matches its description.

Can I use Scrape Webpage in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add NeverSight/learn-skills.dev --skill scrape-webpage -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/scrape-webpage, .gemini/skills/scrape-webpage, .github/skills/scrape-webpage and .opencode/skills/scrape-webpage in your project.

What does Scrape Webpage need to run?

Going by SKILL.md and its folder, Scrape Webpage needs the command-line tools its instructions call (npm, npx and node). Our summary lists: Node.js.

Does Scrape Webpage access the network?

SKILL.md contains no URLs. Its commands use npm and npx, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Scrape Webpage safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Scrape Webpage use?

No licence was found for Scrape Webpage or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Scrape Webpage use?

About 1.2k tokens (SKILL.md is roughly 4.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Scrape Webpage?

Skills that share tags, products or a category with Scrape Webpage: Anti Detect Browser (antibrow/anti-detect-browser-skills, 914 stars), Anakinscraper (Anakin-Inc/anakin, 4.5k stars), Python Executor (cortega26/chile-hub, 113 stars) and Chrome Devtools (einverne/dotfiles, 121 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Scrape Webpage?

NeverSight (a GitHub organization) maintains it in NeverSight/learn-skills.dev, which has 216 GitHub stars. The repository holds 43 skills in this directory. The repository was last updated on October 6, 2026.

Source: NeverSight/learn-skills.dev on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.