Agent skill

Web Content Fetcher

by yaomindong1996 in yaomindong1996/forge-admin

Extract article content from any URL as clean Markdown. An agent skill from yaomindong1996/forge-admin.

Apache-2.0Auto-check passedWriting & Content

Install Web Content Fetcher

skills CLI
$ npx skills add yaomindong1996/forge-admin --skill web-content-fetcher -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install yaomindong1996/forge-admin web-content-fetcher --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/yaomindong1996/forge-admin.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/web-content-fetcher .claude/skills/web-content-fetcher && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
web-content-fetcher
GitHub stars
127
Token cost
~898 tokens
SKILL.md length
251 words
Files
2 (incl. scripts)
Skills in repo
4
Repo updated
First seen
Licence
Apache-2.0

At a glance

Extract article content from any URL as clean Markdown. An agent skill from yaomindong1996/forge-admin.

  • The user wants to fetch
  • SKILL.md covers Extraction Strategy, Domain Routing, Script Options and Install Dependencies, plus 1 more section
  • Runs Python scripts from its folder; calls python3 and pip; reaches r.jina.ai and sspai.com
  • Summarize content from a URL — including blog posts

What it does

Web Content Fetcher is an agent skill from yaomindong1996/forge-admin. Extract article content from any URL as clean Markdown. Uses Scrapling script as primary method (with auto fast→stealth fallback), Jina Reader as alternative for simple pages. Preserves headings, links, images, lists, and code blocks. Use this skill whenever the user wants to fetch, read, extract, scrape, or summarize content from a URL — including blog posts, news articles, WeChat articles (微信公众号), documentation pages, or any web page. Also trigger when the user says things like "帮我读一下这篇文章", "抓取这个网页", "提取正文", or…

Its SKILL.md is about 900 tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/fetch.py`).

It sits in Writing & Content, covering Web scraping, Blog and article writing and Messaging and chat bots. It works with WeChat. The repository describes itself as: AI编码,SpringBoot 3.x + JDK 17 构建的轻量化企业级管理系统基础框架,以配置驱动为核心设计理念,追求简洁高效、开箱即用,助力开发者快速搭建稳定可靠的企业级应用,极简开发、高效迭代、生产可用. The licence is Apache-2.0.

When your agent uses it

  • The user wants to fetch
  • Summarize content from a URL — including blog posts
  • WeChat articles (微信公众号)
  • Documentation pages

Example prompts

  • “帮我读一下这篇文章”
  • “抓取这个网页”
  • “read this page for me”
  • “/web-content-fetcher”

Requirements

  • Python 3

What it can do on your machine

Read from SKILL.md and the folder at commit 719a2fc. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python3
    • pip

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • r.jina.ai
    • sspai.com
    • mp.weixin.qq.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Web Content Fetcher loads about 898 tokens when it runs. Until then it costs about 141 tokens; SKILL.md has 251 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~141
When it runs · the whole SKILL.md, loaded when a task matches
~898

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from yaomindong1996/forge-admin at commit 719a2fc, republished under its Apache-2.0 licence (© yaomindong1996). 251 words, ~898 tokens.

Download SKILL.mdSave it as .claude/skills/web-content-fetcher/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
web-content-fetcher
description
Extract article content from any URL as clean Markdown. Uses Scrapling script as primary method (with auto fast→stealth fallback), Jina Reader as alternative for simple pages. Preserves headings, links, images, lists, and code blocks. Use this skill whenever the user wants to fetch, read, extract, scrape, or summarize content from a URL — including blog posts, news articles, WeChat articles (微信公众号), documentation pages, or any web page. Also trigger when the user says things like "帮我读一下这篇文章", "抓取这个网页", "提取正文", or "read this page for me".

Web Content Fetcher

Given a URL, return its main content as clean Markdown — headings, links, images, lists, code blocks all preserved.

Extraction Strategy

Always try one method per URL — don't cascade blindly. Pick the right one upfront.

URL
 │
 ├─ 1. Scrapling script (preferred)
 │     Run fetch.py — check the domain routing table to decide fast vs --stealth.
 │     Works for most sites. Returns clean Markdown directly.
 │
 └─ 2. Jina Reader (fallback — only if Scrapling fails or dependencies not installed)
       web_fetch("https://r.jina.ai/<url>")
       Free tier: 200 req/day. Fast (~1-2s), good Markdown output.
       Does NOT work for: WeChat (403), some Chinese platforms.
Scrapling script
bash
python3 <SKILL_DIR>/scripts/fetch.py "<url>" [max_chars] [--stealth]

<SKILL_DIR> is the directory where this SKILL.md lives. Resolve it before calling the script.

The script has two modes built in:

  • Default (fast): HTTP fetch, ~1-3s, works for most sites
  • --stealth: Headless browser, ~5-15s, for JS-rendered or anti-scraping sites

When run without --stealth, the script automatically falls back to stealth if the fast result has too little content. So you rarely need to specify --stealth manually — the only reason to force it is when you already know the site needs it (see routing table), which saves the initial fast attempt.

Domain Routing

Use this table to pick the right mode on the first call:

DomainCommandWhy
mp.weixin.qq.comfetch.py <url> --stealthJS-rendered content
zhuanlan.zhihu.comfetch.py <url> --stealthAnti-scraping + JS
juejin.cnfetch.py <url> --stealthJS-rendered SPA
sspai.comfetch.py <url>Static HTML
blog.csdn.netfetch.py <url>Static HTML
ruanyifeng.comfetch.py <url>Static blog
openai.comfetch.py <url>Static HTML
blog.googlefetch.py <url>Static HTML
Everything elsefetch.py <url>Auto-fallback handles it

Script Options

bash
# Basic — auto-selects fast or stealth
python3 <SKILL_DIR>/scripts/fetch.py "https://sspai.com/post/73145"

# Force stealth for known JS-heavy sites
python3 <SKILL_DIR>/scripts/fetch.py "https://mp.weixin.qq.com/s/xxx" --stealth

# Limit output to 15000 characters (default: 30000)
python3 <SKILL_DIR>/scripts/fetch.py "https://example.com/article" 15000

# JSON output with metadata (url, mode, selector, content_length)
python3 <SKILL_DIR>/scripts/fetch.py "https://example.com" --json

Install Dependencies

First use only — the script checks and tells you if anything is missing:

bash
pip install scrapling html2text

If on system-managed Python (macOS/Linux), add --break-system-packages or use a venv.

Failure Rules

  • Same URL fails once → give up, tell the user "unable to extract content from this URL"
  • Do not retry — each failed call wastes context tokens

© yaomindong1996, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (scripts) in .agents/skills/web-content-fetcher of yaomindong1996/forge-admin.

  • SKILL.md
  • scripts/fetch.py

Open the folder on GitHubat commit 719a2fc

Compare with similar skills

Web Content Fetcher next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Web Content Fetcher compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Web Content Fetcher this skillyaomindong1996/forge-admin127—~898Automated safety check: PassApache-2.0
Xiaohongshu Note Creatorwpsnote/wpsnote-skills179—~3kAutomated safety check: PassNone
Outreach Communicationinfometa/workbuddyskills348—~1.3kAutomated safety check: PassNone
WeChat Article Publisherjiji262/wechat-publisher274—~4.4kAutomated safety check: PassNone
WeChat Official Account Pipelineaiworkskills/wechat-article-skills672—~2kAutomated safety check: PassApache-2.0
Wechat Article Search API Skillbrowser-act/skills6.1k1 repos~1.4kAutomated safety check: PassMIT

Similar skills

  • Xiaohongshu Note Creator

    wpsnote/wpsnote-skills

    【笔记/文章转小红书】将用户已有的 WPS 笔记或文章内容,改写压缩为小红书图文方案. An agent skill from wpsnote/wpsnote-skills.

    179 GitHub stars~3k tokensUpdated 4 mo ago
    Writing & ContentAuto-check passed
  • Outreach Communication

    infometa/workbuddyskills

    This skill should be used when the user is NOT a parent seeking counseling, but instead a researcher, journalist, editor, public speaker, policy maker, content creator, or training instructor who…

    348 GitHub stars~1.3k tokensUpdated yesterday
    Writing & ContentAuto-check passed
  • WeChat Article Publisher

    jiji262/wechat-publisher

    Researches a topic, writes an illustrated WeChat Official Account article in a chosen author voice and sends it to the account's draft box.

    274 GitHub stars~4.4k tokensUpdated 2 mo ago
    Writing & ContentAuto-check passed
  • WeChat Official Account Pipeline

    aiworkskills/wechat-article-skills

    Orchestrates the full WeChat official account workflow from topic to draft, chaining sub-skills for writing, review, layout, images and publishing.

    672 GitHub stars~2k tokensUpdated 18 days ago
    Writing & ContentAuto-check passed
  • This skill helps users extract full article contents from WeChat using the BrowserAct API.

    6.1k GitHub starsUsed in 1 repo~1.4k tokens
    Productivity & AutomationAuto-check passed
  • Wechat Article Writer

    LeoYeAI/openclaw-master-skills

    WeChat Official Account article writing assistant. An agent skill from LeoYeAI/openclaw-master-skills.

    2.2k GitHub stars~2.5k tokensUpdated 2 mo ago
    Writing & ContentAuto-check passed

More from yaomindong1996/forge-admin

  • Forge Project Init

    yaomindong1996/forge-admin

    Bootstrap or maintain projects generated from Forge, initialize a clean template database, and install, upgrade or remove Forge source plugins in template or generated Admin projects.

    127 GitHub stars~2.2k tokensUpdated today
    Auto-check: notes
  • Forge Business Flow Development

    yaomindong1996/forge-admin

    Develop Forge business approval workflows backed by Flowable, using the current sample purchase order approval as the reference.

    127 GitHub stars~1.3k tokensUpdated today
    Auto-check passed
  • Forge Codegen Crud

    yaomindong1996/forge-admin

    Generate or review Forge project code-generation output for CRUD modules.

    127 GitHub stars~1.3k tokensUpdated today
    Auto-check passed

Works with

Questions about Web Content Fetcher

What does Web Content Fetcher do?

Extract article content from any URL as clean Markdown. An agent skill from yaomindong1996/forge-admin. Web Content Fetcher is an agent skill from yaomindong1996/forge-admin. Extract article content from any URL as clean Markdown.

When should I use Web Content Fetcher?

Web Content Fetcher fits situations like: the user wants to fetch; summarize content from a URL — including blog posts; weChat articles (微信公众号); documentation pages.

How do I install Web Content Fetcher in Claude Code?

Run `npx skills add yaomindong1996/forge-admin --skill web-content-fetcher -a claude-code`. Or copy the skill folder (.agents/skills/web-content-fetcher in yaomindong1996/forge-admin) into .claude/skills/web-content-fetcher in your project. Claude Code loads it when a task matches its description.

How do I install Web Content Fetcher in Codex?

Run `npx skills add yaomindong1996/forge-admin --skill web-content-fetcher -a codex`. Or copy the skill folder (.agents/skills/web-content-fetcher in yaomindong1996/forge-admin) into .agents/skills/web-content-fetcher in your project. Codex loads it when a task matches its description.

Can I use Web Content Fetcher in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add yaomindong1996/forge-admin --skill web-content-fetcher -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/web-content-fetcher, .gemini/skills/web-content-fetcher, .github/skills/web-content-fetcher and .opencode/skills/web-content-fetcher in your project.

What does Web Content Fetcher need to run?

Going by SKILL.md and its folder, Web Content Fetcher needs Python for the scripts in its folder and the command-line tools its instructions call (python3 and pip). Our summary lists: Python 3.

Does Web Content Fetcher access the network?

SKILL.md names 3 domains. In commands or code: r.jina.ai, sspai.com and mp.weixin.qq.com; the agent is likely to contact these when it follows the instructions. This is read from the text; nothing was executed.

Is Web Content Fetcher safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Web Content Fetcher use?

Web Content Fetcher is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Web Content Fetcher use?

About 898 tokens (SKILL.md is roughly 3.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Web Content Fetcher?

Skills that share tags, products or a category with Web Content Fetcher: Xiaohongshu Note Creator (wpsnote/wpsnote-skills, 179 stars), Outreach Communication (infometa/workbuddyskills, 348 stars), WeChat Article Publisher (jiji262/wechat-publisher, 274 stars) and WeChat Official Account Pipeline (aiworkskills/wechat-article-skills, 672 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Web Content Fetcher?

yaomindong1996 (a GitHub user) maintains it in yaomindong1996/forge-admin, which has 127 GitHub stars. The repository holds 4 skills in this directory. The repository was last updated on October 11, 2026.

Source: yaomindong1996/forge-admin on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.