Agent skill

Browser Scrape

by ChanningLua in ChanningLua/prax-agent

用 AutoCLI 二进制驱动用户已登录的 Chrome 抓取 Twitter/X、知乎、Bilibili、Reddit 等 55+ 站点

MITAuto-check: notesData & Analytics

Install Browser Scrape

skills CLI
$ npx skills add ChanningLua/prax-agent --skill browser-scrape -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install ChanningLua/prax-agent browser-scrape --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/ChanningLua/prax-agent.git skills-src && mkdir -p .claude/skills && cp -r skills-src/src/prax/skills/browser-scrape .claude/skills/browser-scrape && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
browser-scrape
GitHub stars
273
Token cost
~587 tokens
SKILL.md length
159 words
Files
1
Skills in repo
12
Repo updated
First seen
Licence
MIT

At a glance

用 AutoCLI 二进制驱动用户已登录的 Chrome 抓取 Twitter/X、知乎、Bilibili、Reddit 等 55+ 站点

  • Works in 4 steps: 安装 AutoCLI Rust 二进制:参见… → 装 autocli Chrome 扩展(仓库 README 里有下载链接和加载步骤) → 保持 Chrome 运行并在目标站点处于登录态 → …
  • Tasks that involve Web scraping
  • SKILL.md covers 前置要求(由用户完成一次即可), 能力范围, 常用命令(按频次排序) and 典型抓取流程, plus 3 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Browser Scrape is an agent skill from ChanningLua/prax-agent. 用 AutoCLI 二进制驱动用户已登录的 Chrome 抓取 Twitter/X、知乎、Bilibili、Reddit 等 55+ 站点

Its SKILL.md is about 590 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Data & Analytics, covering Web scraping. It works with X (Twitter), Bilibili and Reddit. The repository describes itself as: Self-improving agent runtime that learns from experience — test-verify-fix loops, correction detection, cross-project memory, multi-model orchestration. The licence is MIT.

When your agent uses it

  • Tasks that involve Web scraping

Example prompts

  • “/browser-scrape”

Requirements

  • Pre-approved tools (allowed-tools): Bash, Write, Read, Glob

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. 安装 AutoCLI Rust 二进制:参见 https://github.com/nashsu/AutoCLI(单文件 ~4.7MB,无运行时依赖)
  2. 装 autocli Chrome 扩展(仓库 README 里有下载链接和加载步骤)
  3. 保持 Chrome 运行并在目标站点处于登录态
  4. autocli doctor 应报告 OK

What it can do on your machine

Read from SKILL.md and the folder at commit 19d016b. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Bash
    • Write
    • Read
    • Glob

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are bash).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • github.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Browser Scrape loads about 587 tokens when it runs. Until then it costs about 21 tokens; SKILL.md has 159 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~21
When it runs · the whole SKILL.md, loaded when a task matches
~587

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NotePre-approves every shell command (allowed-tools: Bash)SKILL.md
    allowed-tools: Bash, Write, Read, Glob

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from ChanningLua/prax-agent at commit 19d016b, republished under its MIT licence (© ChanningLua). 159 words, ~587 tokens.

Download SKILL.mdSave it as .claude/skills/browser-scrape/SKILL.md (or your agent's skills folder).
name
browser-scrape
description
用 AutoCLI 二进制驱动用户已登录的 Chrome 抓取 Twitter/X、知乎、Bilibili、Reddit 等 55+ 站点
allowed-tools
Bash, Write, Read, Glob
triggers
抓取, 爬取, 登录态, Chrome, scrape, 推文, timeline, hot, 热榜, zhihu, bilibili, twitter, X, reddit, autocli
tags
browser, scraping, data-collection, autocli
priority
7

浏览器抓取技能(AutoCLI)

复用用户本机 Chrome 的登录态抓取需要登录的站点(推特/X、知乎、Bilibili、Reddit、小红书等 55+ 平台)。不需要 cookie 配置、不需要 API key。

前置要求(由用户完成一次即可)

  1. 安装 AutoCLI Rust 二进制:参见 https://github.com/nashsu/AutoCLI(单文件 ~4.7MB,无运行时依赖)
  2. 装 autocli Chrome 扩展(仓库 README 里有下载链接和加载步骤)
  3. 保持 Chrome 运行并在目标站点处于登录态
  4. autocli doctor 应报告 OK

如果 autocli doctor 失败,先提醒用户修复前置条件,不要继续抓取。

能力范围

AutoCLI 在 Prax 里通过普通 Bash 工具调用——它不是 MCP server,只是一个 CLI。只要 PATH 里能找到 autocli,任何 Prax agent 都能用。

常用命令(按频次排序)

bash
# 诊断(第一次使用必跑)
autocli doctor

# 推特/X
autocli twitter timeline --limit 20 --format json
autocli twitter search --query "AI safety" --limit 10 --format json

# 知乎
autocli zhihu hot --limit 20 --format json

# Bilibili
autocli bilibili hot --limit 10 --format json

# Reddit
autocli reddit subreddit --name "MachineLearning" --limit 15 --format json

# 任意网页文章(不走登录态)
autocli read https://example.com/article --format md
autocli read https://example.com/article --format text -o /tmp/article.txt
  • --format json|md|text|yaml|csv 任选;编程任务优先 json,归档任务用 md
  • --limit N 控制条数
  • autocli --help 查所有子命令;单平台用 autocli twitter --help

典型抓取流程

用户让我"抓今天 X 上 AI 相关点赞 top 10 存到 Obsidian",我的动作:

  1. autocli doctor 确认前置条件
  2. autocli twitter timeline --limit 50 --format json 拉最近推文
  3. 本地过滤(按关键词/点赞数),用 Write 工具存到 .prax/vault/ai-news-hub/YYYY-MM-DD/ 下
  4. 每条一个 markdown 文件,header 带 tweet_id / author / likes / url / scraped_at
  5. 完成后总结文件路径给用户

产出约定(配合 knowledge-compile 技能)

  • 目录命名:.prax/vault/<topic>/YYYY-MM-DD/(方便 Obsidian 按日期归档)
  • 文件命名:<source>-<id>.md,例如 twitter-17xxx.md
  • 文件 frontmatter:至少包含 source / url / scraped_at,供下游 knowledge-compile 编译 wiki 用
  • 原始 JSON 保留:把 autocli ... --format json 的原始输出同步存一份到 raw/<source>-<stamp>.json,便于溯源

边界与禁区

  • 不要用 AutoCLI 做 写操作(发帖/点赞/关注)除非用户明确要求;无脑批量操作会导致账号封禁
  • 不要把登录 cookie 或扩展私钥外传
  • 被站点频控或返回异常时先停下来报告,不要循环重试——可能会触发风控

配合其他技能

  • 抓完 → knowledge-compile 把散文件整理成 wiki
  • 每天定时运行 → cron 里配 prax cron add 用本技能
  • 出成品 → Notify 工具把结果推到飞书/邮箱

© ChanningLua, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in src/prax/skills/browser-scrape of ChanningLua/prax-agent.

Open the folder on GitHubat commit 19d016b

Compare with similar skills

Browser Scrape next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Browser Scrape compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Browser Scrape this skillChanningLua/prax-agent273—~587Automated safety check: NotesMIT
Agent ReachPanniantong/Agent-Reach93k—~1.4kAutomated safety check: PassMIT
bb-browser Site Commands for OpenClawepiral/bb-browser6.2k—~1kAutomated safety check: PassMIT
Opinions Crawlerinfometa/workbuddyskills342—~2.2kAutomated safety check: PassNone
Pp Apifymvanhorn/printing-press-library2.1k—~6.4kAutomated safety check: NotesApache-2.0
Funding Signal Monitorgooseworks-ai/goose-skills1.2k1 repos~2.3kAutomated safety check: PassMIT

Similar skills

  • Agent Reach

    Panniantong/Agent-Reach

    Routes web research and platform lookups across 16 sites, including Twitter, Reddit, YouTube, Bilibili, Xiaohongshu and GitHub, through one command-line tool.

    93k GitHub stars~1.4k tokensUpdated 22 days ago
    Productivity & AutomationAuto-check passed
  • Runs structured data commands against sites such as Twitter, Reddit, GitHub, YouTube and Zhihu through OpenClaw's browser, reusing your existing login state.

    6.2k GitHub stars~1k tokensUpdated 4 mo ago
    Productivity & AutomationAuto-check passed
  • Opinions Crawler

    infometa/workbuddyskills

    基于 OpenCLI 的舆情数据抓取技能。覆盖国内外主流社媒(搜索、用户信息、帖子/视频列表、视频详情及互动量、评论、弹幕等)和商店平台数据爬取。当用户需要抓取舆情数据、采集社媒内容、获取商店评分评论、或者需要安装和配置 OpenCLI 时使用。

    342 GitHub stars~2.2k tokensUpdated yesterday
    Data & AnalyticsAuto-check passed
  • Pp Apify

    mvanhorn/printing-press-library

    Every Apify platform feature, plus a local SQLite store, cross-Actor search, novelty diffing, cost-aware runs, and...

    2.1k GitHub stars~6.4k tokensUpdated yesterday
    Data & AnalyticsAuto-check: notes
  • Funding Signal Monitor

    gooseworks-ai/goose-skills

    Monitor web sources for Series A-C funding announcements. An agent skill from gooseworks-ai/goose-skills.

    1.2k GitHub starsUsed in 1 repo~2.3k tokens
    Productivity & AutomationAuto-check passed
  • Trending Ad Hook Spotter

    gooseworks-ai/goose-skills

    Monitor Twitter/X, Reddit, LinkedIn, and Hacker News for trending narratives, viral posts, and hot-button topics in your space.

    1.2k GitHub starsUsed in 1 repo~2.5k tokens
    Productivity & AutomationAuto-check passed

More from ChanningLua/prax-agent

All 12 skills in this repo
  • Prax Shift

    ChanningLua/prax-agent

    Hand coding work or explicitly authorized one-shot verification to Prax Shift, inspect a shift, or schedule supported coding work with Claude as the worker.

    273 GitHub stars~1.4k tokensUpdated 26 days ago
    Auto-check passed
  • Prax Shift

    ChanningLua/prax-agent

    Hand coding work or explicitly authorized one-shot verification to Prax Shift, inspect a shift, or schedule supported coding work with Codex as the worker.

    273 GitHub stars~1.6k tokensUpdated 26 days ago
    Auto-check passed
  • AI News Daily

    ChanningLua/prax-agent

    端到端 pipeline —— 抓 X/知乎/Bilibili AI 相关热门 → 整理成 wiki → 推送飞书日报. An agent skill from ChanningLua/prax-agent.

    273 GitHub stars~1.3k tokensUpdated 26 days ago
    Auto-check: notes
  • Hotspot Article

    ChanningLua/prax-agent

    从近期大事件、真实需求和常青决策中选题,完成多源研究、业务落地、实测、事实核验和精选文章. An agent skill from ChanningLua/prax-agent.

    273 GitHub stars~2.7k tokensUpdated 26 days ago
    Auto-check: notes
  • Knowledge Compile

    ChanningLua/prax-agent

    把一堆 raw markdown(抓取/笔记/文章)压成 Obsidian 风格 wiki —— 有 TOC、有按主题聚合、有日简报

    273 GitHub stars~914 tokensUpdated 26 days ago
    Auto-check: notes
  • PR Triage

    ChanningLua/prax-agent

    给单个 PR 做代码感知 triage —— 分类 / 跑测试 / 扫依赖 / 产出审查笔记. An agent skill from ChanningLua/prax-agent.

    273 GitHub stars~1.4k tokensUpdated 26 days ago
    Auto-check: notes

Questions about Browser Scrape

What does Browser Scrape do?

用 AutoCLI 二进制驱动用户已登录的 Chrome 抓取 Twitter/X、知乎、Bilibili、Reddit 等 55+ 站点. Browser Scrape is an agent skill from ChanningLua/prax-agent.

When should I use Browser Scrape?

Browser Scrape fits situations like: tasks that involve Web scraping.

How do I install Browser Scrape in Claude Code?

Run `npx skills add ChanningLua/prax-agent --skill browser-scrape -a claude-code`. Or copy the skill folder (src/prax/skills/browser-scrape in ChanningLua/prax-agent) into .claude/skills/browser-scrape in your project. Claude Code loads it when a task matches its description.

How do I install Browser Scrape in Codex?

Run `npx skills add ChanningLua/prax-agent --skill browser-scrape -a codex`. Or copy the skill folder (src/prax/skills/browser-scrape in ChanningLua/prax-agent) into .agents/skills/browser-scrape in your project. Codex loads it when a task matches its description.

Can I use Browser Scrape in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add ChanningLua/prax-agent --skill browser-scrape -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/browser-scrape, .gemini/skills/browser-scrape, .github/skills/browser-scrape and .opencode/skills/browser-scrape in your project.

What does Browser Scrape need to run?

SKILL.md names no scripts, command-line tools or credentials: Browser Scrape is instructions for the agent only. Its frontmatter pre-approves these tools: Bash, Write, Read, Glob.

Does Browser Scrape access the network?

SKILL.md names 1 domain. As links in the text: github.com. This is read from the text; nothing was executed.

Is Browser Scrape safe to install?

Our automated static check of SKILL.md found notes only (pre-approves every shell command (allowed-tools: bash)), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.

What licence does Browser Scrape use?

Browser Scrape is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Browser Scrape use?

About 587 tokens (SKILL.md is roughly 2.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Browser Scrape?

Skills that share tags, products or a category with Browser Scrape: Agent Reach (Panniantong/Agent-Reach, 93k stars), bb-browser Site Commands for OpenClaw (epiral/bb-browser, 6.2k stars), Opinions Crawler (infometa/workbuddyskills, 342 stars) and Pp Apify (mvanhorn/printing-press-library, 2.1k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Browser Scrape?

ChanningLua (a GitHub user) maintains it in ChanningLua/prax-agent, which has 273 GitHub stars. The repository holds 12 skills in this directory. The repository was last updated on September 11, 2026.

Source: ChanningLua/prax-agent on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.