Agent skill

Tiered Web Browsing and Scraping

by code-yeongyu in code-yeongyu/oh-my-openagent

Routes a web request through the cheapest tier that can finish it, from headless extraction with WAF bypass up to a real stealth or signed-in browser, with screenshots as proof.

Custom licenceAuto-check: warningsProductivity & Automation

Install Tiered Web Browsing and Scraping

The automated check flagged lines worth reading first. See the safety section below.

skills CLI
$ npx skills add code-yeongyu/oh-my-openagent --skill ultimate-browsing -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install code-yeongyu/oh-my-openagent ultimate-browsing --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/code-yeongyu/oh-my-openagent.git skills-src && mkdir -p .claude/skills && cp -r skills-src/packages/shared-skills/skills/ultimate-browsing .claude/skills/ultimate-browsing && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
ultimate-browsing
GitHub stars
70k
Token cost
~2.7k tokens
SKILL.md length
830 words
Files
65 (incl. scripts, references)
Skills in repo
44
Repo updated
First seen
Licence
Custom licence

At a glance

Routes a web request through the cheapest tier that can finish it, from headless extraction with WAF bypass up to a real stealth or signed-in browser, with screenshots as proof.

  • Extracting content from a page blocked by a WAF or Cloudflare
  • SKILL.md covers PHASE 0 — ROUTE FIRST…, Tier 1 — insane-search…, Tier 1.5 — agent-reach… and Tier 2 — a real browser (real…, plus 3 more sections
  • Runs Python and JavaScript scripts from its folder; calls python3, yt-dlp and curl; reaches r.jina.ai and v2ex.com; needs JINA_API_KEY
  • Rendering and interacting with a page that needs JavaScript, a click or a form

What it does

This skill covers everything a plain fetch cannot finish: a page that renders in JavaScript, a click or a form, a screenshot, a login that must persist across pages, or a host that blocks generic fetchers with a WAF or a 403. A mandatory routing phase picks the cheapest tier first and climbs only when it cannot do the job: Tier 1 is headless extraction with WAF bypass, Tier 1.5 reaches for platform-native APIs, especially on Chinese platforms, and Tier 2 opens a real browser, either an owned stealth-mode engine the code launches or the user's own already signed-in browser.

Tier 1, called insane-search, is roughly ten times faster than spinning up a browser and handles most blocked-page requests through curl_cffi TLS impersonation, yt-dlp across more than a thousand sites, official public APIs, mobile URL transforms, provenance-tagged Wayback and archive.today snapshots, a key-gated Jina Reader, and a Playwright real-Chrome fallback, with the engine living inside the skill as a Python module invoked with a target URL and options such as a CSS selector for positive-proof validation or a device type. A result pulled from an archive snapshot is explicitly a dated copy and must be reported with its snapshot timestamp, never presented as the live page.

The skill ships its own engine code, including fetch-chain, referer and surrogate-archive modules plus Playwright templates for desktop and mobile Chrome, and an ATTRIBUTION.md file, with a reference file for each tier to read before acting.

When your agent uses it

  • Extracting content from a page blocked by a WAF or Cloudflare
  • Rendering and interacting with a page that needs JavaScript, a click or a form
  • Taking a screenshot as provenance for a research or browsing task
  • Reaching a page that needs a persistent login across requests

Example prompts

  • “Fetch the text of this blocked article page using the cheapest tier that works.”
  • “This site needs a login that persists across pages. Open a real browser session for it.”
  • “Pull this page from a Wayback snapshot and report its timestamp clearly.”
  • “Take a screenshot of this rendered page as proof for the research dossier.”

Requirements

  • Python with curl_cffi and yt-dlp for the headless extraction tier
  • Playwright with a real or stealth Chrome build for the browser tier
  • Optionally a Jina Reader API key

What it can do on your machine

Read from SKILL.md and the folder at commit cbd7dd2. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python and JavaScript, from the files we listed), which the agent can run.

    Shell commands in SKILL.md call:

    • python3
    • yt-dlp
    • curl
    • gh

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • r.jina.ai
    • v2ex.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • JINA_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Tiered Web Browsing and Scraping loads about 2.7k tokens when it runs, and up to ~25k if it reads all its reference files. Until then it costs about 75 tokens; SKILL.md has 830 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~75
When it runs · the whole SKILL.md, loaded when a task matches
~2.7k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~25k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: warnings

The automated check found patterns that need a careful read before installing.

  • WarningMentions a credentials file (SSH keys, cloud or package-manager tokens)SKILL.md:124
    n3 scripts/extract_cookies.py --browser chrome --domain youtube.com --output ~/.local/state/omo-cookies/youtube.cookies.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 830 words (~2,722 tokens).

“Web access for everything a plain fetch cannot finish: a page that renders in JS, a click or a form, a screenshot, a login that must persist across pages, or a host that blocks generic fetchers (WAF / 403 /…”

— opening of SKILL.md by code-yeongyu, Custom licence
name
ultimate-browsing

Read the full SKILL.md on GitHub

Files

SKILL.md and 64 other files (scripts, references) in packages/shared-skills/skills/ultimate-browsing of code-yeongyu/oh-my-openagent.

  • SKILL.md
  • .gitignore
  • ATTRIBUTION.md
  • engine/AGENTS.md
  • engine/__init__.py
  • engine/__main__.py
  • engine/bias_check.py
  • engine/curl_probe.py
  • engine/executor.py
  • engine/fetch_chain.py
  • engine/referers.py
  • engine/result_schema.py
  • engine/summary.py
  • engine/surrogate.py
  • engine/surrogates.yaml
  • engine/templates/package.json
  • engine/templates/playwright_mobile_chrome.js
  • engine/templates/playwright_real_chrome.js
  • engine/tests
  • … and 46 more

Open the folder on GitHubat commit cbd7dd2

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders. This page covers the copy in code-yeongyu/oh-my-openagent, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Tiered Web Browsing and Scraping next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Tiered Web Browsing and Scraping compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Tiered Web Browsing and Scraping this skillcode-yeongyu/oh-my-openagent70k—~2.7kAutomated safety check: WarnCustom licence
Anti Detect Browserantibrow/anti-detect-browser-skills914—~9.8kAutomated safety check: WarnMIT
Skyvern Browser AutomationSkyvern-AI/skyvern23k—~2.9kAutomated safety check: PassAGPL-3.0
Skyvern Browser AutomationSkyvern-AI/skyvern23k—~1.9kAutomated safety check: PassAGPL-3.0
Camofox Browserredf0x1/camofox-browser410—~4.6kAutomated safety check: PassMIT
Playwright Bowserdisler/bowser265—~1.1kAutomated safety check: NotesNone

Similar skills

  • Anti Detect Browser

    antibrow/anti-detect-browser-skills

    Drive Chromium from standard Playwright APIs with a real-device fingerprint applied in the kernel, one persistent isolated profile per identity, and a per-profile proxy whose exit IP sets timezone…

    914 GitHub stars~9.8k tokensUpdated 1 mo ago
    Testing & QAAuto-check: warnings
  • Skyvern Browser Automation

    Skyvern-AI/skyvern

    Picks the right Skyvern CLI command for a web task, from quick yes/no checks to reusable multi-page workflows, instead of falling back to plain page fetching.

    23k GitHub stars~2.9k tokensUpdated yesterday
    Productivity & AutomationAuto-check passed
  • Skyvern Browser Automation

    Skyvern-AI/skyvern

    Automates websites with Skyvern's AI browser agent to fill forms, extract data, download files, log in and run multi-step workflows through SDKs, REST, MCP or a CLI.

    23k GitHub stars~1.9k tokensUpdated yesterday
    Productivity & AutomationAuto-check passed
  • Camofox Browser

    redf0x1/camofox-browser

    Anti-detection browser automation for AI agents. An agent skill from redf0x1/camofox-browser.

    410 GitHub stars~4.6k tokensUpdated 15 days ago
    Productivity & AutomationAuto-check passed
  • Playwright Bowser

    disler/bowser

    Headless browser automation using Playwright CLI. An agent skill from disler/bowser.

    265 GitHub stars~1.1k tokensUpdated 7 mo ago
    Productivity & AutomationAuto-check: notes
  • Browser Automation

    alirezarezvani/claude-skills

    A skill your agent uses when the user asks to automate browser tasks, scrape websites, fill forms, capture screenshots, extract structured data from web pages, or build web automation workflows.

    28k GitHub stars~3.4k tokensUpdated 1 mo ago
    Productivity & AutomationAuto-check: notes

More from code-yeongyu/oh-my-openagent

All 44 skills in this repo
  • ast-grep Structural Search

    code-yeongyu/oh-my-openagent

    Searches and rewrites code by syntax-tree shape across 25 languages with ast-grep, for codemods, structural queries and YAML lint rules, using a Python wrapper script.

    70k GitHub stars~3.3k tokensUpdated today
    Auto-check passed
  • Browser Control with Omowright

    code-yeongyu/oh-my-openagent

    Drives a real browser through the omowright library, either the user's own signed-in browser or a separate browser the code launches, for forms, QA, screenshots and scraping.

    70k GitHub stars~2.2k tokensUpdated today
    Auto-check passed
  • Codex Plugin QA

    code-yeongyu/oh-my-openagent

    Tests the omo Codex plugin in an isolated CODEX_HOME with a local mock model, proving hooks fired through app-server notifications without touching ~/.codex.

    70k GitHub stars~1.9k tokensUpdated today
    Auto-check passed
  • Coding Agent Session Finder

    code-yeongyu/oh-my-openagent

    Finds, reads and reconstructs past coding-agent sessions across Codex, Claude, OpenCode, Senpi and many other local agent logs.

    70k GitHub stars~2.8k tokensUpdated today
    Auto-check passed
  • LSP Setup

    code-yeongyu/oh-my-openagent

    Detects which languages a project uses, installs the matching language server, writes its config and checks it with a real call so diagnostics and go-to-definition work.

    70k GitHub stars~1.4k tokensUpdated today
    Auto-check: notes
  • OpenCode QA Toolkit

    code-yeongyu/oh-my-openagent

    Tests the opencode coding agent itself: its CLI, server, plugin hooks and events, the terminal UI under tmux, and its SQLite session database, using tested helper scripts.

    70k GitHub stars~2.9k tokensUpdated today
    Auto-check passed

Questions about Tiered Web Browsing and Scraping

What does Tiered Web Browsing and Scraping do?

Routes a web request through the cheapest tier that can finish it, from headless extraction with WAF bypass up to a real stealth or signed-in browser, with screenshots as proof. This skill covers everything a plain fetch cannot finish: a page that renders in JavaScript, a click or a form, a screenshot, a login that must persist across pages, or a host that blocks generic fetchers with a WAF or a 403.5 reaches for platform-native APIs, especially on Chinese platforms, and Tier 2 opens a real browser, either an owned stealth-mode engine the code launches or the user's own already signed-in browser.

When should I use Tiered Web Browsing and Scraping?

Tiered Web Browsing and Scraping fits situations like: extracting content from a page blocked by a WAF or Cloudflare; rendering and interacting with a page that needs JavaScript, a click or a form; taking a screenshot as provenance for a research or browsing task; reaching a page that needs a persistent login across requests.

How do I install Tiered Web Browsing and Scraping in Claude Code?

Run `npx skills add code-yeongyu/oh-my-openagent --skill ultimate-browsing -a claude-code`. Or copy the skill folder (packages/shared-skills/skills/ultimate-browsing in code-yeongyu/oh-my-openagent) into .claude/skills/ultimate-browsing in your project. Claude Code loads it when a task matches its description.

How do I install Tiered Web Browsing and Scraping in Codex?

Run `npx skills add code-yeongyu/oh-my-openagent --skill ultimate-browsing -a codex`. Or copy the skill folder (packages/shared-skills/skills/ultimate-browsing in code-yeongyu/oh-my-openagent) into .agents/skills/ultimate-browsing in your project. Codex loads it when a task matches its description.

Can I use Tiered Web Browsing and Scraping in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add code-yeongyu/oh-my-openagent --skill ultimate-browsing -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/ultimate-browsing, .gemini/skills/ultimate-browsing, .github/skills/ultimate-browsing and .opencode/skills/ultimate-browsing in your project.

What does Tiered Web Browsing and Scraping need to run?

Going by SKILL.md and its folder, Tiered Web Browsing and Scraping needs Python and JavaScript for the scripts in its folder, the command-line tools its instructions call (python3, yt-dlp, curl and gh) and credentials named JINA_API_KEY. Our summary lists: Python with curl_cffi and yt-dlp for the headless extraction tier; Playwright with a real or stealth Chrome build for the browser tier; Optionally a Jina Reader API key.

Does Tiered Web Browsing and Scraping access the network?

SKILL.md names 2 domains. In commands or code: r.jina.ai and v2ex.com; the agent is likely to contact these when it follows the instructions. This is read from the text; nothing was executed.

Is Tiered Web Browsing and Scraping safe to install?

Our automated static check of SKILL.md flagged 1 warning(s): mentions a credentials file (ssh keys, cloud or package-manager tokens). Read the flagged lines before installing; the check is not a guarantee either way. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Tiered Web Browsing and Scraping use?

Tiered Web Browsing and Scraping has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Tiered Web Browsing and Scraping use?

About 2.7k tokens (SKILL.md is roughly 11k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 22k tokens, read only when the agent opens those files.

What are the alternatives to Tiered Web Browsing and Scraping?

Skills that share tags, products or a category with Tiered Web Browsing and Scraping: Anti Detect Browser (antibrow/anti-detect-browser-skills, 914 stars), Skyvern Browser Automation (Skyvern-AI/skyvern, 23k stars), Skyvern Browser Automation (Skyvern-AI/skyvern, 23k stars) and Camofox Browser (redf0x1/camofox-browser, 410 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Tiered Web Browsing and Scraping?

code-yeongyu (a GitHub user) maintains it in code-yeongyu/oh-my-openagent, which has 69,850 GitHub stars. The repository holds 44 skills in this directory. The repository was last updated on October 7, 2026.

Source: code-yeongyu/oh-my-openagent on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.