Topic · Data & Analytics

Best web scraping skills for Claude Code, Codex and other agents.

Skills that collect structured data from websites and web APIs.
skills
717
official
41

Web scraping skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Web scraping skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Gets Firecrawl working in a project: signs you in through the browser, saves FIRECRAWL_API_KEY to .env and picks the first SDK or REST path.

firecrawl/firecrawl189k1 repo~1.4kAutomated safety check: NotesISCtoday
2

Routes every web task, from searching to logged-in browsing, through a tiered choice of search, fetch, curl or a real Chrome or Edge session driven over CDP.

eze-is/web-access9.1k4 repos~2.2kAutomated safety check: PassMIT1 mo ago
3

Automates a real browser through short TypeScript scripts that keep page state between runs, for navigating, filling forms, taking screenshots and extracting data.

MemTensor/MemOS12k3 repos~1.7kAutomated safety check: PassApache-2.08 days ago
4

Routes web research and platform lookups across 16 sites, including Twitter, Reddit, YouTube, Bilibili, Xiaohongshu and GitHub, through one command-line tool.

Panniantong/Agent-Reach93k—~1.4kAutomated safety check: PassMIT22 days ago
5

Adds Firecrawl's /scrape endpoint to application code to pull markdown, HTML, links, screenshots or structured data from a single known URL.

firecrawl/firecrawl189k1 repo~944Automated safety check: PassISCtoday
6

Walks through writing an OpenCLI adapter for a new site or a new command on an existing one, from first recon and field decoding to coding and verification.

jackwener/OpenCLI30k1 repo~3.4kAutomated safety check: PassApache-2.012 days ago
7

Adds web search, scraping, structured extraction and browser interaction to application code using Firecrawl's scrape, search and interact endpoints.

firecrawl/firecrawl189k—~2.1kAutomated safety check: NotesISCtoday
8

Core usage guide for the agent-browser CLI: the snapshot-and-ref workflow for navigating, clicking, filling forms, extracting data and running parallel sessions.

vercel-labs/agent-browser44k4 repos~9.5kAutomated safety check: PassApache-2.0yesterday
9

Guides adding Firecrawl's /interact endpoint to product code for pages that need clicks, forms, pagination or logged-in flows beyond plain scraping.

firecrawl/firecrawl189k1 repo~731Automated safety check: PassISCtoday
10
10.Tmux

Remote-control tmux sessions for interactive CLIs by sending keystrokes and scraping pane output.

trpc-group/trpc-agent-go1.8k23 repos~868Automated safety check: PassApache-2.0today
11

Drives a real Chrome window through the opencli CLI to inspect pages, fill forms, click through logged-in flows and extract data from structured command responses.

jackwener/OpenCLI30k2 repos~7.3kAutomated safety check: PassApache-2.012 days ago
12

Drives headless Chrome through Cloudflare Browser Rendering over a CDP WebSocket to take screenshots, navigate and scrape pages, and record multi-page videos.

cloudflare/moltworker9.9k—~742Automated safety check: PassApache-2.05 mo ago
13

Drives a real browser through the omowright library, either the user's own signed-in browser or a separate browser the code launches, for forms, QA, screenshots and scraping.

code-yeongyu/oh-my-openagent70k—~2.2kAutomated safety check: PassUnknowntoday
14

Guidance for adding Firecrawl's /search endpoint to product code and agent workflows when a feature starts from a query rather than a URL.

firecrawl/firecrawl189k1 repo~1.1kAutomated safety check: PassISCtoday
15

Picks the right Skyvern CLI command for a web task, from quick yes/no checks to reusable multi-page workflows, instead of falling back to plain page fetching.

Skyvern-AI/skyvern23k—~2.9kAutomated safety check: PassAGPL-3.0today
16

Drives a web browser from the shell with the agent-browser CLI: open pages, read an element snapshot, click and fill by reference, grab text and screenshots.

nanocoai/nanoclaw31k3 repos~1.6kAutomated safety check: PassMITyesterday
17

Fetches structured Amazon product details such as title, price, ratings and availability for a given ASIN through BrowserAct's lookup API template.

browser-act/skills6.1k2 repos~1.5kAutomated safety check: PassMIT1 mo ago
18

Browser automation with persistent named pages via the dev-browser CLI. Use when users ask to navigate websites, fill forms, take screenshots…

SawyerHood/dev-browser6.7k1 repo~455Automated safety check: PassMIT1 mo ago
19

Controls a real Chrome browser over CDP for clicking, typing, navigation, logged-in sessions and JavaScript-heavy or bot-protected pages.

browser-use/browser-harness18k—~3.6kAutomated safety check: PassMIT9 days ago
20

Detects the type of a knowledge source and uses the Skill Seekers MCP tools to turn docs, repos, PDFs or videos into packaged AI skills.

yusufkaraaslan/Skill_Seekers15k—~760Automated safety check: PassMIT6 days ago
21

Extracts structured Amazon product data for a keyword and marketplace through the BrowserAct API, including titles, prices, ratings, reviews, sales volume and promotions.

browser-act/skills6.1k2 repos~1.5kAutomated safety check: PassMIT1 mo ago
22

How to scan a job source through an Apify actor as a keyed provider.

career-ops-hq/career-ops74k—~246Automated safety check: NotesMITtoday
23

Drives a real browser from the shell with the agent-browser CLI: open pages, snapshot elements by ref, click, fill and extract, for testing web flows.

slopus/happy24k—~1.7kAutomated safety check: PassMITtoday
24

Research skill for ketch — a fast stateless CLI for web search, OSS code search, curated library docs, page scraping, and site crawling; an optional MCP server exists for operators who want it, but…

1broseidon/ketch6961 repo~3.9kAutomated safety check: PassMITtoday
25

Extract structured data from websites and produce an executable Playwright script plus extracted data.

actionbook/actionbook1.6k2 repos~3.4kAutomated safety check: PassApache-2.029 days ago
26

Scrape and extract public data from 27+ social media platforms using the ScrapeCreators REST API.

ScrapeCreators/social-media-research-skills3.3k1 repo~4kAutomated safety check: NotesMIT1 mo ago
27

Kimi Browser Extension(Kimi 浏览器扩展,原 Kimi WebBridge)lets AI control the user's real browser — navigate, click, type, read, screenshot, and interact with any website using the user's actual login…

MoonshotAI/kimi-code7.8k—~3.6kAutomated safety check: PassMIT5 days ago
28

Controls a browser through BrowserWing's local HTTP API: navigation, clicks, typing, data extraction, accessibility snapshots, screenshots and batch operations.

MemTensor/MemOS12k2 repos~3.9kAutomated safety check: PassApache-2.08 days ago
29

Drives your real, logged-in Chrome browser from the command line to read pages, fill forms, fetch authenticated URLs and run ready-made site commands.

epiral/bb-browser6.2k—~2.2kAutomated safety check: PassMIT4 mo ago
30

Turns a plain request such as finding dentists in a city into a local Google Maps crawl run with Docker, then monitors it and helps you work with the results.

gosom/google-maps-scraper6.3k—~1.8kAutomated safety check: WarnMIT13 days ago
31

Scrapes sites, handles JavaScript-heavy pages and extracts structured data with Crawl4AI, through its crwl CLI or Python SDK, including schema-based extraction without an LLM.

smallnest/goclaw5981 repo~2.5kAutomated safety check: PassMIT6 mo ago
32

Scrape BOSS直聘 (job listing site) via Chrome CDP. An agent skill from eatmoreduck/boss-zhipin-scraper.

eatmoreduck/boss-zhipin-scraper1.5k—~2.6kAutomated safety check: PassMIT8 days ago
33

Automates websites with Skyvern's AI browser agent to fill forms, extract data, download files, log in and run multi-step workflows through SDKs, REST, MCP or a CLI.

Skyvern-AI/skyvern23k—~1.9kAutomated safety check: PassAGPL-3.0today
34
34.Ax

Use the ax CLI instead of curl + throwaway parsing scripts whenever you fetch a URL, explore an unknown web page, or extract structured data from HTML.

yusukebe/ax7191 repo~918Automated safety check: PassMIT2 mo ago
35

Scrape any website into clean markdown or structured JSON. An agent skill from Anakin-Inc/anakin.

Anakin-Inc/anakin4.5k—~859Automated safety check: PassAGPL-3.01 mo ago
36

PROACTIVELY use whenever the user needs data outside your training set or requires a live network call — web search, URL scraping, news, social media (any platform), market prices…

chainbase-labs/Agentkey655—~2.3kAutomated safety check: PassMIT1 mo ago
37

Guides building Tavily integrations for web search, URL extraction, site crawling and AI-assisted research in Python or JavaScript agent and RAG projects.

andrewyng/context-hub14k—~1.1kAutomated safety check: PassMIT4 mo ago
38

Fetches Reddit posts, threads and search results as JSON through a browser session, using a DuckDuckGo redirect to get past Reddit's automated-access block.

ykdojo/claude-code-tips10k—~1.4kAutomated safety check: PassUnknown12 days ago
39

Finds new job postings that match your profile through installed portal-search CLIs, dedupes against past runs and your application tracker, and rates each one's fit.

MadsLorentzen/ai-job-search45k—~5.7kAutomated safety check: PassMITyesterday
40

A skill your agent uses when helping a Chinese job seeker, especially students, interns, or early-career users, prepare job applications with an agent without fabricating experience, auto-submitting…

GresonKwan/JobOK411—~745Automated safety check: PassMIT3 mo ago
41

Live SEO data via DataForSEO MCP server. An agent skill from AgriciDaniel/codex-seo.

AgriciDaniel/codex-seo7872 repos~4.6kAutomated safety check: PassMIT25 days ago
42

Browser automation platform with 78 built-in scripts and full CLI.

browserwing/browserwing1.4k—~1.8kAutomated safety check: PassMIT2 mo ago
43

Scrape Google Maps business listings (name, address, phone, website, rating, reviews, lat/lng, hours, emails) via the local gosom google-maps-scraper REST API.

Mahanaicoach/google-maps-scraper-kit1.3k—~2.8kAutomated safety check: PassMIT2 days ago
44

Runs structured data commands against sites such as Twitter, Reddit, GitHub, YouTube and Zhihu through OpenClaw's browser, reusing your existing login state.

epiral/bb-browser6.2k—~1kAutomated safety check: PassMIT4 mo ago
45

Implement web page content extraction capabilities using the z-ai-web-dev-sdk.

jjyaoao/HelloAgents3.2k1 repo~7.1kAutomated safety check: PassMITyesterday
46

Pulls structured Amazon search results (titles, ASINs, prices, ratings, specifications) for a keyword and brand through BrowserAct's Amazon Product API template.

browser-act/skills6.1k2 repos~1.5kAutomated safety check: PassMIT1 mo ago
47

Searches WeChat public account articles by keyword through Sogou WeChat Search and returns titles, summaries, dates, source accounts and links.

zjp1997720/wechat-article-search2681 repo~730Automated safety check: NotesMIT2 mo ago
48

A skill your agent uses when opening, navigating, inspecting, testing, clicking, typing, filling, screenshotting, or verifying web pages and local HTTP targets (localhost, 127.0.0.1, ::1) inside…

zai-org/ZCode7.5k—~4.6kAutomated safety check: PassApache-2.08 days ago

Questions, answered from the data.

What is the best web scraping skill?

Firecrawl Build Onboarding from firecrawl/firecrawl ranks first of the 717 web scraping skills listed here, with the highest score: its repository has 189k GitHub stars, 1 other GitHub owner carry a copy, its SKILL.md loads about 1.4k tokens and it has informational notes only in the automated safety check. Next come Web Access via Browser CDP and Dev Browser Automation.

Which web scraping skills are official?

41 of the 717 web scraping skills are official, published by the vendor's own GitHub organization: Core Guide for agent-browser, Cloudflare Browser Rendering, Apify CLI, Apify Actor Development, Apify Multi-Platform Scraper and 36 more.

How are these skills ranked?

By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.