Topic · Data & Analytics
Best web scraping skills for Claude Code, Codex and other agents.
- skills
- 717
- official
- 41
Web scraping skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Gets Firecrawl working in a project: signs you in through the browser, saves FIRECRAWL_API_KEY to .env and picks the first SDK or REST path. | firecrawl/ | 189k | 1 repo | ~1.4k | Automated safety check: Notes | ISC | today |
| 2 | Routes every web task, from searching to logged-in browsing, through a tiered choice of search, fetch, curl or a real Chrome or Edge session driven over CDP. | eze-is/ | 9.1k | 4 repos | ~2.2k | Automated safety check: Pass | MIT | 1 mo ago |
| 3 | Automates a real browser through short TypeScript scripts that keep page state between runs, for navigating, filling forms, taking screenshots and extracting data. | MemTensor/ | 12k | 3 repos | ~1.7k | Automated safety check: Pass | Apache-2.0 | 8 days ago |
| 4 | Routes web research and platform lookups across 16 sites, including Twitter, Reddit, YouTube, Bilibili, Xiaohongshu and GitHub, through one command-line tool. | Panniantong/ | 93k | — | ~1.4k | Automated safety check: Pass | MIT | 22 days ago |
| 5 | Adds Firecrawl's /scrape endpoint to application code to pull markdown, HTML, links, screenshots or structured data from a single known URL. | firecrawl/ | 189k | 1 repo | ~944 | Automated safety check: Pass | ISC | today |
| 6 | Walks through writing an OpenCLI adapter for a new site or a new command on an existing one, from first recon and field decoding to coding and verification. | jackwener/ | 30k | 1 repo | ~3.4k | Automated safety check: Pass | Apache-2.0 | 12 days ago |
| 7 | Adds web search, scraping, structured extraction and browser interaction to application code using Firecrawl's scrape, search and interact endpoints. | firecrawl/ | 189k | — | ~2.1k | Automated safety check: Notes | ISC | today |
| 8 | Core usage guide for the agent-browser CLI: the snapshot-and-ref workflow for navigating, clicking, filling forms, extracting data and running parallel sessions. | vercel-labs/ | 44k | 4 repos | ~9.5k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 9 | Guides adding Firecrawl's /interact endpoint to product code for pages that need clicks, forms, pagination or logged-in flows beyond plain scraping. | firecrawl/ | 189k | 1 repo | ~731 | Automated safety check: Pass | ISC | today |
| 10 | 10.Tmux Remote-control tmux sessions for interactive CLIs by sending keystrokes and scraping pane output. | trpc-group/ | 1.8k | 23 repos | ~868 | Automated safety check: Pass | Apache-2.0 | today |
| 11 | Drives a real Chrome window through the opencli CLI to inspect pages, fill forms, click through logged-in flows and extract data from structured command responses. | jackwener/ | 30k | 2 repos | ~7.3k | Automated safety check: Pass | Apache-2.0 | 12 days ago |
| 12 | Drives headless Chrome through Cloudflare Browser Rendering over a CDP WebSocket to take screenshots, navigate and scrape pages, and record multi-page videos. | cloudflare/ | 9.9k | — | ~742 | Automated safety check: Pass | Apache-2.0 | 5 mo ago |
| 13 | Drives a real browser through the omowright library, either the user's own signed-in browser or a separate browser the code launches, for forms, QA, screenshots and scraping. | code-yeongyu/ | 70k | — | ~2.2k | Automated safety check: Pass | Unknown | today |
| 14 | Guidance for adding Firecrawl's /search endpoint to product code and agent workflows when a feature starts from a query rather than a URL. | firecrawl/ | 189k | 1 repo | ~1.1k | Automated safety check: Pass | ISC | today |
| 15 | Picks the right Skyvern CLI command for a web task, from quick yes/no checks to reusable multi-page workflows, instead of falling back to plain page fetching. | Skyvern-AI/ | 23k | — | ~2.9k | Automated safety check: Pass | AGPL-3.0 | today |
| 16 | Drives a web browser from the shell with the agent-browser CLI: open pages, read an element snapshot, click and fill by reference, grab text and screenshots. | nanocoai/ | 31k | 3 repos | ~1.6k | Automated safety check: Pass | MIT | yesterday |
| 17 | Fetches structured Amazon product details such as title, price, ratings and availability for a given ASIN through BrowserAct's lookup API template. | browser-act/ | 6.1k | 2 repos | ~1.5k | Automated safety check: Pass | MIT | 1 mo ago |
| 18 | Browser automation with persistent named pages via the dev-browser CLI. Use when users ask to navigate websites, fill forms, take screenshots… | SawyerHood/ | 6.7k | 1 repo | ~455 | Automated safety check: Pass | MIT | 1 mo ago |
| 19 | Controls a real Chrome browser over CDP for clicking, typing, navigation, logged-in sessions and JavaScript-heavy or bot-protected pages. | browser-use/ | 18k | — | ~3.6k | Automated safety check: Pass | MIT | 9 days ago |
| 20 | Detects the type of a knowledge source and uses the Skill Seekers MCP tools to turn docs, repos, PDFs or videos into packaged AI skills. | yusufkaraaslan/ | 15k | — | ~760 | Automated safety check: Pass | MIT | 6 days ago |
| 21 | Extracts structured Amazon product data for a keyword and marketplace through the BrowserAct API, including titles, prices, ratings, reviews, sales volume and promotions. | browser-act/ | 6.1k | 2 repos | ~1.5k | Automated safety check: Pass | MIT | 1 mo ago |
| 22 | How to scan a job source through an Apify actor as a keyed provider. | career-ops-hq/ | 74k | — | ~246 | Automated safety check: Notes | MIT | today |
| 23 | Drives a real browser from the shell with the agent-browser CLI: open pages, snapshot elements by ref, click, fill and extract, for testing web flows. | slopus/ | 24k | — | ~1.7k | Automated safety check: Pass | MIT | today |
| 24 | 24.Ketch Research skill for ketch — a fast stateless CLI for web search, OSS code search, curated library docs, page scraping, and site crawling; an optional MCP server exists for operators who want it, but… | 1broseidon/ | 696 | 1 repo | ~3.9k | Automated safety check: Pass | MIT | today |
| 25 | 25.Extract Extract structured data from websites and produce an executable Playwright script plus extracted data. | actionbook/ | 1.6k | 2 repos | ~3.4k | Automated safety check: Pass | Apache-2.0 | 29 days ago |
| 26 | Scrape and extract public data from 27+ social media platforms using the ScrapeCreators REST API. | ScrapeCreators/ | 3.3k | 1 repo | ~4k | Automated safety check: Notes | MIT | 1 mo ago |
| 27 | Kimi Browser Extension(Kimi 浏览器扩展,原 Kimi WebBridge)lets AI control the user's real browser — navigate, click, type, read, screenshot, and interact with any website using the user's actual login… | MoonshotAI/ | 7.8k | — | ~3.6k | Automated safety check: Pass | MIT | 5 days ago |
| 28 | Controls a browser through BrowserWing's local HTTP API: navigation, clicks, typing, data extraction, accessibility snapshots, screenshots and batch operations. | MemTensor/ | 12k | 2 repos | ~3.9k | Automated safety check: Pass | Apache-2.0 | 8 days ago |
| 29 | Drives your real, logged-in Chrome browser from the command line to read pages, fill forms, fetch authenticated URLs and run ready-made site commands. | epiral/ | 6.2k | — | ~2.2k | Automated safety check: Pass | MIT | 4 mo ago |
| 30 | Turns a plain request such as finding dentists in a city into a local Google Maps crawl run with Docker, then monitors it and helps you work with the results. | gosom/ | 6.3k | — | ~1.8k | Automated safety check: Warn | MIT | 13 days ago |
| 31 | Scrapes sites, handles JavaScript-heavy pages and extracts structured data with Crawl4AI, through its crwl CLI or Python SDK, including schema-based extraction without an LLM. | smallnest/ | 598 | 1 repo | ~2.5k | Automated safety check: Pass | MIT | 6 mo ago |
| 32 | Scrape BOSS直聘 (job listing site) via Chrome CDP. An agent skill from eatmoreduck/boss-zhipin-scraper. | eatmoreduck/ | 1.5k | — | ~2.6k | Automated safety check: Pass | MIT | 8 days ago |
| 33 | Automates websites with Skyvern's AI browser agent to fill forms, extract data, download files, log in and run multi-step workflows through SDKs, REST, MCP or a CLI. | Skyvern-AI/ | 23k | — | ~1.9k | Automated safety check: Pass | AGPL-3.0 | today |
| 34 | 34.Ax Use the ax CLI instead of curl + throwaway parsing scripts whenever you fetch a URL, explore an unknown web page, or extract structured data from HTML. | yusukebe/ | 719 | 1 repo | ~918 | Automated safety check: Pass | MIT | 2 mo ago |
| 35 | Scrape any website into clean markdown or structured JSON. An agent skill from Anakin-Inc/anakin. | Anakin-Inc/ | 4.5k | — | ~859 | Automated safety check: Pass | AGPL-3.0 | 1 mo ago |
| 36 | 36.Agentkey PROACTIVELY use whenever the user needs data outside your training set or requires a live network call — web search, URL scraping, news, social media (any platform), market prices… | chainbase-labs/ | 655 | — | ~2.3k | Automated safety check: Pass | MIT | 1 mo ago |
| 37 | Guides building Tavily integrations for web search, URL extraction, site crawling and AI-assisted research in Python or JavaScript agent and RAG projects. | andrewyng/ | 14k | — | ~1.1k | Automated safety check: Pass | MIT | 4 mo ago |
| 38 | Fetches Reddit posts, threads and search results as JSON through a browser session, using a DuckDuckGo redirect to get past Reddit's automated-access block. | ykdojo/ | 10k | — | ~1.4k | Automated safety check: Pass | Unknown | 12 days ago |
| 39 | Finds new job postings that match your profile through installed portal-search CLIs, dedupes against past runs and your application tracker, and rates each one's fit. | MadsLorentzen/ | 45k | — | ~5.7k | Automated safety check: Pass | MIT | yesterday |
| 40 | 40.Job Ok A skill your agent uses when helping a Chinese job seeker, especially students, interns, or early-career users, prepare job applications with an agent without fabricating experience, auto-submitting… | GresonKwan/ | 411 | — | ~745 | Automated safety check: Pass | MIT | 3 mo ago |
| 41 | Live SEO data via DataForSEO MCP server. An agent skill from AgriciDaniel/codex-seo. | AgriciDaniel/ | 787 | 2 repos | ~4.6k | Automated safety check: Pass | MIT | 25 days ago |
| 42 | 42.Browserwing Browser automation platform with 78 built-in scripts and full CLI. | browserwing/ | 1.4k | — | ~1.8k | Automated safety check: Pass | MIT | 2 mo ago |
| 43 | Scrape Google Maps business listings (name, address, phone, website, rating, reviews, lat/lng, hours, emails) via the local gosom google-maps-scraper REST API. | Mahanaicoach/ | 1.3k | — | ~2.8k | Automated safety check: Pass | MIT | 2 days ago |
| 44 | Runs structured data commands against sites such as Twitter, Reddit, GitHub, YouTube and Zhihu through OpenClaw's browser, reusing your existing login state. | epiral/ | 6.2k | — | ~1k | Automated safety check: Pass | MIT | 4 mo ago |
| 45 | 45.Web Reader Implement web page content extraction capabilities using the z-ai-web-dev-sdk. | jjyaoao/ | 3.2k | 1 repo | ~7.1k | Automated safety check: Pass | MIT | yesterday |
| 46 | Pulls structured Amazon search results (titles, ASINs, prices, ratings, specifications) for a keyword and brand through BrowserAct's Amazon Product API template. | browser-act/ | 6.1k | 2 repos | ~1.5k | Automated safety check: Pass | MIT | 1 mo ago |
| 47 | Searches WeChat public account articles by keyword through Sogou WeChat Search and returns titles, summaries, dates, source accounts and links. | zjp1997720/ | 268 | 1 repo | ~730 | Automated safety check: Notes | MIT | 2 mo ago |
| 48 | A skill your agent uses when opening, navigating, inspecting, testing, clicking, typing, filling, screenshotting, or verifying web pages and local HTTP targets (localhost, 127.0.0.1, ::1) inside… | zai-org/ | 7.5k | — | ~4.6k | Automated safety check: Pass | Apache-2.0 | 8 days ago |
Questions, answered from the data.
What is the best web scraping skill?
Firecrawl Build Onboarding from firecrawl/firecrawl ranks first of the 717 web scraping skills listed here, with the highest score: its repository has 189k GitHub stars, 1 other GitHub owner carry a copy, its SKILL.md loads about 1.4k tokens and it has informational notes only in the automated safety check. Next come Web Access via Browser CDP and Dev Browser Automation.
Which web scraping skills are official?
41 of the 717 web scraping skills are official, published by the vendor's own GitHub organization: Core Guide for agent-browser, Cloudflare Browser Rendering, Apify CLI, Apify Actor Development, Apify Multi-Platform Scraper and 36 more.
How are these skills ranked?
By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.