Topic · Data & Analytics

Best web scraping skills, page 2

Skills #49–96 of 717, ranked by score.

Web scraping skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Web scraping skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
49

Multi-engine web search (SearXNG) + browsing/scraping (Camofox, CloakBrowser).

Johell1NS/browser-search528—~2.5kAutomated safety check: PassMIT1 mo ago
50

Gate a build on browser fingerprint regressions with liarjs - save a baseline scan as JSON, diff later runs against it, and fail the job when the consistency score falls below a floor.

liarjsdev/liarjs-skills5181 repo~986Automated safety check: NotesMIT2 mo ago
51

When the user wants to research, profile, or analyze competitors from their URLs.

Nexus-JPF/note-companion8693 repos~3.5kAutomated safety check: PassMIT4 days ago
52

Access 2,000+ AI models and API tools through one MCP interface for inference, media generation, search, scraping, embeddings, social data, and structured retrieval.

iflytek/skillhub5.2k2 repos~2.1kAutomated safety check: PassApache-2.06 days ago
53

Direct browser control via CDP. An agent skill from davidondrej/skills.

davidondrej/skills4.1k2 repos~3kAutomated safety check: PassMITyesterday
54

Execute Python code in a safe sandboxed environment via [inference.sh](https://inference.sh).

cortega26/chile-hub1132 repos~1.5kAutomated safety check: PassMITyesterday
55

Analyzes ranking charts from Chinese web novel platforms such as Qidian, Fanqie and Jinjiang to spot market trends and promising genres, with scraper scripts per site.

uu201/character-arc5771 repo~2kAutomated safety check: PassMIT4 days ago
56

Official skill for twitterapi.io — query Twitter/X data (tweets, profiles, followers, advanced search, trends, spaces, communities, lists) and perform authenticated actions (post, reply, like…

kaitoInfra/twitterapi-io449—~3.1kAutomated safety check: PassMIT4 mo ago
57

使用 scrapling 进行网页抓取和数据提取。根据目标网站特征自动选择最佳 Fetcher, 生成并执行 Python 脚本完成任务。Use when: (1) 抓取/爬取网页内容或数据(scrape, crawl, fetch page, extract data) (2) 需要绕过 Cloudflare/WAF 等反爬保护 (3) 登录后抓取受保护页面 (4) 解析已有 HTML…

Cedriccmh/claude-code-skill-scrapling440—~1.1kAutomated safety check: PassMIT3 mo ago
58

Collects structured product data from Amazon search results for a keyword and optional brand, using a BrowserAct script, for market and competitor research.

browser-act/skills6.1k2 repos~1.6kAutomated safety check: PassMIT1 mo ago
59

A skill your agent uses when fetching, searching, or analyzing transcripts from Lenny's Podcast, Dwarkesh Podcast, Cheeky Pint, 20VC, or A16z Podcast.

Varnan-Tech/opendirectory6721 repo~1.9kAutomated safety check: NotesMIT1 mo ago
60

Query 57 Indonesian government APIs and data sources — BPJPH halal certification, BPOM food safety, OJK financial legality, BPS statistics, BMKG weather/earthquakes, Bank Indonesia exchange rates…

suryast/indonesia-gov-apis172—~997Automated safety check: PassMITyesterday
61

Query Spot/preemptible VM prices, savings and interruption risk across AWS, GCP and Azure with the spotinfo CLI.

alexei-led/spotinfo164—~1.8kAutomated safety check: PassApache-2.02 days ago
62

Scrape video stats from the Douyin creator center (creator.douyin.com) using AppleScript to automate the user's logged-in Chrome.

TradingAi666/TzFilm-Douyin-Tool405—~7.2kAutomated safety check: PassMIT4 mo ago
63

Scaffold a new PEP (Politically Exposed Persons) crawler — members of a parliament, legislature, senate, chamber of deputies, cabinet, judiciary, or an asset-declaration register — from a source URL…

opensanctions/opensanctions831—~1.9kAutomated safety check: NotesMITtoday
64

Downloads and manages WeChat official account articles from URLs, an account's history or subscriptions, adding engagement metrics and top comments when a valid session exists.

Moore-developers/moore-wechat-article-downloader294—~3.4kAutomated safety check: PassMIT2 mo ago
65

Surveys open Korean government startup and R&D support programs and sorts them by fit with your project, checking eligibility against the original notices.

djfksjd/ir-search392—~3.5kAutomated safety check: NotesMIT2 mo ago
66

Scrape AI papers published by 28 big tech companies and AI labs in a given date window, with institutional attribution (lead vs.

tigerless-labs/paper-radar217—~2.4kAutomated safety check: PassUnknown1 mo ago
67

Analyzes ranking-list data from Chinese web novel platforms to spot repeating genres, title patterns and opening hooks, then writes a market report for authors.

zenstory-ai/oh-story-claudecode7.3k1 repo~1.4kAutomated safety check: PassMIT4 days ago
68

Connects to Oxylabs remote agent browsers over the Chrome DevTools Protocol (CDP) with Playwright or Puppeteer.

oxylabs/agent-skills875—~3kAutomated safety check: PassMIT7 days ago
69

Install, authenticate, configure, operate, and troubleshoot the external MediaCrawler client shared by Douyin, Kuaishou, Bilibili, Weibo, Tieba, and Zhihu collectors.

tsingyuai/growth-lab2k—~731Automated safety check: PassApache-2.09 days ago
70
70.Neo

Browse websites, read web pages, interact with web apps, call website APIs, and automate web tasks.

4ier/neo756—~1.8kAutomated safety check: PassNo licence5 mo ago
71

Drive a signed-in Chrome / Brave / Safari session via the interceptor CLI: open/read pages, click, type, inspect DOM/text/network, automate rich browser editors and scene graphs, capture…

Hacker-Valley-Media/Interceptor514—~4.8kAutomated safety check: PassUnknown5 days ago
72

Generate beautiful, on-brand HTML — presentation decks or full responsive websites — in any brand's design language, combining brand DNA extracted via Firecrawl with codified, research-backed design…

ItsssssJack/power-design713—~2.8kAutomated safety check: PassMIT1 mo ago
73

Conduct deep OSINT research on individuals. An agent skill from smixs/osint-skill.

smixs/osint-skill140—~5.5kAutomated safety check: PassMIT7 mo ago
74

Enter a friendly OpenSEO coach mode that explains workflows, recommends next steps, and helps users use agents, web search, scraping, and MCP data effectively.

petera2c/simple-table2293 repos~1.2kAutomated safety check: PassMIT3 days ago
75

Pulls structured Amazon product reviews for an ASIN through BrowserAct's Amazon Reviews API, with no Amazon login, using a bundled Python script.

browser-act/skills6.1k2 repos~1.4kAutomated safety check: PassMIT1 mo ago
76

Extracts data from web pages using browser automation and CSS/JavaScript selectors.

platonai/Browser41.2k—~1.1kAutomated safety check: PassApache-2.0today
77

Your For You page for content creators. An agent skill from bradautomates/content-ideas.

bradautomates/content-ideas130—~5.4kAutomated safety check: NotesMIT4 mo ago
78

Anti-detection browser automation for AI agents. An agent skill from redf0x1/camofox-browser.

redf0x1/camofox-browser410—~4.6kAutomated safety check: PassMIT15 days ago
79

Compares two or more products or companies on pricing, features and positioning by scraping their sites, and returns a normalized JSON matrix.

firecrawl/web-agent1.2k—~1.1kAutomated safety check: PassMIT5 mo ago
80

Searches and fetches pages through a real logged-in Chromium session when ordinary API or HTML retrieval cannot get past login walls, scripts or anti-bot checks.

taxueseek/argo184—~2.1kAutomated safety check: PassMITyesterday
81

Activate when the user needs to interact with any website — browser automation, web scraping, screenshots, form filling, UI testing, monitoring, or building AI agents.

actionbook/actionbook1.6k—~1.5kAutomated safety check: PassApache-2.029 days ago
82

browser-based page capture and text extraction for public-opinion research.

123321kk/opinion-agent-ultimate107—~631Automated safety check: PassNo licence6 mo ago
83

Generate, revise, translate, and manage App Store / Google Play marketing screenshots.

hypersocialinc/shots240—~2kAutomated safety check: PassNo licence5 mo ago
84

Generate working code that routes HTTP requests through Bright Data proxy networks (Datacenter, ISP, Residential, Mobile) and help users decide which network and IP pool type to use (shared pool…

brightdata/skills264—~5.1kAutomated safety check: PassMITyesterday
85

Browser automation, debugging, and performance analysis using Puppeteer CLI scripts.

einverne/dotfiles1211 repo~1.6kAutomated safety check: NotesApache-2.028 days ago
86

DEFAULT search skill for OpenClaw. An agent skill from skernelx/MySearch-Proxy.

skernelx/MySearch-Proxy159—~1.4kAutomated safety check: NotesMIT6 mo ago
87
87.Cmux

Manage cloud development sandboxes with cmux. An agent skill from manaflow-ai/manaflow.

manaflow-ai/manaflow1.1k—~2kAutomated safety check: PassMITyesterday
88

Oxylabs proxy networks: Residential, Mobile, shared Datacenter/ISP, and Dedicated Datacenter/ISP proxies with geo-targeting, IP rotation, session persistence, and port-based sticky IPs.

oxylabs/agent-skills875—~1.7kAutomated safety check: PassMIT7 days ago
89

Use this skill as the default XCrawl entry point for direct XCrawl requests, including single-URL fetch, format selection, sync or async execution, and JSON extraction with prompt or jsonschema.

xcrawl-api/xcrawl-skills449—~2.4kAutomated safety check: PassNo licence6 mo ago
90

用 DrissionPage 爬取 BOSS 直聘职位,解析简历,用规则打分 + LLM 语义分析做岗位匹配,生成 HTML 可视化报告,自动投递匹配岗位,并启动内嵌的 Markdown 简历编辑器(ShowCV)。当用户想搜索 BOSS 直聘职位、上传简历做岗位匹配、生成岗位匹配报告、自动投递匹配岗位、针对特定 JD 优化简历,或打开简历编辑器写/改简历("打开简历编辑器"、"启动…

zhansan379/boss-crawler-skill100—~2.6kAutomated safety check: PassNo licence1 mo ago
91

Fetch a URL directly with scrapling's bot-detection bypass. An agent skill from cyberchitta/scrapling-fetch-mcp.

cyberchitta/scrapling-fetch-mcp120—~637Automated safety check: NotesApache-2.028 days ago
92

Native Rust browser automation CLI for AI agents. An agent skill from gsd-build/gsd-browser.

gsd-build/gsd-browser267—~1.3kAutomated safety check: PassApache-2.05 mo ago
93

智能读取任意URL内容,支持微信公众号、小红书、今日头条、抖音、淘宝、天猫、京东、百度等中国主流平台,自动识别平台类型并提取核心内容。自动保存内容为Markdown,下载图片到本地。

yhslgg-arch/url-reader188—~1.3kAutomated safety check: PassNo licence8 mo ago
94

Audit websites for Google AdSense application readiness and ad-serving compliance.

yantoumu/adsense-site-auditor-skill236—~1kAutomated safety check: PassNo licence3 mo ago
95

Fetches a web page through a 9Router server's web fetch endpoint and returns it as markdown, plain text or HTML, using one of several extraction providers.

decolua/9router30k—~941Automated safety check: PassMIT6 days ago
96

Iteratively improves cookie-popup button regex patterns in button-patterns.js against labelled-button-texts.csv.

duckduckgo/tracker-radar-collector169—~1.6kAutomated safety check: PassUnknownyesterday