Topic · Data & Analytics
Best web scraping skills, page 15
Web scraping skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 673 | 673.Brightdata Progressive four-tier URL content scraping with automatic fallback strategy. | Microck/ | 401 | — | ~1.4k | Automated safety check: Pass | Unknown | 1 mo ago |
| 674 | 674.Algo SEO Crawl Implement a web crawler pipeline covering URL discovery, fetching, parsing, and storage. | asgard-ai-platform/ | 241 | — | ~993 | Automated safety check: Pass | MIT | 4 mo ago |
| 675 | 675.Social Listening A skill your agent uses when you need to turn a supplied set of comments, DMs, reviews, community posts, or call notes into an evidence-backed picture of what people ask, resist, and repeat. | Ootto-AI/ | 101 | — | ~882 | Automated safety check: Pass | MIT | 1 mo ago |
| 676 | 676.Browser Tools Security wrapper over the upstream agent-browser skill, adding URL blocklisting, rate limiting, robots.txt enforcement, and scraping guardrails. | yonatangross/ | 288 | — | ~6.1k | Automated safety check: Pass | MIT | yesterday |
| 677 | 677.Pp Context Dev Printing Press CLI for Context.dev. An agent skill from mvanhorn/printing-press-library. | mvanhorn/ | 2.1k | — | ~4.3k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 678 | 678.Pp Firecrawl Printing Press CLI for Firecrawl. An agent skill from mvanhorn/printing-press-library. | mvanhorn/ | 2.1k | — | ~2.2k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 679 | Enterprise-grade browser automation using WebDriver protocol. | aiskillstore/ | 430 | — | ~1.5k | Automated safety check: Notes | No licence | today |
| 680 | 680.Firecrawl Firecrawl handles all web operations with superior accuracy, speed, and LLM-optimized output. | aiskillstore/ | 430 | 1 repo | ~2.6k | Automated safety check: Pass | No licence | today |
| 681 | Find a direct path to structured data through ready-made workflows, data APIs, and indexes. | firecrawl/ | 115 | — | ~1.6k | Automated safety check: Pass | ISC | yesterday |
| 682 | 682.Firecrawl 2 Web scraping and crawling with Firecrawl API. An agent skill from sundial-org/awesome-openclaw-skills. | sundial-org/ | 663 | — | ~698 | Automated safety check: Pass | No licence | 7 mo ago |
| 683 | 683.Serper Google search via Serper API with full page content extraction. | sundial-org/ | 663 | — | ~1.6k | Automated safety check: Notes | No licence | 7 mo ago |
| 684 | Configure static outbound IP addresses for Google Cloud Run jobs/services. | divinevideo/ | 265 | — | ~1.5k | Automated safety check: Pass | MPL-2.0 | today |
| 685 | Fix rsky-pds createRecord/putRecord hanging indefinitely with no error logs. | divinevideo/ | 265 | — | ~1.1k | Automated safety check: Pass | MPL-2.0 | today |
| 686 | A skill your agent uses when handling research ethics, transparency, and data sharing for a New Media & Society (NM&S) manuscript — consent and anonymization, the ethics of scraping and platform… | brycewang-stanford/ | 1.2k | — | ~1.6k | Automated safety check: Pass | MIT | 10 days ago |
| 687 | 687.Firecrawl 专业网页抓取和数据提取。使用 Firecrawl API 抓取网页、提取结构化数据、批量爬取网站。当用户需要抓取复杂网页、提取结构化数据、批量爬取时使用此技能。 | aAAaqwq/ | 105 | 1 repo | ~260 | Automated safety check: Notes | MIT | 10 days ago |
| 688 | When the user wants to choose or optimize rendering strategy for SEO. | kostja94/ | 1k | — | ~1.6k | Automated safety check: Pass | MIT | yesterday |
| 689 | 689.Coffee Chat Generate a personalized coffee chat playbook for networking conversations. | LeoYeAI/ | 2.2k | — | ~6.6k | Automated safety check: Pass | MIT | 2 mo ago |
| 690 | 690.Firecrawl Firecrawl API integration with managed authentication. An agent skill from LeoYeAI/openclaw-master-skills. | LeoYeAI/ | 2.2k | — | ~6.2k | Automated safety check: Pass | MIT | 2 mo ago |
| 691 | 691.Firecrawl MCP Auto-generated skill for firecrawl-mcp tools via OneKey Gateway. | LeoYeAI/ | 2.2k | — | ~7.4k | Automated safety check: Pass | MIT | 2 mo ago |
| 692 | 692.Init Kb Initialize or update a knowledge base for a project, business, or client. | LeoYeAI/ | 2.2k | — | ~7k | Automated safety check: Pass | MIT | 2 mo ago |
| 693 | 693.Extract Crawl an existing website (capped, multi-page) and seed stardust/current/ with PRODUCT.md, DESIGN.md, DESIGN.json, a per-page inventory, and the consolidated brand surface — the captured design… | adobe/ | 195 | — | ~11k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 694 | Use Xquik to fetch X (Twitter) data or act through a connected account: search, profiles, followers, replies, threads, timelines, media downloads, bulk exports, trends, monitors, signed webhooks… | hashgraph-online/ | 1.2k | — | ~2.5k | Automated safety check: Pass | MIT | yesterday |
| 695 | Extract clean, de-noised job-posting data from LinkedIn, Indeed, Glassdoor, and 20+ boards in one Apify run — deduplicated across boards, with likely ghost jobs and reposts flagged (heuristic, not… | apify/ | 262 | — | ~5.5k | Automated safety check: Pass | Apache-2.0 | 15 days ago |
| 696 | Answer classic system design problems as constraint-to-solution sketches and coach interview practice: URL shortener, rate limiter, news feed, chat, notification, autocomplete, crawler, unique id. | HoangNguyen0403/ | 570 | — | ~1.1k | Automated safety check: Pass | MIT | yesterday |
| 697 | Browser automation via Playwright for web testing, screenshots, form filling, scraping, and verification. | ArabelaTso/ | 253 | — | ~930 | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 698 | Clean and extract the main body content from a webpage URL. An agent skill from OpenMinis/MinisSkills. | OpenMinis/ | 440 | — | ~564 | Automated safety check: Pass | MIT | 2 days ago |
| 699 | Assesses lawful basis for AI training data processing per EDPB April 2025 report on LLMs and general-purpose AI. | mukul975/ | 295 | — | ~3.7k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 700 | 700.Steel Browser Use this skill by default for browser or web tasks that can run in the cloud: site navigation, scraping, structured extraction, screenshots/PDFs, form flows, and anti-bot-sensitive automation. | aiskillstore/ | 430 | — | ~2k | Automated safety check: Pass | No licence | today |
| 701 | 701.Tiktok API A TikTok API alternative on fetcher.sh — pay-per-call in USDC via x402, or prepaid credits with a Bearer key, no login and no app review. | aiskillstore/ | 430 | — | ~2k | Automated safety check: Pass | No licence | today |
| 702 | 702.Twitter API A Twitter API alternative and X API alternative on fetcher.sh — pay-per-call in USDC via x402, or prepaid credits with a Bearer key, no OAuth and no developer application. | aiskillstore/ | 430 | — | ~2.2k | Automated safety check: Pass | No licence | today |
| 703 | A skill your agent uses when the user needs X (Twitter) data or confirmation-gated X actions through Xquik: tweet search, user lookup, follower extraction, media download, monitoring, webhooks, MCP… | aiskillstore/ | 430 | — | ~2.6k | Automated safety check: Pass | MIT | today |
| 704 | Playwright-based browser automation for scraping JavaScript-rendered scientific databases | lamm-mit/ | 244 | — | ~414 | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 705 | Web scraping of JavaScript-rendered scientific websites using Firecrawl API | lamm-mit/ | 244 | — | ~425 | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 706 | Ingest and normalize web pages for critical-minerals intelligence, with optional Firecrawl fetching, deduplication manifest, and JSONL export | lamm-mit/ | 244 | — | ~512 | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 707 | 707.Abm Outbound Multi-channel ABM automation that turns LinkedIn URLs into coordinated outbound campaigns. | sundial-org/ | 663 | — | ~1.8k | Automated safety check: Pass | No licence | 7 mo ago |
| 708 | 708.Scrape Webpage Scrape webpage content, extract metadata, download images, and prepare for import/migration to AEM Edge Delivery Services. | NeverSight/ | 216 | 1 repo | ~1.2k | Automated safety check: Pass | No licence | yesterday |
| 709 | 709.Image Scraper Scrape and download all images from a given URL. An agent skill from aAAaqwq/AGI-Super-Team. | aAAaqwq/ | 105 | — | ~914 | Automated safety check: Pass | MIT | 10 days ago |
| 710 | 自动化爬取网站数据和 API 接口。当用户需要抓取网页内容、调用 API、解析数据或创建爬虫脚本时使用此技能. An agent skill from aAAaqwq/AGI-Super-Team. | aAAaqwq/ | 105 | 1 repo | ~795 | Automated safety check: Notes | MIT | 10 days ago |
| 711 | 711.Web Search 网络搜索与网页内容获取。当用户需要搜索互联网信息、获取网页内容、查找实时数据、进行 websearch 时使用此技能。支持多种搜索工具:WebFetch、Firecrawl skill、Tavily skill。 | aAAaqwq/ | 105 | 1 repo | ~703 | Automated safety check: Notes | MIT | 10 days ago |
| 712 | X API & Twitter automation skill. An agent skill from LeoYeAI/openclaw-master-skills. | LeoYeAI/ | 2.2k | — | ~7.1k | Automated safety check: Pass | MIT | 2 mo ago |
| 713 | 713.Web Scraper Full pipeline for building web scraping systems with agent team collaboration. | revfactory/ | 1.3k | — | ~1.7k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 714 | This skill should be used when the user asks to "analyze a competitor", "compare pricing", "competitive landscape", "market research", "what do customers think", "review intelligence", "hiring… | apify/ | 262 | — | ~2.5k | Automated safety check: Notes | Apache-2.0 | 15 days ago |
| 715 | Use before adding or changing Python code that calls a third-party HTTP API from PostHog (a vendor REST call, a vendor SDK client, a scraping or enrichment service), and before adding or changing a… | PostHog/ | 721 | — | ~2.7k | Automated safety check: Pass | MIT | today |
| 716 | 716.Content Monitor Monitor web pages for changes and extract updated content with scheduling support. | CoWork-OS/ | 473 | — | ~456 | Automated safety check: Pass | MIT | yesterday |
| 717 | 717.Web Scraper Scrape web pages and extract structured data using Scrapling. | CoWork-OS/ | 473 | — | ~452 | Automated safety check: Pass | MIT | yesterday |