Topic · Data & Analytics
Best web scraping skills, page 11
Web scraping skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 481 | Find paid open-source work, OSS bounties, open source grants, or ways to get paid contributing to open source. | tinyfish-io/ | 2.2k | — | ~4.6k | Automated safety check: Pass | MIT | 7 days ago |
| 482 | 482.Tenders Finder Find open Singapore government tenders for any sector in real time using parallel TinyFish agents scraping multiple government tender portals simultaneously. | tinyfish-io/ | 2.2k | — | ~2.5k | Automated safety check: Pass | MIT | 7 days ago |
| 483 | 483.Use Tinyfish Use TinyFish for web search, fetching URLs, reading pages, current information, source-backed answers, research, docs, pricing/product pages, extraction, scraping, and browser automation. | tinyfish-io/ | 2.2k | — | ~2.1k | Automated safety check: Pass | MIT | 7 days ago |
| 484 | Run the AI blogger crawler daily update check for Bilibili, Douyin, Xiaohongshu, and YouTube with Chrome producers only where required, the merged Xiaohongshu profile-video skill, and the yt-dlp… | dragon-hh/ | 106 | — | ~1.3k | Automated safety check: Pass | No licence | 1 mo ago |
| 485 | 485.Worldcup CLI FIFA World Cup 2026 data via the cli-web-worldcup command — fixtures (nations games), the 48 qualified nations, national-team squads, group standings, and bookmaker odds. | ItamarZand88/ | 231 | — | ~731 | Automated safety check: Pass | MIT | 7 days ago |
| 486 | 486.Web Scraping Authorized web scraping with fallback cascades and access-failure handling. | jamditis/ | 416 | — | ~7.2k | Automated safety check: Pass | MIT | 4 days ago |
| 487 | 487.Roguelike Build a roguelike: turn-based grid movement, procedural dungeons, permadeath, field-of-view, and loot tables. | gamedev-skills/ | 1.4k | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | 12 days ago |
| 488 | Expert legal research agent for finding and scraping expungement data state by state. | curiositech/ | 243 | 1 repo | ~2.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 489 | Firecrawl v2.5 API for web scraping/crawling to LLM-ready markdown. | secondsky/ | 227 | 1 repo | ~1.9k | Automated safety check: Notes | MIT | 10 days ago |
| 490 | Scrape Amazon product data using Pangolin APIs. An agent skill from LeoYeAI/openclaw-master-skills. | LeoYeAI/ | 2.2k | — | ~4.8k | Automated safety check: Pass | MIT | 2 mo ago |
| 491 | Automate repeatable workflows with WhatsApp/Telegram notifications, Excel/CSV processing, browser automation, and flexible scheduling. | LeoYeAI/ | 2.2k | — | ~4.7k | Automated safety check: Pass | MIT | 2 mo ago |
| 492 | Scrape Google Maps for local businesses by category and location, output CSV ready for cold email enrichment. | growthenginenowoslawski/ | 742 | — | ~6.2k | Automated safety check: Notes | MIT | 3 days ago |
| 493 | 493.Scraper Builder Guide AI agents to generate complete PageObject pattern web scraper projects using Playwright and TypeScript with Docker deployment. | jwynia/ | 166 | — | ~4k | Automated safety check: Pass | MIT | 7 mo ago |
| 494 | 494.Firecrawl Map Discover and list all URLs on a website, with optional search filtering. | aiskillstore/ | 430 | 2 repos | ~533 | Automated safety check: Pass | No licence | yesterday |
| 495 | 495.Firecrawl Scrape Extract clean markdown from any URL, including JavaScript-rendered SPAs. | aiskillstore/ | 430 | 2 repos | ~907 | Automated safety check: Pass | No licence | yesterday |
| 496 | This skill should be used when the user asks to "build web automation scripts", "check browser automation for detection", "generate web scraping code", "create form filling automation", or "build… | borghei/ | 881 | — | ~897 | Automated safety check: Pass | MIT | yesterday |
| 497 | Search the Yandex Images vertical and get full-size image URLs as structured JSON with the Apify Yandex Search Scraper Actor (johnvc/Scrape-Yandex). | apify/ | 264 | — | ~1.9k | Automated safety check: Pass | MIT | 16 days ago |
| 498 | 498.SEO Images Image SEO audit for a URL or domain. An agent skill from seranking/seo-skills. | seranking/ | 160 | — | ~6k | Automated safety check: Pass | MIT | 3 mo ago |
| 499 | 499.SEO Sitemap Pull a domain's XML sitemap (and sitemap-of-sitemaps), then compare against the most recent SE Ranking website audit. | seranking/ | 160 | — | ~2.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 500 | AI-powered web scraping - extract data using natural language prompts | gooseworks-ai/ | 1.2k | 1 repo | ~3.1k | Automated safety check: Pass | MIT | today |
| 501 | Browser automation - control browser sessions, scrape pages, and run AI agents | gooseworks-ai/ | 1.2k | 1 repo | ~2.5k | Automated safety check: Pass | MIT | today |
| 502 | Extract structured data from web pages using AI. An agent skill from gooseworks-ai/goose-skills. | gooseworks-ai/ | 1.2k | 1 repo | ~1.9k | Automated safety check: Pass | MIT | today |
| 503 | Generate a personal voice guide for X (Twitter) and/or LinkedIn by scanning a user's past posts and iteratively refining with sample-and-feedback loops. | gooseworks-ai/ | 1.2k | 1 repo | ~2.1k | Automated safety check: Pass | MIT | today |
| 504 | Get Instagram profiles, posts, and reels. An agent skill from gooseworks-ai/goose-skills. | gooseworks-ai/ | 1.2k | 1 repo | ~584 | Automated safety check: Pass | MIT | today |
| 505 | Research LinkedIn profiles and write personalized messages for any LinkedIn message type — connection requests, InMails, DMs, message requests, post comments, and comment replies. | gooseworks-ai/ | 1.2k | 1 repo | ~3.4k | Automated safety check: Notes | MIT | today |
| 506 | 506.Linkedin Scraper Get LinkedIn profiles, company pages, and posts. An agent skill from gooseworks-ai/goose-skills. | gooseworks-ai/ | 1.2k | 1 repo | ~640 | Automated safety check: Pass | MIT | today |
| 507 | Comprehensive SEO footprint analysis. An agent skill from gooseworks-ai/goose-skills. | gooseworks-ai/ | 1.2k | 1 repo | ~2.8k | Automated safety check: Pass | MIT | today |
| 508 | Web scraping with structured data extraction - define your output schema | gooseworks-ai/ | 1.2k | 1 repo | ~1.5k | Automated safety check: Pass | MIT | today |
| 509 | Monitor Twitter/X, Reddit, LinkedIn, and Hacker News for trending narratives, viral posts, and hot-button topics in your space. | gooseworks-ai/ | 1.2k | 1 repo | ~2.5k | Automated safety check: Pass | MIT | today |
| 510 | Web scraping, crawling, and AI-powered answer extraction at scale | gooseworks-ai/ | 1.2k | 1 repo | ~3.1k | Automated safety check: Pass | MIT | today |
| 511 | 511.Geo Crawlers AI crawler access analysis. | sickn33/ | 47k | 1 repo | ~4.8k | Automated safety check: Notes | MIT | yesterday |
| 512 | Monitor competitor content across blogs, LinkedIn, and Twitter/X on a recurring basis. | majiayu000/ | 666 | 2 repos | ~1.4k | Automated safety check: Pass | MIT | yesterday |
| 513 | Search Hacker News stories and comments using the free Algolia API. | majiayu000/ | 666 | 2 repos | ~538 | Automated safety check: Pass | MIT | yesterday |
| 514 | Scrape and search Reddit posts using Apify. An agent skill from majiayu000/claude-skill-registry. | majiayu000/ | 666 | 2 repos | ~1.2k | Automated safety check: Pass | MIT | yesterday |
| 515 | Search and scrape Twitter/X posts using Apify. An agent skill from majiayu000/claude-skill-registry. | majiayu000/ | 666 | 2 repos | ~805 | Automated safety check: Pass | MIT | yesterday |
| 516 | Scrape blog posts via RSS feeds (free, no API key) with Apify fallback for JS-heavy sites. | majiayu000/ | 666 | 2 repos | ~578 | Automated safety check: Pass | MIT | yesterday |
| 517 | Extract speaker names, titles, companies, and bios from conference websites. | majiayu000/ | 666 | 2 repos | ~846 | Automated safety check: Pass | MIT | yesterday |
| 518 | Scrape competitor ads from Google Ads by domain. An agent skill from majiayu000/claude-skill-registry. | majiayu000/ | 666 | 2 repos | ~977 | Automated safety check: Pass | MIT | yesterday |
| 519 | Extract commenters from LinkedIn posts via Apify. An agent skill from majiayu000/claude-skill-registry. | majiayu000/ | 666 | 2 repos | ~722 | Automated safety check: Pass | MIT | yesterday |
| 520 | Search LinkedIn posts by keywords, sorted by engagement or date. | majiayu000/ | 666 | 2 repos | ~1.3k | Automated safety check: Notes | MIT | yesterday |
| 521 | Scrape recent posts from LinkedIn profiles using Apify. An agent skill from majiayu000/claude-skill-registry. | majiayu000/ | 666 | 2 repos | ~495 | Automated safety check: Pass | MIT | yesterday |
| 522 | 522.Meta Ad Scraper Scrape competitor ads from Meta's Ad Library (Facebook, Instagram, Messenger, Threads, WhatsApp). | majiayu000/ | 666 | 2 repos | ~1.2k | Automated safety check: Pass | MIT | yesterday |
| 523 | 523.Review Scraper Scrape product reviews from G2, Capterra, and Trustpilot using Apify. | majiayu000/ | 666 | 2 repos | ~573 | Automated safety check: Pass | MIT | yesterday |
| 524 | Scrape product reviews from G2, Capterra, and Trustpilot using Apify. | majiayu000/ | 666 | 2 repos | ~882 | Automated safety check: Pass | MIT | yesterday |
| 525 | Find TikTok influencers using Apify's Influencer Discovery Agent. | majiayu000/ | 666 | 2 repos | ~1.1k | Automated safety check: Pass | MIT | yesterday |
| 526 | A toolkit for web data extraction and search intent analysis. | kennyzir/ | 322 | — | ~432 | Automated safety check: Pass | MIT | 9 days ago |
| 527 | 从 Roblox 游戏的公开 Trello 看板采集卡片,并提取物品与兑换码结构化数据。适用于无需登录的公开看板资料整理和 JSON 数据准备。 | kennyzir/ | 322 | — | ~492 | Automated safety check: Pass | MIT | 9 days ago |
| 528 | [omh] Expensive or unconfigured web search: diagnose scraper and auxiliary extract-model configuration, guide account setup, and apply each change as its own diff approval. | rlaope/ | 3.2k | — | ~1.9k | Automated safety check: Notes | MIT | today |