Topic · Data & Analytics
Best web scraping skills, page 7
Web scraping skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 289 | Scrapes second-hand item search results from Goofish (闲鱼/xianyu, goofish.com) — China's largest second-hand marketplace. | browser-act/ | 6.1k | — | ~2k | Automated safety check: Pass | MIT | 1 mo ago |
| 290 | Extracts business contact details from Google Maps search results and place detail pages, then visits each business website to collect emails, phone numbers, and social media profiles (Facebook… | browser-act/ | 6.1k | — | ~2.9k | Automated safety check: Pass | MIT | 1 mo ago |
| 291 | Extracts Google Search results page (SERP) data including organic results, paid ads, related searches, People Also Ask questions, AI Overview text, and total result count from google.com. | browser-act/ | 6.1k | — | ~2k | Automated safety check: Pass | MIT | 1 mo ago |
| 292 | Scrape job listings from Indeed.com by keyword, location, and country. | browser-act/ | 6.1k | — | ~2.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 293 | Scrapes Instagram posts by hashtag, returning media items with captions, like/comment counts, media URLs and user info from the hashtag explore feed. | browser-act/ | 6.1k | — | ~1.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 294 | Scrapes Instagram posts tagged at a specific location or place, returning media items with captions, like/comment counts, media URLs and user info. | browser-act/ | 6.1k | — | ~1.9k | Automated safety check: Pass | MIT | 1 mo ago |
| 295 | Fetches comments from an Instagram post including comment text, username, timestamp, like count and reply count. | browser-act/ | 6.1k | — | ~1.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 296 | Fetches Instagram user profile metadata including bio, follower count, following count, post count, verification status and other profile details. | browser-act/ | 6.1k | — | ~1.2k | Automated safety check: Pass | MIT | 1 mo ago |
| 297 | Scrapes posts from an Instagram user's profile feed including captions, media URLs, like/comment counts, timestamps and location tags. | browser-act/ | 6.1k | — | ~1.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 298 | Search LinkedIn job listings and extract full job details. An agent skill from browser-act/skills. | browser-act/ | 6.1k | — | ~2.5k | Automated safety check: Pass | MIT | 1 mo ago |
| 299 | Scrape Product Hunt daily/weekly/monthly/yearly leaderboard launches with full product details, maker profiles, and website contact info. | browser-act/ | 6.1k | — | ~2.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 300 | Search Taobao and Tmall product listings by keyword, returning paginated product cards with title, price, shop, image, sales, and tags. | browser-act/ | 6.1k | — | ~1.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 301 | Fetch full product detail from a Taobao or Tmall product page by itemId, returning title, price, shop info, images, SKU variants, and product attributes. | browser-act/ | 6.1k | — | ~1.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 302 | Fetch customer reviews for a Taobao or Tmall product by itemId, returning reviewer name, date, purchased variant, review text, and photo URLs. | browser-act/ | 6.1k | — | ~1.8k | Automated safety check: Pass | MIT | 1 mo ago |
| 303 | Browse a Taobao or Tmall shop's product catalog by shopId, returning paginated product listings with itemId and title. | browser-act/ | 6.1k | — | ~1.4k | Automated safety check: Pass | MIT | 1 mo ago |
| 304 | Searches Threads posts by keyword or hashtag and returns matching posts with engagement metrics, extracted from SSR-embedded JSON. | browser-act/ | 6.1k | — | ~1.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 305 | Fetches public posts from a Threads user's profile page, extracting post text, engagement metrics, and media info from SSR-embedded JSON. | browser-act/ | 6.1k | — | ~2.3k | Automated safety check: Pass | MIT | 1 mo ago |
| 306 | TikTok hashtag video scraper: input a hashtag name → output paginated video list with full metadata (author profile, engagement stats, music, video meta, hashtag list). | browser-act/ | 6.1k | — | ~2k | Automated safety check: Pass | MIT | 1 mo ago |
| 307 | TikTok single video detail scraper: input a TikTok video URL → output full video metadata (author profile, engagement stats, music, video meta, hashtags, mentions, slideshow images). | browser-act/ | 6.1k | — | ~1.5k | Automated safety check: Pass | MIT | 1 mo ago |
| 308 | Walmart keyword search scraper: input a search keyword and page number, navigate to walmart.com search results, extract paginated product listings with itemId, url, title, brand, image, price… | browser-act/ | 6.1k | — | ~1.8k | Automated safety check: Pass | MIT | 1 mo ago |
| 309 | Deep-crawl any website from start URLs, return per-page LLM-ready text/markdown/HTML plus metadata (title, description, author, language, canonical URL, OG) and in-scope outbound links. | browser-act/ | 6.1k | — | ~4.3k | Automated safety check: Pass | MIT | 1 mo ago |
| 310 | 310.X Tweet Search Scrapes tweets from X (Twitter) by search query, user handle, or direct URL — returns full tweet data including text, author info, engagement metrics, media, and hashtags. | browser-act/ | 6.1k | — | ~2.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 311 | Fetch Xiaohongshu (RedNote / xhs) note detail and comments by note ID, returning title, description, author info, engagement stats, tags, and paginated comment list. | browser-act/ | 6.1k | — | ~2.1k | Automated safety check: Pass | MIT | 1 mo ago |
| 312 | Search Xiaohongshu (RedNote / xhs) notes by keyword and return a paginated list with title, author, engagement stats (likes, collects, comments), cover image URL, and xsecToken for detail lookup. | browser-act/ | 6.1k | — | ~2.1k | Automated safety check: Pass | MIT | 1 mo ago |
| 313 | Fetch Xiaohongshu (RedNote / xhs) user profile information and their published notes list by user ID, returning nickname, bio, follower/following counts, engagement totals, tags, and paginated notes… | browser-act/ | 6.1k | — | ~2k | Automated safety check: Pass | MIT | 1 mo ago |
| 314 | Scrapes Amazon product data from ASINs using browseract.com automation API and performs surgical competitive analysis. | browser-act/ | 6.1k | 1 repo | ~1k | Automated safety check: Notes | MIT | 1 mo ago |
| 315 | 315.Dev Pain Finder Scrape real developer pain points for any keyword, technology, or problem space from Reddit, Hacker News, dev.to, and GitHub Discussions simultaneously — then group complaints by theme, score them… | tinyfish-io/ | 2.2k | — | ~2.6k | Automated safety check: Pass | MIT | today |
| 316 | 316.Deep Research 使用firecrawl和exa MCPs进行多源深度研究。搜索网络、综合发现并交付带有来源引用的报告。适用于用户希望对任何主题进行有证据和引用的彻底研究时。 | affaan-m/ | 276k | 2 repos | ~590 | Automated safety check: Pass | MIT | 4 days ago |
| 317 | Use when the user wants to interact with Facebook Marketplace — search visible listings, extract listing details, and summarize seller/profile… | browsing-skills/ | 117 | — | ~459 | Automated safety check: Pass | MIT | 4 mo ago |
| 318 | 318.Browser Use Direct browser control via CDP for web interaction: automation, scraping, testing, screenshots, and site/app work. | davidondrej/ | 4.1k | 1 repo | ~2.6k | Automated safety check: Pass | MIT | today |
| 319 | Turn a single LinkedIn post URL into a paused cold email campaign in Instantly. | gethouston/ | 117 | — | ~2.1k | Automated safety check: Pass | MIT | today |
| 320 | 320.Junta Leiloeiros Coleta e consulta dados de leiloeiros oficiais de todas as 27 Juntas Comerciais do Brasil. | sickn33/ | 47k | 2 repos | ~1.6k | Automated safety check: Pass | MIT | today |
| 321 | Automate Scrape Do tasks via Rube MCP (Composio). An agent skill from ComposioHQ/awesome-claude-skills. | ComposioHQ/ | 77k | 3 repos | ~738 | Automated safety check: Pass | No licence | 21 days ago |
| 322 | 322.Page Prep Prepare any webpage for clean interaction by detecting and removing disruptive overlays (cookie banners, GDPR consent, modals, popups, newsletter signups, paywalls, login walls). | adobe/ | 197 | — | ~2.1k | Automated safety check: Pass | Apache-2.0 | today |
| 323 | Read your Instagram niche and profile from real data via Apify, no login. | sergebulaev/ | 328 | — | ~1.3k | Automated safety check: Pass | MIT | 2 days ago |
| 324 | Build a local-business lead database from Google Maps in one Apify pipeline: search by target audience + geography, enrich each place with company contacts from its website, leads enrichment (names… | apify/ | 265 | — | ~3.8k | Automated safety check: Pass | Apache-2.0 | 16 days ago |
| 325 | 325.Reels Scripting Turn a reference Instagram Reel into a script for your own Reel, tuned to your voice and repurposed from your newsletter content. | charlie947/ | 3.8k | — | ~2.6k | Automated safety check: Pass | MIT | 24 days ago |
| 326 | 326.Data Feeds Extract structured data from 40+ supported platforms (Amazon, LinkedIn, Instagram, TikTok, Facebook, YouTube, Reddit, and more) via the Bright Data CLI (bdata pipelines). | brightdata/ | 264 | — | ~2.2k | Automated safety check: Pass | MIT | yesterday |
| 327 | 327.Tradingview Datos de mercado de TradingView via APIs publicas internas sin auth: Scanner (~300 columnas con quote/indicadores tecnicos/financials/earnings/ratings/targets), Symbol Search v3 (ISIN/CUSIP/CIK)… | gauss314/ | 247 | — | ~4.7k | Automated safety check: Pass | MIT | 3 mo ago |
| 328 | Search Xiaohongshu (XHS / RedNote) notes by keyword with full field extraction including body text, topics/tags, image list URLs, video stream URL, publish timestamp, and all engagement stats… | browser-act/ | 6.1k | — | ~5.3k | Automated safety check: Pass | MIT | 1 mo ago |
| 329 | Extracts structured data from websites and APIs, delivering clean datasets in multiple formats. | ertugrulakben/ | 303 | — | ~3k | Automated safety check: Pass | MIT | 2 days ago |
| 330 | Scrape business data from Google Maps — names, phones, emails, websites, ratings, reviews. | gmapsscraper/ | 132 | — | ~1.2k | Automated safety check: Pass | MIT | 4 mo ago |
| 331 | Review and tune Prometheus configuration and performance. An agent skill from prometheus/prometheus-mcp. | prometheus/ | 120 | — | ~724 | Automated safety check: Pass | Apache-2.0 | 4 days ago |
| 332 | Produce an intensive, cited analytical report: executive summary, multi-angle findings, contrarian views, open questions, and full sources. | firecrawl/ | 117 | — | ~1.4k | Automated safety check: Pass | ISC | today |
| 333 | 333.Debug Crawler Investigate a failing crawler and propose a fix, starting from a dataset name or an issues.json artifact URL. | opensanctions/ | 832 | — | ~1.2k | Automated safety check: Notes | MIT | today |
| 334 | 334.Threads Writer Threads Content Specialist. An agent skill from stevenflanagan1/social-ai-team. | stevenflanagan1/ | 242 | — | ~2.6k | Automated safety check: Pass | No licence | 2 days ago |
| 335 | 335.Scrapling Web scraping with Scrapling - HTTP fetching, stealth browser automation, Cloudflare bypass, and spider crawling via CLI and Python. | Tommy-yw/ | 546 | 3 repos | ~2.3k | Automated safety check: Pass | MIT | 4 mo ago |
| 336 | Deprecated. An agent skill from apify/agent-skills. | apify/ | 2.4k | — | ~309 | Automated safety check: Pass | No licence | today |