Search
Python · Web scraping
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Adds Firecrawl's /scrape endpoint to application code to pull markdown, HTML, links, screenshots or structured data from a single known URL. | firecrawl/ | 190k | 1 repo | ~944 | Automated safety check: Pass | ISC | today |
| 2 | Guides adding Firecrawl's /interact endpoint to product code for pages that need clicks, forms, pagination or logged-in flows beyond plain scraping. | firecrawl/ | 190k | 1 repo | ~731 | Automated safety check: Pass | ISC | today |
| 3 | Scrapes sites, handles JavaScript-heavy pages and extracts structured data with Crawl4AI, through its crwl CLI or Python SDK, including schema-based extraction without an LLM. | smallnest/ | 599 | 1 repo | ~2.5k | Automated safety check: Pass | MIT | 6 mo ago |
| 4 | Scrape BOSS直聘 (job listing site) via Chrome CDP. An agent skill from eatmoreduck/boss-zhipin-scraper. | eatmoreduck/ | 1.5k | — | ~2.6k | Automated safety check: Pass | MIT | 10 days ago |
| 5 | Automates websites with Skyvern's AI browser agent to fill forms, extract data, download files, log in and run multi-step workflows through SDKs, REST, MCP or a CLI. | Skyvern-AI/ | 23k | — | ~1.9k | Automated safety check: Pass | AGPL-3.0 | today |
| 6 | 6.Ax Use the ax CLI instead of curl + throwaway parsing scripts whenever you fetch a URL, explore an unknown web page, or extract structured data from HTML. | yusukebe/ | 719 | 1 repo | ~918 | Automated safety check: Pass | MIT | 2 mo ago |
| 7 | Guides building Tavily integrations for web search, URL extraction, site crawling and AI-assisted research in Python or JavaScript agent and RAG projects. | andrewyng/ | 14k | — | ~1.1k | Automated safety check: Pass | MIT | 4 mo ago |
| 8 | Finds new job postings that match your profile through installed portal-search CLIs, dedupes against past runs and your application tracker, and rates each one's fit. | MadsLorentzen/ | 45k | — | ~5.7k | Automated safety check: Pass | MIT | today |
| 9 | Scrape Google Maps business listings (name, address, phone, website, rating, reviews, lat/lng, hours, emails) via the local gosom google-maps-scraper REST API. | Mahanaicoach/ | 1.3k | — | ~2.8k | Automated safety check: Pass | MIT | 4 days ago |
| 10 | Execute Python code in a safe sandboxed environment via [inference.sh](https://inference.sh). | cortega26/ | 113 | 2 repos | ~1.5k | Automated safety check: Pass | MIT | today |
| 11 | 11.Scrapling 使用 scrapling 进行网页抓取和数据提取。根据目标网站特征自动选择最佳 Fetcher, 生成并执行 Python 脚本完成任务。Use when: (1) 抓取/爬取网页内容或数据(scrape, crawl, fetch page, extract data) (2) 需要绕过 Cloudflare/WAF 等反爬保护 (3) 登录后抓取受保护页面 (4) 解析已有 HTML… | Cedriccmh/ | 444 | — | ~1.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 12 | Query 57 Indonesian government APIs and data sources — BPJPH halal certification, BPOM food safety, OJK financial legality, BPS statistics, BMKG weather/earthquakes, Bank Indonesia exchange rates… | suryast/ | 172 | — | ~997 | Automated safety check: Pass | MIT | yesterday |
| 13 | Scrape video stats from the Douyin creator center (creator.douyin.com) using AppleScript to automate the user's logged-in Chrome. | TradingAi666/ | 406 | — | ~7.2k | Automated safety check: Pass | MIT | 4 mo ago |
| 14 | Downloads and manages WeChat official account articles from URLs, an account's history or subscriptions, adding engagement metrics and top comments when a valid session exists. | Moore-developers/ | 294 | — | ~3.4k | Automated safety check: Pass | MIT | 2 mo ago |
| 15 | Surveys open Korean government startup and R&D support programs and sorts them by fit with your project, checking eligibility against the original notices. | djfksjd/ | 392 | — | ~3.5k | Automated safety check: Notes | MIT | 2 mo ago |
| 16 | Generate working code that routes HTTP requests through Bright Data proxy networks (Datacenter, ISP, Residential, Mobile) and help users decide which network and IP pool type to use (shared pool… | brightdata/ | 264 | — | ~5.1k | Automated safety check: Pass | MIT | 2 days ago |
| 17 | Selectively monitor important long-running or resource-intensive commands with Haoleme by prefixing them with hao, so status, output, and completion notifications sync to the mobile app. | HaolemeApp/ | 157 | — | ~1.3k | Automated safety check: Pass | AGPL-3.0 | 1 mo ago |
| 18 | Pulls structured Amazon product reviews for an ASIN through BrowserAct's Amazon Reviews API, with no Amazon login, using a bundled Python script. | browser-act/ | 6.1k | 1 repo | ~1.4k | Automated safety check: Pass | MIT | 1 mo ago |
| 19 | Creates, changes, debugs and deploys Apify Actors, including their input and output schemas, using the Apify CLI. | apify/ | 2.4k | — | ~2.9k | Automated safety check: Pass | No licence | today |
| 20 | 20.Web Re Web reverse engineering tools and workflows. An agent skill from schlarpc/re-shell. | schlarpc/ | 532 | — | ~1.4k | Automated safety check: Pass | No licence | 1 mo ago |
| 21 | Search Google, scrape web pages, Amazon product pages, YouTube subtitles, or Reddit (post/subreddit) using the Decodo Scraper OpenClaw Skill. | Decodo/ | 153 | — | ~1.5k | Automated safety check: Notes | No licence | 7 mo ago |
| 22 | Crawls a site with Crawl4AI to audit titles, meta tags, H1s, canonicals, navigation and internal links, and to compare landing pages and competitor sites. | artwist-polyakov/ | 206 | — | ~1.7k | Automated safety check: Pass | MIT | yesterday |
| 23 | Finds authoritative public data sources for modeling tasks, prefers official APIs and bulk downloads, and outputs a reproducible fetch and cleaning plan with citations. | yushui2022/ | 453 | 1 repo | ~1.1k | Automated safety check: Pass | MIT | 2 days ago |
| 24 | 24.Leetcode Py Generates Python LeetCode practice environments and manages a 307-problem catalog with the lcpy CLI. | wislertt/ | 142 | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 25 | Runs web-page JavaScript without a browser, in a V8 sandbox with a browser-environment shim, for scripts that only probe the environment and compute a result. | taxueseek/ | 186 | — | ~400 | Automated safety check: Pass | MIT | 3 days ago |
| 26 | Build high-quality literature reviews from a research topic using a 10-phase workflow. | Drchronx/ | 137 | — | ~2.5k | Automated safety check: Pass | Unknown | 4 mo ago |
| 27 | 27.Tmux Remote control tmux sessions for interactive CLIs (python, gdb, etc.) by sending keystrokes and scraping pane output. | mitsuhiko/ | 3.2k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | 11 days ago |
| 28 | Web data extraction and discovery using the Bright Data JavaScript/TypeScript SDK (@brightdata/sdk). | brightdata/ | 264 | — | ~3k | Automated safety check: Pass | MIT | 2 days ago |
| 29 | Routes a web request through the cheapest tier that can finish it, from headless extraction with WAF bypass up to a real stealth or signed-in browser, with screenshots as proof. | code-yeongyu/ | 70k | — | ~2.7k | Automated safety check: Warn | Unknown | today |
| 30 | 30.Tmux Remote control tmux sessions for interactive CLIs (python, gdb, etc.) by sending keystrokes and scraping pane output. | archibate/ | 108 | — | ~1.4k | Automated safety check: Pass | No licence | 5 mo ago |
| 31 | 31.Geo Measure Measure GEO visibility from an approved, file-backed engine observation bundle. | yaojingang/ | 165 | — | ~348 | Automated safety check: Pass | AGPL-3.0 | 1 mo ago |
| 32 | Extracts wholesale product data from 1688.com offer pages, including tiered prices, SKU variants, seller info, promotions and review stats, with bundled Python scripts. | browser-act/ | 6.1k | — | ~3.3k | Automated safety check: Pass | MIT | 1 mo ago |
| 33 | Fetches full Airbnb listing details for a numeric listing ID, including amenities, photos, house rules and ratings, through the site's internal GraphQL API. | browser-act/ | 6.1k | — | ~1.4k | Automated safety check: Pass | MIT | 1 mo ago |
| 34 | Extracts Airbnb search results - listing id, name, coordinates, rating, price, photos and badges - from the page's own embedded search data, with pagination. | browser-act/ | 6.1k | — | ~1.4k | Automated safety check: Pass | MIT | 1 mo ago |
| 35 | eBay sold-listings scraper across 8 marketplaces (ebay.com/.co.uk/.de/.fr/.it/.es/.ca/.com.au). | browser-act/ | 6.1k | — | ~4.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 36 | Extract complete product information from any e-commerce product page. | browser-act/ | 6.1k | — | ~1.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 37 | Etsy category page scraper: given an Etsy category URL (e.g. | browser-act/ | 6.1k | — | ~2k | Automated safety check: Pass | MIT | 1 mo ago |
| 38 | Etsy product detail scraper: given an Etsy listing URL, returns full product detail including listingId, title, priceCurrent, priceOriginal, currency, images (all), description, shopName, shopUrl… | browser-act/ | 6.1k | — | ~2.3k | Automated safety check: Pass | MIT | 1 mo ago |
| 39 | Etsy shop catalog scraper: given an Etsy shop URL (e.g. An agent skill from browser-act/skills. | browser-act/ | 6.1k | — | ~1.9k | Automated safety check: Pass | MIT | 1 mo ago |
| 40 | Searches Meta Ad Library (Facebook/Instagram/WhatsApp ads) by keyword or Facebook page ID and extracts ad details including creatives, copy, CTA, publisher platforms, spend, impressions, reach… | browser-act/ | 6.1k | — | ~2.3k | Automated safety check: Pass | MIT | 1 mo ago |
| 41 | Scrapes posts from any public Facebook Page timeline, returning structured data including post text, author info, engagement metrics (likes/comments/shares), reaction breakdowns… | browser-act/ | 6.1k | — | ~2.3k | Automated safety check: Pass | MIT | 1 mo ago |
| 42 | Scrapes posts from any public Facebook Page or personal Profile timeline, returning structured data including post text, author info with profile picture, engagement metrics (likes/comments/shares)… | browser-act/ | 6.1k | — | ~2.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 43 | Scrapes second-hand item search results from Goofish (闲鱼/xianyu, goofish.com) — China's largest second-hand marketplace. | browser-act/ | 6.1k | — | ~2k | Automated safety check: Pass | MIT | 1 mo ago |
| 44 | Extracts Google Search results page (SERP) data including organic results, paid ads, related searches, People Also Ask questions, AI Overview text, and total result count from google.com. | browser-act/ | 6.1k | — | ~2k | Automated safety check: Pass | MIT | 1 mo ago |
| 45 | Scrapes Instagram posts tagged at a specific location or place, returning media items with captions, like/comment counts, media URLs and user info. | browser-act/ | 6.1k | — | ~1.9k | Automated safety check: Pass | MIT | 1 mo ago |
| 46 | Fetches comments from an Instagram post including comment text, username, timestamp, like count and reply count. | browser-act/ | 6.1k | — | ~1.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 47 | Fetches Instagram user profile metadata including bio, follower count, following count, post count, verification status and other profile details. | browser-act/ | 6.1k | — | ~1.2k | Automated safety check: Pass | MIT | 1 mo ago |
| 48 | Scrapes posts from an Instagram user's profile feed including captions, media URLs, like/comment counts, timestamps and location tags. | browser-act/ | 6.1k | — | ~1.7k | Automated safety check: Pass | MIT | 1 mo ago |