Topic · Data & Analytics

Best web scraping skills, page 10

Skills #433–480 of 777, ranked by score.

Web scraping skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Web scraping skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
433

OPS on-demand: This skill should be used when the user asks to "leadgen drafts", "cold email approve"…

Lifecycle-Innovations-Limited/claude-ops540—~990Automated safety check: NotesMITyesterday
434

Generate a personalized outbound message (LinkedIn DM or cold email) for a single prospect by combining a template with the lead's profile, company, and recent signals.

Othmane-Khadri/YALC-the-GTM-operating-system317—~1.3kAutomated safety check: NotesMIT1 mo ago
435

Pull the audience that engaged with a LinkedIn post — likers (reactors) and commenters — via Unipile, dedupe across endpoints, and persist them as a result set ready for qualification or campaign…

Othmane-Khadri/YALC-the-GTM-operating-system317—~1.4kAutomated safety check: NotesMIT1 mo ago
436

Generate a structured interview preparation guide for any company by scraping real candidate experiences from Glassdoor, Blind, and Reddit in real time using parallel TinyFish agents.

tinyfish-io/tinyfish-cookbook2.2k—~2.3kAutomated safety check: PassMIT7 days ago
437
437.InfrastructureOfficial

Ship Kubernetes, host, container, and cloud-provider telemetry into Grafana Cloud — k8s-monitoring Helm chart for K8s clusters (metrics + logs + traces + events + cost), Alloy…

grafana/skills279—~1.2kAutomated safety check: PassApache-2.02 days ago
438

Search Yelp for local businesses, get contact info, ratings, and hours.

letta-ai/skills147—~1.4kAutomated safety check: NotesMIT7 days ago
439

Faithfully archive public WPS/KDocs/金山文档 links, especially embedded ProcessOn .pof mind maps and canvases, as raw source data, original SVG/PNG, and Markdown.

daymade/claude-code-skills1.4k—~1.1kAutomated safety check: PassMITyesterday
440

Non-testing browser automation - web scraping, form filling, screenshot capture, PDF generation, workflow automation.

ynulihao/AgentSkillOS617—~2.2kAutomated safety check: PassNo licence7 mo ago
441

Debug Bright Data Scraping Browser sessions using the Browser Sessions API.

brightdata/skills264—~2kAutomated safety check: PassMITyesterday
442

Web data extraction and discovery using the Bright Data Python SDK.

brightdata/skills264—~5.2kAutomated safety check: PassMITyesterday
443

Build production-ready web scrapers for any website using Bright Data infrastructure.

brightdata/skills264—~7.2kAutomated safety check: PassMITyesterday
444

Monitor competitor pricing pages via live web scrape and Web Archive snapshots.

gooseworks-ai/goose-skills1.2k1 repo~2.2kAutomated safety check: PassMITyesterday
445

Extract leads from competitor product activity — Product Hunt commenters/upvoters, HN posts about competitors, case studies, testimonials, tech press, and switching signals.

gooseworks-ai/goose-skills1.2k1 repo~3kAutomated safety check: NotesMITyesterday
446

Monitor web sources for Series A-C funding announcements. An agent skill from gooseworks-ai/goose-skills.

gooseworks-ai/goose-skills1.2k1 repo~2.3kAutomated safety check: PassMITyesterday
447

Prepare for investor calls by pulling upcoming meetings from Google Calendar, deeply researching each investor and their firm (website scraping, portfolio analysis, thesis extraction), checking for…

gooseworks-ai/goose-skills1.2k1 repo~3.8kAutomated safety check: PassMITyesterday
448

Search for job postings across LinkedIn and Indeed. An agent skill from gooseworks-ai/goose-skills.

gooseworks-ai/goose-skills1.2k1 repo~2.6kAutomated safety check: NotesMITyesterday
449

Track what key opinion leaders (KOLs) in your space are posting on LinkedIn and Twitter/X.

gooseworks-ai/goose-skills1.2k1 repo~1.7kAutomated safety check: PassMITyesterday
450

Lead qualification engine with conversational intake. An agent skill from gooseworks-ai/goose-skills.

gooseworks-ai/goose-skills1.2k1 repo~3.8kAutomated safety check: PassMITyesterday
451

Generates Instagram-ready product reels from any e-commerce product page URL.

gooseworks-ai/goose-skills1.2k1 repo~2.3kAutomated safety check: NotesMITyesterday
452

Scrape G2, Capterra, and Trustpilot reviews for your product and competitors, then extract recurring themes, objections, proof points, and exact customer language for use in messaging.

gooseworks-ai/goose-skills1.2k1 repo~1.7kAutomated safety check: PassMITyesterday
453

Pull real SEO metrics for any domain using Apify scrapers for Semrush and Ahrefs data.

gooseworks-ai/goose-skills1.2k1 repo~2.2kAutomated safety check: PassMITyesterday
454

Detect buying signals across TAM companies and watchlist personas.

gooseworks-ai/goose-skills1.2k1 repo~1.4kAutomated safety check: NotesMITyesterday
455

Scrape websites, extract structured data, and automate browsers.

gooseworks-ai/goose-skills1.2k1 repo~4.6kAutomated safety check: PassMITyesterday
456

Bulk extract content from an entire website or site section.

aiskillstore/marketplace4303 repos~673Automated safety check: PassNo licenceyesterday
457

Web search and scraping via Firecrawl API. An agent skill from sundial-org/awesome-openclaw-skills.

sundial-org/awesome-openclaw-skills6631 repo~250Automated safety check: NotesNo licence7 mo ago
458

Discover and list a site's URLs, with search filtering. An agent skill from firecrawl/skills.

firecrawl/skills1161 repo~412Automated safety check: PassISC2 days ago
459

Use before adding or changing Python code that calls a third-party HTTP API from PostHog (a vendor REST call, a vendor SDK client, a scraping or enrichment service), and before adding or changing a…

PostHog/posthog40k—~2.7kAutomated safety check: PassUnknownyesterday
460

Scrape Reddit with the harshmaur/reddit-scraper Apify Actor: keyword search across all of Reddit or inside one subreddit, full subreddit listings, post permalinks with complete comment threads, user…

apify/awesome-skills264—~3.9kAutomated safety check: PassApache-2.016 days ago
461

Add new sources to your knowledge base or re-scrape existing ones to pick up changes.

techwolf-ai/ai-first-toolkit1321 repo~1.3kAutomated safety check: PassMIT9 days ago
462
462.Scrape

Scrape any webpage as clean markdown via Bright Data Web Unlocker API.

davila7/claude-code-templates32k—~392Automated safety check: PassMITyesterday
463

Check whether a website's robots.txt allows the AI crawlers that decide visibility in ChatGPT Search, Perplexity, Claude, Gemini, and Microsoft Copilot.

davepoon/buildwithclaude3.6k1 repo~1.5kAutomated safety check: PassMIT2 days ago
464

DEPRECATED in v0.2.0 -- use browser-extract instead; this is a thin shim for backward compatibility, removed in v0.3.0

ruvnet/ruflo74k—~405Automated safety check: NotesMITyesterday
465

小红书作品爬取工具。根据关键词爬取小红书热门作品数据,支持按日期范围、排序方式筛选,结果以结构化表格展示。当用户需要爬取小红书作品、查询小红书热门内容、搜索小红书爆款笔记时使用。触发词:小红书爬取、小红书作品、小红书爆款、小红书搜索、小红书热门、小红书笔记查询。

redfox-data/redfox-community425—~1.5kAutomated safety check: PassNo licenceyesterday
466
466.X Twitter ScraperOfficial

Build GitHub Copilot workflows with Xquik X API SDKs, REST endpoints, hosted Apify Actor runs, MCP tools, TweetClaw OpenClaw plugin installs, signed webhooks, tweet search, user lookup, follower…

github/awesome-copilot40k—~2.2kAutomated safety check: PassMITyesterday
467

Quotes, fundamentals, insider, analyst ratings, earnings estimates, financial summary, income/balance/cashflow detail y company profile (delayed 15-20min).

gauss314/skills246—~5.4kAutomated safety check: PassMIT3 mo ago
468

Run Xquik's Apify Actor for X followers, following, verified audiences, lists, communities, and overlap research.

Varnan-Tech/opendirectory674—~1.6kAutomated safety check: PassMIT1 mo ago
469

Run Xquik's Apify Actor for X searches, posts, timelines, conversations, lists, articles, and engagement research.

Varnan-Tech/opendirectory674—~1.7kAutomated safety check: PassMIT1 mo ago
470

A skill your agent uses when writing Playwright automation code, building web scrapers, or creating E2E tests - provides best practices for selector strategies, waiting patterns, and robust…

ed3dai/ed3d-plugins250—~3.6kAutomated safety check: PassNo licence1 mo ago
471

Run ad-hoc web research on a prospect (company URL or person) — Firecrawl scrapes the marketing site, key pages (pricing, careers, blog), and surfaces a brief.

Othmane-Khadri/YALC-the-GTM-operating-system317—~490Automated safety check: NotesMIT1 mo ago
472

Pull competitive intelligence on a competitor — pricing, positioning, recent changes, customer stories — via the competitive-intel CLI.

Othmane-Khadri/YALC-the-GTM-operating-system317—~445Automated safety check: NotesMIT1 mo ago
473

Produce cited research reports from multiple web sources using firecrawl and exa MCP tools — plan sub-questions, search and deep-read sources, then synthesize findings with inline citations and…

affaan-m/ECC275k—~1.5kAutomated safety check: WarnMIT4 days ago
474

Universal archivist for personal file archives (Dropbox/B2/Gmail-takeout/local-mount/hard-drive-dump).

inbrainfun/inbrain1421 repo~2.6kAutomated safety check: PassUnknown2 mo ago
475

Ingest public or authenticated knowledge bases and docs portals with Firecrawl browser.

firecrawl/skills116—~565Automated safety check: PassISC2 days ago
476

Analyze competitor strategies, content, pricing, ads, and market positioning across Google Maps, Booking.com, Facebook, Instagram, YouTube, and TikTok.

majiayu000/claude-skill-registry6664 repos~1.3kAutomated safety check: NotesMITyesterday
477

Crawl entire websites using Cloudflare Browser Rendering /crawl API.

davila7/claude-code-templates32k—~2.6kAutomated safety check: NotesMITyesterday
478

X API & Twitter scraper skill for AI coding agents. An agent skill from davila7/claude-code-templates.

davila7/claude-code-templates32k—~1.8kAutomated safety check: PassMITyesterday
479

A skill your agent uses to report Omni-Channel Pending Service Routing usage without changing the org: query the current PendingServiceRouting count, compare it with an admin-supplied maximum or a…

forcedotcom/sf-skills1.1k—~964Automated safety check: NotesApache-2.0yesterday
480

Extracts Feishu/Lark Docs, Wiki, Sheets, and Minutes (妙记) transcripts into faithful local Markdown via the lark-cli API — no LLM paraphrasing, browser-DOM fallback when lark-cli can't reach content.

daymade/claude-code-skills1.4k—~6.9kAutomated safety check: PassMITyesterday