Agent skill

Conference Speaker Scraper

by gooseworks-ai in gooseworks-ai/goose-skills

Extract speaker names, titles, companies, and bios from conference websites.

MITAuto-check passedData & Analytics

Install Conference Speaker Scraper

skills CLI
$ npx skills add gooseworks-ai/goose-skills --skill conference-speaker-scraper -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install gooseworks-ai/goose-skills conference-speaker-scraper --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/gooseworks-ai/goose-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/lead-generation/capabilities/conference-speaker-scraper .claude/skills/conference-speaker-scraper && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
conference-speaker-scraper
GitHub stars
1.2k
Used in
1 other repo
Token cost
~846 tokens
SKILL.md length
241 words
Files
3 (incl. scripts)
Skills in repo
273
Repo updated
First seen
Licence
MIT

At a glance

Extract speaker names, titles, companies, and bios from conference websites.

  • Works in 4 steps: Strategy A -- CSS class hints: Looks for… → Strategy B -- Heading + paragraph… → Strategy C -- JSON-LD structured data:… → …
  • Pre-event research and outreach targeting
  • SKILL.md covers Quick Start, How It Works, CLI Reference and Output Schema, plus 2 more sections
  • Runs Python scripts from its folder; calls python3; reaches linkedin.com and sagefuture2026.com

What it does

Conference Speaker Scraper is an agent skill from gooseworks-ai/goose-skills. Extract speaker names, titles, companies, and bios from conference websites. Supports direct HTML scraping and Apify web scraper fallback for JS-heavy sites. Use for pre-event research and outreach targeting.

Its SKILL.md is about 850 tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files, including scripts (for example `scripts/scrape_speakers.py` and `skill.meta.json`).

It sits in Data & Analytics, covering Web scraping. It works with Apify. The repository describes itself as: Library of Growth & GTM skills + data APIs for Claude Code, Codex, Cursor to run ads, social, content, lead gen, seo and data scraping. The licence is MIT.

When your agent uses it

  • Pre-event research and outreach targeting
  • Tasks that involve Web scraping

Example prompts

  • “/conference-speaker-scraper”

Requirements

  • Python 3

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Strategy A -- CSS class hints: Looks for speaker cards with class names containing "speaker", "presenter", "faculty", "panelist"…
  2. Strategy B -- Heading + paragraph patterns: Looks for repeated / + structures
  3. Strategy C -- JSON-LD structured data: Checks for with speaker data
  4. Strategy D -- Platform embeds: Detects Sched.com/Sessionize patterns used by many conferences

What it can do on your machine

Read from SKILL.md and the folder at commit 4bbe1ef. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • linkedin.com
    • sagefuture2026.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Conference Speaker Scraper loads about 846 tokens when it runs. Until then it costs about 59 tokens; SKILL.md has 241 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~59
When it runs · the whole SKILL.md, loaded when a task matches
~846

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from gooseworks-ai/goose-skills at commit 4bbe1ef, republished under its MIT licence (© gooseworks-ai). 241 words, ~846 tokens.

Download SKILL.mdSave it as .claude/skills/conference-speaker-scraper/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.
name
conference-speaker-scraper
description
Extract speaker names, titles, companies, and bios from conference websites. Supports direct HTML scraping and Apify web scraper fallback for JS-heavy sites. Use for pre-event research and outreach targeting.

Conference Speaker Scraper

Extract speaker names, titles, companies, and bios from conference website /speakers pages. Supports direct HTML scraping with multiple extraction strategies, plus Apify fallback for JS-heavy sites.

Quick Start

No API key needed for direct scraping mode.

bash
# Scrape speakers from a conference page
python3 skills/conference-speaker-scraper/scripts/scrape_speakers.py \
  --url "https://example.com/speakers"

# Use Apify for JS-heavy sites
python3 skills/conference-speaker-scraper/scripts/scrape_speakers.py \
  --url "https://example.com/speakers" --mode apify

# Custom conference name (otherwise inferred from URL)
python3 skills/conference-speaker-scraper/scripts/scrape_speakers.py \
  --url "https://example.com/speakers" --conference "Sage Future 2026"

# Output formats
python3 skills/conference-speaker-scraper/scripts/scrape_speakers.py --url URL --output json     # default
python3 skills/conference-speaker-scraper/scripts/scrape_speakers.py --url URL --output csv
python3 skills/conference-speaker-scraper/scripts/scrape_speakers.py --url URL --output summary

How It Works

Direct Mode (default)

Fetches the page HTML and tries multiple extraction strategies in order, using whichever returns the most results:

  1. Strategy A -- CSS class hints: Looks for speaker cards with class names containing "speaker", "presenter", "faculty", "panelist", "team-member"
  2. Strategy B -- Heading + paragraph patterns: Looks for repeated <h2>/<h3> + <p> structures
  3. Strategy C -- JSON-LD structured data: Checks for <script type="application/ld+json"> with speaker data
  4. Strategy D -- Platform embeds: Detects Sched.com/Sessionize patterns used by many conferences
Apify Mode

Uses apify/cheerio-scraper actor with a custom page function that targets common speaker card selectors. Standard POST/poll/GET dataset pattern.

CLI Reference

FlagDefaultDescription
--urlrequiredConference speakers page URL
--conferenceinferredConference name (otherwise inferred from URL domain)
--modedirectdirect (HTML scraping) or apify (Apify cheerio scraper)
--outputjsonOutput format: json, csv, or summary
--tokenenv varApify token (only needed for apify mode)
--timeout300Max seconds for Apify run

Output Schema

json
{
  "name": "Jane Smith",
  "title": "VP of Finance",
  "company": "Acme Corp",
  "bio": "Jane leads the finance transformation at...",
  "linkedin_url": "https://linkedin.com/in/janesmith",
  "image_url": "https://...",
  "conference": "Sage Future 2026",
  "source_url": "https://sagefuture2026.com/speakers"
}

Cost

  • Direct mode: Free (no API, no tokens)
  • Apify mode: Uses apify/cheerio-scraper -- minimal Apify credits

Testing Notes

HTML scraping is inherently fragile across conference sites. The multi-strategy approach maximizes coverage, but JS-heavy sites will require Apify mode. When direct scraping returns 0 results, try --mode apify.

© gooseworks-ai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 2 other files (scripts) in skills/lead-generation/capabilities/conference-speaker-scraper of gooseworks-ai/goose-skills.

  • SKILL.md
  • scripts/scrape_speakers.py
  • skill.meta.json

Open the folder on GitHubat commit 4bbe1ef

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in gooseworks-ai/goose-skills, which our catalogue first saw on October 9, 2026.

Compare with similar skills

Conference Speaker Scraper next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Conference Speaker Scraper compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Conference Speaker Scraper this skillgooseworks-ai/goose-skills1.2k1 repos~846Automated safety check: PassMIT
Apify CLIapify/apify-cli256—~1.5kAutomated safety check: PassApache-2.0
Apify Collectextrasmall0/dear-hiring-manager111—~1.1kAutomated safety check: NotesMIT
Apify Lead Scoring Enrichmentapify/awesome-skills266—~4.4kAutomated safety check: NotesApache-2.0
Carousel Benchmarknestyme/awesome-prompts151—~2.6kAutomated safety check: NotesNone
Linkedin Thread Monitorsergebulaev/linkedin-skills4.4k1 repos~1.4kAutomated safety check: PassMIT

Similar skills

  • Apify CLI

    apify/apify-cli

    Official

    Patterns for invoking the Apify CLI (apify) from agents. An agent skill from apify/apify-cli.

    256 GitHub stars~1.5k tokensUpdated yesterday
    Data & AnalyticsAuto-check passed
  • Apify Collect

    extrasmall0/dear-hiring-manager

    Collect fresh job-posting URLs into ~/.dear-hiring-manager/urls.txt by running an Apify job scraper — the discovery source for /batch.

    111 GitHub stars~1.1k tokensUpdated 2 mo ago
    Data & AnalyticsAuto-check: notes
  • Apify Lead Scoring Enrichment

    apify/awesome-skills

    Official

    Score and enrich a CSV of B2B leads using Apify Actors. An agent skill from apify/awesome-skills.

    266 GitHub stars~4.4k tokensUpdated 18 days ago
    Data & AnalyticsAuto-check: notes
  • Carousel Benchmark

    nestyme/awesome-prompts

    Find the TikTok photo-mode carousels (slideshows) that actually go viral in a niche and the accounts behind them, rank them by organic quality (save-rate, like-rate, boost detection) instead of raw…

    151 GitHub stars~2.6k tokensUpdated 11 days ago
    Data & AnalyticsAuto-check: notes
  • Linkedin Thread Monitor

    sergebulaev/linkedin-skills

    Track which of your LinkedIn comments earned author replies.

    4.4k GitHub starsUsed in 1 repo~1.4k tokens
    Data & AnalyticsAuto-check passed
  • Apify Ashby Jobs Scraper

    apify/awesome-skills

    Official

    Scrape Ashby jobs or discover companies using Ashby with the Apify Ashby Job Board API Actor (johnvc/ashby-job-board-scraper).

    266 GitHub stars~3.7k tokensUpdated 18 days ago
    Data & AnalyticsAuto-check passed

More from gooseworks-ai/goose-skills

All 273 skills in this repo
  • Reddit Post Finder

    gooseworks-ai/goose-skills

    Scrape and search Reddit posts using Apify. An agent skill from gooseworks-ai/goose-skills.

    1.2k GitHub starsUsed in 1 repo~1.2k tokens
    Auto-check passed
  • Create Image Fal

    gooseworks-ai/goose-skills

    Generate or edit an image via any FAL image model (nano-banana edit, gpt-image, flux, ...), ROUTED THROUGH THE fal-proxy so it bills the Ads agent.

    1.2k GitHub stars~1.3k tokensUpdated today
    Auto-check passed
  • Render Hook Replacement

    gooseworks-ai/goose-skills

    Replace an existing video's opening with a supplied clip or free kinetic text hook while retaining and verifying every original body frame, audio, captions and ending.

    1.2k GitHub stars~2.3k tokensUpdated today
    Auto-check passed
  • Blog Feed Monitor

    gooseworks-ai/goose-skills

    Scrape blog posts via RSS feeds (free, no API key) with Apify fallback for JS-heavy sites.

    1.2k GitHub starsUsed in 1 repo~578 tokens
    Auto-check passed
  • Competitor Post Engagers

    gooseworks-ai/goose-skills

    Find leads by scraping engagers from a competitor's top LinkedIn posts.

    1.2k GitHub starsUsed in 1 repo~1.8k tokens
    Auto-check: notes
  • Render Chatgpt Chat

    gooseworks-ai/goose-skills

    Assemble a ChatGPT chat-reveal video ad from a thread + timeline JSON — one continuous Playwright recording of a ChatGPT mobile chat (user types with the iOS keyboard up → taps send → keyboard…

    1.2k GitHub stars~2.3k tokensUpdated today
    Auto-check passed

Works with

Questions about Conference Speaker Scraper

What does Conference Speaker Scraper do?

Extract speaker names, titles, companies, and bios from conference websites. Conference Speaker Scraper is an agent skill from gooseworks-ai/goose-skills. Extract speaker names, titles, companies, and bios from conference websites.

When should I use Conference Speaker Scraper?

Conference Speaker Scraper fits situations like: pre-event research and outreach targeting; tasks that involve Web scraping.

How do I install Conference Speaker Scraper in Claude Code?

Run `npx skills add gooseworks-ai/goose-skills --skill conference-speaker-scraper -a claude-code`. Or copy the skill folder (skills/lead-generation/capabilities/conference-speaker-scraper in gooseworks-ai/goose-skills) into .claude/skills/conference-speaker-scraper in your project. Claude Code loads it when a task matches its description.

How do I install Conference Speaker Scraper in Codex?

Run `npx skills add gooseworks-ai/goose-skills --skill conference-speaker-scraper -a codex`. Or copy the skill folder (skills/lead-generation/capabilities/conference-speaker-scraper in gooseworks-ai/goose-skills) into .agents/skills/conference-speaker-scraper in your project. Codex loads it when a task matches its description.

Can I use Conference Speaker Scraper in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add gooseworks-ai/goose-skills --skill conference-speaker-scraper -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/conference-speaker-scraper, .gemini/skills/conference-speaker-scraper, .github/skills/conference-speaker-scraper and .opencode/skills/conference-speaker-scraper in your project.

What does Conference Speaker Scraper need to run?

Going by SKILL.md and its folder, Conference Speaker Scraper needs Python for the scripts in its folder and the command-line tools its instructions call (python3). Our summary lists: Python 3.

Does Conference Speaker Scraper access the network?

SKILL.md names 2 domains. In commands or code: linkedin.com and sagefuture2026.com; the agent is likely to contact these when it follows the instructions. This is read from the text; nothing was executed.

Is Conference Speaker Scraper safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Conference Speaker Scraper use?

Conference Speaker Scraper is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Conference Speaker Scraper use?

About 846 tokens (SKILL.md is roughly 3.4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Conference Speaker Scraper?

Skills that share tags, products or a category with Conference Speaker Scraper: Apify CLI (apify/apify-cli, 256 stars), Apify Collect (extrasmall0/dear-hiring-manager, 111 stars), Apify Lead Scoring Enrichment (apify/awesome-skills, 266 stars) and Carousel Benchmark (nestyme/awesome-prompts, 151 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Conference Speaker Scraper?

gooseworks-ai (a GitHub organization) maintains it in gooseworks-ai/goose-skills, which has 1,242 GitHub stars. The repository holds 273 skills in this directory. The repository was last updated on October 10, 2026.

Source: gooseworks-ai/goose-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.