Agent skill

Job Scraper

by gooseworks-ai in gooseworks-ai/goose-skills

Search for job postings across LinkedIn and Indeed. An agent skill from gooseworks-ai/goose-skills.

MITAuto-check: notesData & Analytics

Install Job Scraper

skills CLI
$ npx skills add gooseworks-ai/goose-skills --skill job-scraper -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install gooseworks-ai/goose-skills job-scraper --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/gooseworks-ai/goose-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/lead-generation/capabilities/job-scraper .claude/skills/job-scraper && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
job-scraper
GitHub stars
1.2k
Used in
1 other repo
Token cost
~2.6k tokens
SKILL.md length
1,065 words
Files
2
Skills in repo
273
Repo updated
First seen
Licence
MIT

At a glance

Search for job postings across LinkedIn and Indeed. An agent skill from gooseworks-ai/goose-skills.

  • Works in 5 steps: Understand the Request → Search → Filter & Deduplicate → …
  • Users want to find open roles
  • SKILL.md covers When to Auto-Load, Prerequisites, Sources and Workflow, plus 3 more sections
  • Calls curl and stripe; reaches api.apify.com; needs APIFY_API_TOKEN

What it does

Job Scraper is an agent skill from gooseworks-ai/goose-skills. Search for job postings across LinkedIn and Indeed. Use when users want to find open roles, monitor hiring signals, identify companies hiring for specific positions, or research competitor hiring activity. Returns job title, company, location, salary, description, seniority level, and direct apply URLs. No login or cookies required.

Its SKILL.md is about 2.6k tokens, which your agent loads only when the skill is triggered. The skill folder holds 1 other file (for example `skill.meta.json`).

It sits in Data & Analytics, covering Web scraping. It works with LinkedIn and Apify. The repository describes itself as: Library of Growth & GTM skills + data APIs for Claude Code, Codex, Cursor to run ads, social, content, lead gen, seo and data scraping. The licence is MIT.

When your agent uses it

  • Users want to find open roles
  • Monitor hiring signals
  • Identify companies hiring for specific positions
  • Research competitor hiring activity

Example prompts

  • “/job-scraper”

Requirements

  • A credential in APIFY_API_TOKEN

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. Understand the Request
  2. Search
  3. Filter & Deduplicate
  4. Present Results
  5. Export (Optional)

What it can do on your machine

Read from SKILL.md and the folder at commit c650c6d. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • curl
    • stripe

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • api.apify.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • APIFY_API_TOKEN

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Job Scraper loads about 2.6k tokens when it runs. Until then it costs about 87 tokens; SKILL.md has 1,065 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~87
When it runs · the whole SKILL.md, loaded when a task matches
~2.6k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NoteMentions a .env fileSKILL.md:29
    th LinkedIn and Indeed scraping. Set in `.env`:
  • NoteMentions a .env fileSKILL.md:264
    _TOKEN` not set | Ask user to add it to `.env` |

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from gooseworks-ai/goose-skills at commit c650c6d, republished under its MIT licence (© gooseworks-ai). 1,065 words, ~2,564 tokens.

Download SKILL.mdSave it as .claude/skills/job-scraper/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
job-scraper
description
Search for job postings across LinkedIn and Indeed. Use when users want to find open roles, monitor hiring signals, identify companies hiring for specific positions, or research competitor hiring activity. Returns job title, company, location, salary, description, seniority level, and direct apply URLs. No login or cookies required.
tags
lead-generation, research

Job Scraper

Search for job postings across LinkedIn and Indeed using Apify. Find open roles by keyword, location, company, or job type. Use for hiring signal detection, GTM research, or competitive intelligence.

No LinkedIn cookies. No Indeed login. Just search queries in, structured job data out.

When to Auto-Load

Load this skill when:

  • User says "find jobs", "who is hiring", "what roles is [company] hiring for"
  • User wants hiring signals ("find companies growing their AI team")
  • User wants competitive intelligence ("what is [competitor] hiring for")
  • User says "job search", "open roles", "job listings", "job postings"

Prerequisites

Apify API Token

Required for both LinkedIn and Indeed scraping. Set in .env:

APIFY_API_TOKEN=your_token_here

No LinkedIn cookies, Indeed login, or any platform credentials needed. That's the only setup.


Sources

This skill searches two job platforms via Apify actors:

SourceApify ActorBest ForCost
LinkedInautomation-lab/linkedin-jobs-scraperB2B, tech, SaaS, enterprise roles. Has seniority level, job function, industries.~$0.002/job
Indeedborderline/indeed-scraperBroadest coverage. Richest data — salary, company details, ratings, contacts, street addresses.~$0.004/job
Source Selection Logic

Do NOT ask the user which source to use unless genuinely ambiguous. Decide based on context:

  1. User specifies a source → use that source only.
  2. Context strongly suggests one source:
    • B2B/tech/SaaS roles, enterprise companies, seniority-level filtering → LinkedIn
    • Hourly/blue-collar roles, local/retail jobs, salary-focused search → Indeed
    • Company hiring research ("what is Stripe hiring for") → LinkedIn (better company filtering)
  3. No clear signal → search both sources, deduplicate results by job title + company name, present combined results.

After deciding, tell the user which source(s) you're searching and why. Don't ask — inform.


Workflow

Phase 0: Understand the Request

Extract from the user's message:

  • Search term — job title, role, or keyword (required)
  • Location — city, state, country, or "Remote" (optional)
  • Company — specific company name (optional)
  • Recency — "recent", "last week", "last 30 days" (optional)
  • Job type — fulltime, parttime, contract, internship (optional)
  • Remote — whether to filter for remote jobs (optional)
  • Result count — how many results they want (default: 25)

If anything is ambiguous, pick reasonable defaults and tell the user what you chose. Do not ask clarifying questions for things you can reasonably infer.

LinkedIn — automation-lab/linkedin-jobs-scraper

API call:

bash
curl -X POST "https://api.apify.com/v2/acts/automation-lab~linkedin-jobs-scraper/runs?token=$APIFY_API_TOKEN" \
  -H "Content-Type: application/json" \
  -d '{
    "searchQuery": "AI engineer",
    "location": "San Francisco",
    "maxItems": 25
  }'

Input fields:

FieldTypeDescription
searchQuerystringJob title or keywords (required)
locationstringCity, state, or country (optional)
maxItemsintegerMax jobs to return (default: 50)

Polling for results:

bash
# Check run status (poll every 10s)
curl "https://api.apify.com/v2/acts/automation-lab~linkedin-jobs-scraper/runs/{RUN_ID}?token=$APIFY_API_TOKEN"

# When status is SUCCEEDED, fetch results
curl "https://api.apify.com/v2/datasets/{DATASET_ID}/items?token=$APIFY_API_TOKEN"

Output fields per job:

  • title — Job title
  • companyName — Company name
  • companyLinkedinUrl — Company LinkedIn page
  • companyLogo — Logo URL
  • location — City, state
  • salary — Salary text (when available)
  • employmentType — Full-time, Part-time, Contract, etc.
  • seniorityLevel — Entry, Mid-Senior, Director, Executive, etc.
  • jobFunction — Engineering, Sales, Marketing, etc.
  • industries — Industry classification
  • descriptionText — Full job description (plain text)
  • descriptionHtml — Full job description (HTML)
  • applicantsCount — Number of applicants
  • postedAt — When posted (e.g., "6 days ago")
  • url — Direct link to the LinkedIn job posting
  • applyUrl — Direct apply URL
Indeed — borderline/indeed-scraper

API call:

bash
curl -X POST "https://api.apify.com/v2/acts/borderline~indeed-scraper/runs?token=$APIFY_API_TOKEN" \
  -H "Content-Type: application/json" \
  -d '{
    "query": "AI engineer",
    "location": "San Francisco, CA",
    "country": "us",
    "maxResults": 25
  }'

Input fields:

FieldTypeDescription
querystringJob title or keywords (required)
locationstringCity and state (optional)
countrystringLowercase 2-letter country code (required). Common: us, uk, ca, de, fr, in, au
maxResultsintegerMax jobs to return

Important: The country field is required for Indeed. If the user doesn't specify a country, default to us. Use lowercase 2-letter codes only.

Output fields per job:

  • title — Job title
  • companyName — Company name
  • companyDescription — Company description
  • companyNumEmployees — Company size
  • companyRevenue — Company revenue range
  • companyUrl — Company Indeed page
  • location — Object with city, postalCode, country, formattedAddressShort, latitude, longitude, streetAddress
  • salary — Object with salaryCurrency, salaryMin, salaryMax, salaryText, salaryType (hourly/yearly)
  • descriptionText — Full job description (plain text)
  • descriptionHtml — Full job description (HTML)
  • datePublished — Posted date (YYYY-MM-DD)
  • age — Human-readable age ("24 days ago")
  • expired — Whether job is still active
  • isRemote — Remote flag
  • jobType — Employment type
  • jobUrl — Direct Indeed job URL
  • applyUrl — Direct apply URL
  • rating — Company rating and review count
  • emails — Contact emails (when available)
  • attributes — Job attributes list (benefits, requirements, etc.)
  • hiringDemand — Urgent hire / high volume hiring flags
Show full SKILL.md (428 more words)Show less
Phase 2: Filter & Deduplicate
Recency Filtering

If the user asked for recent jobs, filter results by date:

  • LinkedIn: Use postedAt field (e.g., "6 days ago") — parse the text to determine recency.
  • Indeed: Use datePublished field (YYYY-MM-DD) — compare against today's date.

Remove jobs older than what the user requested. If no recency filter specified, still remove jobs older than 30 days by default to avoid stale data.

Deduplication (when using both sources)

When searching both LinkedIn and Indeed, the same job may appear on both platforms. Deduplicate by matching:

  1. Normalize company name (lowercase, strip "Inc", "LLC", "Corp", etc.)
  2. Normalize job title (lowercase)
  3. If company name AND job title match, keep the result with richer data (prefer Indeed for salary data, LinkedIn for seniority level)
Phase 3: Present Results

Show results as a summary table:

Source: LinkedIn + Indeed (deduplicated)
Jobs found: {count}
Location: {location}
Search: "{query}"

| # | Title | Company | Location | Salary | Posted | Source |
|---|-------|---------|----------|--------|--------|--------|
| 1 | AI Engineer | Stripe | SF, CA | $200K-$300K | 3 days ago | LinkedIn |
| 2 | ML Engineer | Meta | Menlo Park, CA | $58.65/hr | Mar 14 | Indeed |
| ... |

After the table:

  • Note how many were filtered for recency
  • Note how many duplicates were removed
  • Provide the total cost of the search

If the user wants more detail on a specific job, show the full description.

Phase 4: Export (Optional)

If the user wants to save results:

{search-term}-jobs-{YYYY-MM-DD}.csv

CSV columns:

title, company, location, salary, employment_type, seniority_level, posted_date, job_url, apply_url, description, source

Normalize fields across sources so the CSV has a consistent schema regardless of whether the job came from LinkedIn or Indeed.


Cost Estimates

SearchLinkedIn OnlyIndeed OnlyBoth Sources
25 jobs~$0.05~$0.10~$0.15
50 jobs~$0.10~$0.20~$0.30
100 jobs~$0.20~$0.40~$0.60

LinkedIn is cheaper per job. Indeed returns richer data per job. Both together give the most complete picture.


Common Use Cases

Hiring signal detection: "Find companies hiring AI engineers in SF" → Search both sources, group by company, rank by number of open roles. Companies with 5+ AI roles are actively building.

Competitive intelligence: "What is Anthropic hiring for?" → Search LinkedIn with searchQuery: "Anthropic". Shows their open roles, team growth, and strategic priorities.

Salary research: "What do ML engineers make in NYC?" → Search Indeed (richer salary data). Filter to NYC, aggregate salary ranges.

GTM prospecting: "Find companies hiring for VP of Sales" → These companies are scaling their sales org and may need sales tools. Export the company list for outreach.


Error Handling

ErrorFix
APIFY_API_TOKEN not setAsk user to add it to .env
Indeed: Missing country inputAdd country field with lowercase 2-letter code (default: us)
LinkedIn: 0 resultsBroaden search query or remove location filter
Indeed: 999 results returnedThe maxResults field may not cap results. Filter client-side.
Apify run fails or times outRetry once. If still fails, try the other source.
Stale results (30+ days old)Apply recency filter. Warn user about data freshness.

© gooseworks-ai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in skills/lead-generation/capabilities/job-scraper of gooseworks-ai/goose-skills.

  • SKILL.md
  • skill.meta.json

Open the folder on GitHubat commit c650c6d

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in gooseworks-ai/goose-skills, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Job Scraper next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Job Scraper compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Job Scraper this skillgooseworks-ai/goose-skills1.2k1 repos~2.6kAutomated safety check: NotesMIT
Linkedin Thread Monitorsergebulaev/linkedin-skills4.4k1 repos~1.4kAutomated safety check: PassMIT
Apify Google Maps Leadsapify/awesome-skills266—~3.8kAutomated safety check: PassApache-2.0
Apify Job Boardsapify/awesome-skills266—~3.5kAutomated safety check: PassApache-2.0
Coffee ChatLeoYeAI/openclaw-master-skills2.2k—~6.6kAutomated safety check: PassMIT
Apify Jobs Dataapify/awesome-skills266—~5.5kAutomated safety check: PassApache-2.0

Similar skills

  • Linkedin Thread Monitor

    sergebulaev/linkedin-skills

    Track which of your LinkedIn comments earned author replies.

    4.4k GitHub starsUsed in 1 repo~1.4k tokens
    Data & AnalyticsAuto-check passed
  • Apify Google Maps Leads

    apify/awesome-skills

    Official

    Build a local-business lead database from Google Maps in one Apify pipeline: search by target audience + geography, enrich each place with company contacts from its website, leads enrichment (names…

    266 GitHub stars~3.8k tokensUpdated 18 days ago
    Data & AnalyticsAuto-check passed
  • Apify Job Boards

    apify/awesome-skills

    Official

    Pull job postings from many boards in one run and prepare one validated, deduplicated table.

    266 GitHub stars~3.5k tokensUpdated 18 days ago
    Data & AnalyticsAuto-check passed
  • Coffee Chat

    LeoYeAI/openclaw-master-skills

    Generate a personalized coffee chat playbook for networking conversations.

    2.2k GitHub stars~6.6k tokensUpdated 2 mo ago
    Data & AnalyticsAuto-check passed
  • Apify Jobs Data

    apify/awesome-skills

    Official

    Extract clean, de-noised job-posting data from LinkedIn, Indeed, Glassdoor, and 20+ boards in one Apify run — deduplicated across boards, with likely ghost jobs and reposts flagged (heuristic, not…

    266 GitHub stars~5.5k tokensUpdated 18 days ago
    Data & AnalyticsAuto-check passed
  • Official

    Scrapes public data from social, maps, search and review platforms by choosing from about a hundred Apify Actors and running them through the Apify CLI.

    2.4k GitHub starsUsed in 2 repos~1.4k tokens
    Data & AnalyticsAuto-check: notes

More from gooseworks-ai/goose-skills

All 273 skills in this repo
  • Reddit Post Finder

    gooseworks-ai/goose-skills

    Scrape and search Reddit posts using Apify. An agent skill from gooseworks-ai/goose-skills.

    1.2k GitHub starsUsed in 1 repo~1.2k tokens
    Auto-check passed
  • Create Image Fal

    gooseworks-ai/goose-skills

    Generate or edit an image via any FAL image model (nano-banana edit, gpt-image, flux, ...), ROUTED THROUGH THE fal-proxy so it bills the Ads agent.

    1.2k GitHub stars~1.3k tokensUpdated yesterday
    Auto-check passed
  • Render Hook Replacement

    gooseworks-ai/goose-skills

    Replace an existing video's opening with a supplied clip or free kinetic text hook while retaining and verifying every original body frame, audio, captions and ending.

    1.2k GitHub stars~2.3k tokensUpdated yesterday
    Auto-check passed
  • Blog Feed Monitor

    gooseworks-ai/goose-skills

    Scrape blog posts via RSS feeds (free, no API key) with Apify fallback for JS-heavy sites.

    1.2k GitHub starsUsed in 1 repo~578 tokens
    Auto-check passed
  • Competitor Post Engagers

    gooseworks-ai/goose-skills

    Find leads by scraping engagers from a competitor's top LinkedIn posts.

    1.2k GitHub starsUsed in 1 repo~1.8k tokens
    Auto-check: notes
  • Render Chatgpt Chat

    gooseworks-ai/goose-skills

    Assemble a ChatGPT chat-reveal video ad from a thread + timeline JSON — one continuous Playwright recording of a ChatGPT mobile chat (user types with the iOS keyboard up → taps send → keyboard…

    1.2k GitHub stars~2.3k tokensUpdated yesterday
    Auto-check passed

Works with

Questions about Job Scraper

What does Job Scraper do?

Search for job postings across LinkedIn and Indeed. An agent skill from gooseworks-ai/goose-skills. Job Scraper is an agent skill from gooseworks-ai/goose-skills. Search for job postings across LinkedIn and Indeed.

When should I use Job Scraper?

Job Scraper fits situations like: users want to find open roles; monitor hiring signals; identify companies hiring for specific positions; research competitor hiring activity.

How do I install Job Scraper in Claude Code?

Run `npx skills add gooseworks-ai/goose-skills --skill job-scraper -a claude-code`. Or copy the skill folder (skills/lead-generation/capabilities/job-scraper in gooseworks-ai/goose-skills) into .claude/skills/job-scraper in your project. Claude Code loads it when a task matches its description.

How do I install Job Scraper in Codex?

Run `npx skills add gooseworks-ai/goose-skills --skill job-scraper -a codex`. Or copy the skill folder (skills/lead-generation/capabilities/job-scraper in gooseworks-ai/goose-skills) into .agents/skills/job-scraper in your project. Codex loads it when a task matches its description.

Can I use Job Scraper in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add gooseworks-ai/goose-skills --skill job-scraper -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/job-scraper, .gemini/skills/job-scraper, .github/skills/job-scraper and .opencode/skills/job-scraper in your project.

What does Job Scraper need to run?

Going by SKILL.md and its folder, Job Scraper needs the command-line tools its instructions call (curl and stripe) and credentials named APIFY_API_TOKEN. Our summary lists: A credential in APIFY_API_TOKEN.

Does Job Scraper access the network?

SKILL.md names 1 domain. In commands or code: api.apify.com; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.

Is Job Scraper safe to install?

Our automated static check of SKILL.md found notes only (mentions a .env file), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.

What licence does Job Scraper use?

Job Scraper is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Job Scraper use?

About 2.6k tokens (SKILL.md is roughly 10k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Job Scraper?

Skills that share tags, products or a category with Job Scraper: Linkedin Thread Monitor (sergebulaev/linkedin-skills, 4.4k stars), Apify Google Maps Leads (apify/awesome-skills, 266 stars), Apify Job Boards (apify/awesome-skills, 266 stars) and Coffee Chat (LeoYeAI/openclaw-master-skills, 2.2k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Job Scraper?

gooseworks-ai (a GitHub organization) maintains it in gooseworks-ai/goose-skills, which has 1,240 GitHub stars. The repository holds 273 skills in this directory. The repository was last updated on October 8, 2026.

Source: gooseworks-ai/goose-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.