Agent skill

Yc Jobs Scraper

by Varnan-Tech in Varnan-Tech/opendirectory

Scrape daily job listings from YCombinator's Workatastartup platform without duplicates.

MITAuto-check passedData & Analytics

Install Yc Jobs Scraper

skills CLI
$ npx skills add Varnan-Tech/opendirectory --skill yc-jobs-scraper -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Varnan-Tech/opendirectory yc-jobs-scraper --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Varnan-Tech/opendirectory.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/yc-intent-radar-skill/yc-jobs-scraper .claude/skills/yc-jobs-scraper && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
yc-jobs-scraper
GitHub stars
674
Token cost
~772 tokens
SKILL.md length
290 words
Files
8 (incl. scripts)
Skills in repo
61
Repo updated
First seen
Licence
MIT

At a glance

Scrape daily job listings from YCombinator's Workatastartup platform without duplicates.

  • Works in 4 steps: First-Time Setup → Authentication (Manual Step) → Running the Daily Scraper → …
  • Asked to scrape YC jobs
  • SKILL.md covers Architecture and Workflows
  • Runs JavaScript scripts from its folder; calls node, npm and npx

What it does

Yc Jobs Scraper is an agent skill from Varnan-Tech/opendirectory. Scrape daily job listings from YCombinator's Workatastartup platform without duplicates. Use this skill when asked to scrape YC jobs, update the YC companies list, or retrieve the latest startup jobs. It handles authentication, extracts company slugs via Inertia.js JSON payloads, falls back to public YC job pages when necessary, and maintains a local SQLite database to track historical jobs and prevent duplicates.

Its SKILL.md is about 770 tokens, which your agent loads only when the skill is triggered. The skill folder holds 8 other files, including scripts (for example `scripts/auth.js`, `scripts/db.js` and `scripts/export_radar_candidates.js`).

It sits in Data & Analytics, covering Web scraping. It works with SQLite. The repository describes itself as: AI Agent Skills built for Founders who hate Marketing. The licence is MIT.

When your agent uses it

  • Asked to scrape YC jobs
  • Update the YC companies list
  • Retrieve the latest startup jobs

Example prompts

  • “/yc-jobs-scraper”

Requirements

  • Node.js

Workflow steps

4 steps, taken from the step headings in SKILL.md.

  1. First-Time Setup
  2. Authentication (Manual Step)
  3. Running the Daily Scraper
  4. Querying the Database

What it can do on your machine

Read from SKILL.md and the folder at commit 62e437a. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 6 files in scripts/ (JavaScript), which the agent can run.

    Shell commands in SKILL.md call:

    • node
    • npm
    • npx

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use npm and npx, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Yc Jobs Scraper loads about 772 tokens when it runs. Until then it costs about 108 tokens; SKILL.md has 290 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~108
When it runs · the whole SKILL.md, loaded when a task matches
~772

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from Varnan-Tech/opendirectory at commit 62e437a, republished under its MIT licence (© Varnan-Tech). 290 words, ~772 tokens.

Download SKILL.mdSave it as .claude/skills/yc-jobs-scraper/SKILL.md (or your agent's skills folder). This skill also uses 7 other files; get the full folder from GitHub.
name
yc-jobs-scraper
description
Scrape daily job listings from YCombinator's Workatastartup platform without duplicates. Use this skill when asked to scrape YC jobs, update the YC companies list, or retrieve the latest startup jobs. It handles authentication, extracts company slugs via Inertia.js JSON payloads, falls back to public YC job pages when necessary, and maintains a local SQLite database to track historical jobs and prevent duplicates.

YC Jobs Scraper

This skill provides a robust architecture for scraping jobs from YCombinator and workatastartup.com. It is designed to run automatically, bypass login bottlenecks, and maintain state to never scrape duplicate jobs.

Architecture

The scraper uses a hybrid approach to maximize reliability and minimize bot detection:

  1. Authentication: scripts/auth.js uses Playwright to let a human log in once and saves the session to scripts/state.json.
  2. Database: scripts/db.js uses better-sqlite3 to manage scripts/jobs.db. It tracks every company_slug and job_id ever seen.
  3. Primary Extraction: scripts/scraper.js loads state.json, visits YC query URLs, and extracts company slugs from the hidden Inertia.js data-page JSON payload.
  4. Job Extraction (JSON): It then visits the authenticated company pages (/companies/[slug]) to extract jobs from the backend JSON payload to ensure we get the real job_id for accurate deduplication.
  5. Job Extraction (Fallback): If the JSON extraction fails, it falls back to parsing public HTML job cards from ycombinator.com/companies/[slug]/jobs.

Workflows

1. First-Time Setup

If this is the first time running the scraper in an environment, or if node_modules is missing:

bash
cd @path/scripts
npm install
npx playwright install
2. Authentication (Manual Step)

If scripts/state.json is missing or expired, the scraper will fail. You must instruct the human user to run the authentication script manually:

bash
cd @path/scripts
node auth.js

Tell the user a browser will open, and they must log in. Playwright will automatically save the cookies/tokens to state.json.

3. Running the Daily Scraper

To scrape for new companies and jobs:

bash
cd @path/scripts
node scraper.js

This script will output exactly how many new companies and new jobs were found. Because of jobs.db, running it multiple times consecutively will result in 0 new jobs found.

4. Querying the Database

If you need to analyze the scraped data or view the companies/jobs, you can query scripts/jobs.db directly using better-sqlite3.

Example: Count Companies

bash
cd @path/scripts
node -e "const db = require('better-sqlite3')('jobs.db'); console.log('Companies:', db.prepare('SELECT COUNT(*) as count FROM companies').get().count);"

Example: View Recent Jobs

bash
cd @path/scripts
node -e "const db = require('better-sqlite3')('jobs.db'); const jobs = db.prepare('SELECT title, company_slug, location FROM jobs ORDER BY created_at DESC LIMIT 5').all(); console.table(jobs);"

© Varnan-Tech, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 7 other files (scripts) in skills/yc-intent-radar-skill/yc-jobs-scraper of Varnan-Tech/opendirectory.

  • SKILL.md
  • .gitignore
  • scripts/auth.js
  • scripts/db.js
  • scripts/export_radar_candidates.js
  • scripts/package-lock.json
  • scripts/package.json
  • scripts/scraper.js

Open the folder on GitHubat commit 62e437a

Compare with similar skills

Yc Jobs Scraper next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Yc Jobs Scraper compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Yc Jobs Scraper this skillVarnan-Tech/opendirectory674—~772Automated safety check: PassMIT
Pp Apifymvanhorn/printing-press-library2.1k—~6.4kAutomated safety check: NotesApache-2.0
Pp Zameenmvanhorn/printing-press-library2.1k—~6.8kAutomated safety check: NotesApache-2.0
Junta Leiloeirossickn33/agentic-awesome-skills47k2 repos~1.6kAutomated safety check: PassMIT
Scrape Post EngagersOthmane-Khadri/YALC-the-GTM-operating-system318—~1.4kAutomated safety check: NotesMIT
Tmuxtrpc-group/trpc-agent-go1.9k23 repos~868Automated safety check: PassApache-2.0

Similar skills

  • Pp Apify

    mvanhorn/printing-press-library

    Every Apify platform feature, plus a local SQLite store, cross-Actor search, novelty diffing, cost-aware runs, and...

    2.1k GitHub stars~6.4k tokensUpdated yesterday
    Data & AnalyticsAuto-check: notes
  • Pp Zameen

    mvanhorn/printing-press-library

    Search Zameen.com property listings from your terminal — real filters, an offline SQLite mirror, saved-search monitoring, and price-drop alerts no scraper offers.

    2.1k GitHub stars~6.8k tokensUpdated yesterday
    DatabasesAuto-check: notes
  • Junta Leiloeiros

    sickn33/agentic-awesome-skills

    Coleta e consulta dados de leiloeiros oficiais de todas as 27 Juntas Comerciais do Brasil.

    47k GitHub starsUsed in 2 repos~1.6k tokens
    Backend & APIsAuto-check passed
  • Scrape Post Engagers

    Othmane-Khadri/YALC-the-GTM-operating-system

    Pull the audience that engaged with a LinkedIn post — likers (reactors) and commenters — via Unipile, dedupe across endpoints, and persist them as a result set ready for qualification or campaign…

    318 GitHub stars~1.4k tokensUpdated 1 mo ago
    Writing & ContentAuto-check: notes
  • Tmux

    trpc-group/trpc-agent-go

    Remote-control tmux sessions for interactive CLIs by sending keystrokes and scraping pane output.

    1.9k GitHub starsUsed in 23 repos~868 tokens
    Data & AnalyticsAuto-check passed
  • Ketch

    1broseidon/ketch

    Research skill for ketch — a fast stateless CLI for web search, OSS code search, curated library docs, page scraping, and site crawling; an optional MCP server exists for operators who want it, but…

    702 GitHub starsUsed in 1 repo~3.9k tokens
    Data & AnalyticsAuto-check passed

More from Varnan-Tech/opendirectory

All 61 skills in this repo
  • Graphic Ebook

    Varnan-Tech/opendirectory

    Creates professionally designed B2B SaaS e-books in HTML + CSS, exported as print-ready PDF.

    674 GitHub stars~5k tokensUpdated yesterday
    Auto-check passed
  • Docs From Code

    Varnan-Tech/opendirectory

    Generates and updates README.md and API reference docs by reading your codebase's functions, routes, types, schemas, and architecture.

    674 GitHub stars~1.8k tokensUpdated yesterday
    Auto-check passed
  • Graphic Chart

    Varnan-Tech/opendirectory

    Generates data visualization charts (bar, line, area, pie, doughnut, scatter, radar, treemap) as PNG using Apache ECharts v6.

    674 GitHub stars~2.9k tokensUpdated yesterday
    Auto-check passed
  • Graphic Gif

    Varnan-Tech/opendirectory

    Creates animated looping GIFs from CSS animations (default) or AI image-to-video.

    674 GitHub stars~3k tokensUpdated yesterday
    Auto-check passed
  • Map Your Market

    Varnan-Tech/opendirectory

    Given a product description, category keywords, or competitor names (any combination), searches Reddit, Hacker News, GitHub Issues, G2, and Google Trends for the real pains your market experiences…

    674 GitHub stars~4.3k tokensUpdated yesterday
    Auto-check passed
  • Newsletter Digest

    Varnan-Tech/opendirectory

    Aggregates RSS feeds from the past week, synthesizes the top stories using Gemini, and publishes a newsletter digest to Ghost CMS.

    674 GitHub stars~1.9k tokensUpdated yesterday
    Auto-check passed

Works with

Questions about Yc Jobs Scraper

What does Yc Jobs Scraper do?

Scrape daily job listings from YCombinator's Workatastartup platform without duplicates. Yc Jobs Scraper is an agent skill from Varnan-Tech/opendirectory. Scrape daily job listings from YCombinator's Workatastartup platform without duplicates.

When should I use Yc Jobs Scraper?

Yc Jobs Scraper fits situations like: asked to scrape YC jobs; update the YC companies list; retrieve the latest startup jobs.

How do I install Yc Jobs Scraper in Claude Code?

Run `npx skills add Varnan-Tech/opendirectory --skill yc-jobs-scraper -a claude-code`. Or copy the skill folder (skills/yc-intent-radar-skill/yc-jobs-scraper in Varnan-Tech/opendirectory) into .claude/skills/yc-jobs-scraper in your project. Claude Code loads it when a task matches its description.

How do I install Yc Jobs Scraper in Codex?

Run `npx skills add Varnan-Tech/opendirectory --skill yc-jobs-scraper -a codex`. Or copy the skill folder (skills/yc-intent-radar-skill/yc-jobs-scraper in Varnan-Tech/opendirectory) into .agents/skills/yc-jobs-scraper in your project. Codex loads it when a task matches its description.

Can I use Yc Jobs Scraper in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Varnan-Tech/opendirectory --skill yc-jobs-scraper -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/yc-jobs-scraper, .gemini/skills/yc-jobs-scraper, .github/skills/yc-jobs-scraper and .opencode/skills/yc-jobs-scraper in your project.

What does Yc Jobs Scraper need to run?

Going by SKILL.md and its folder, Yc Jobs Scraper needs JavaScript for the scripts in its folder and the command-line tools its instructions call (node, npm and npx). Our summary lists: Node.js.

Does Yc Jobs Scraper access the network?

SKILL.md contains no URLs. Its commands use npm and npx, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Yc Jobs Scraper safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Yc Jobs Scraper use?

Yc Jobs Scraper is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Yc Jobs Scraper use?

About 772 tokens (SKILL.md is roughly 3.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Yc Jobs Scraper?

Skills that share tags, products or a category with Yc Jobs Scraper: Pp Apify (mvanhorn/printing-press-library, 2.1k stars), Pp Zameen (mvanhorn/printing-press-library, 2.1k stars), Junta Leiloeiros (sickn33/agentic-awesome-skills, 47k stars) and Scrape Post Engagers (Othmane-Khadri/YALC-the-GTM-operating-system, 318 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Yc Jobs Scraper?

Varnan-Tech (a GitHub organization) maintains it in Varnan-Tech/opendirectory, which has 674 GitHub stars. The repository holds 61 skills in this directory. The repository was last updated on August 16, 2026.

Source: Varnan-Tech/opendirectory on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.