Agent skill

Podcast Transcript Fetcher

by Varnan-Tech in Varnan-Tech/opendirectory

A skill your agent uses when fetching, searching, or analyzing transcripts from Lenny's Podcast, Dwarkesh Podcast, Cheeky Pint, 20VC, or A16z Podcast.

MITAuto-check: notesMedia & Creative

Install Podcast Transcript Fetcher

skills CLI
$ npx skills add Varnan-Tech/opendirectory --skill podcast-transcript-fetcher -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Varnan-Tech/opendirectory podcast-transcript-fetcher --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Varnan-Tech/opendirectory.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/podcast-transcript-fetcher .claude/skills/podcast-transcript-fetcher && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
podcast-transcript-fetcher
GitHub stars
674
Used in
1 other repo
Token cost
~1.9k tokens
SKILL.md length
616 words
Files
7 (incl. scripts, references)
Skills in repo
61
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses when fetching, searching, or analyzing transcripts from Lenny's Podcast, Dwarkesh Podcast, Cheeky Pint, 20VC, or A16z Podcast.

  • Works in 3 steps: Install Dependencies → Get a Transcript → Analyze with AI
  • Analyzing transcripts from Lennys Podcast
  • SKILL.md covers Quick Reference, Supported Podcasts, Implementation and Supported Workflows, plus 4 more sections
  • Runs Python scripts from its folder; calls python, pip and winget; reaches console.groq.com and taddy.org; needs TADDY_API_KEY and GROQ_API_KEY

What it does

Podcast Transcript Fetcher is an agent skill from Varnan-Tech/opendirectory. Use when fetching, searching, or analyzing transcripts from Lenny's Podcast, Dwarkesh Podcast, Cheeky Pint, 20VC, or A16z Podcast. Tier 2 (RSS+Groq Whisper) is the recommended approach -- fast, free, and most reliable. Also use when asked to "get transcript", "find episode", "summarize podcast", or "search podcast content". Do not use for general web scraping or non-podcast audio transcription.

Its SKILL.md is about 1.9k tokens, which your agent loads only when the skill is triggered. The skill folder holds 8 other files, including scripts and reference files (for example `README.md`, `package.json` and `references/podcasts.md`).

It sits in Media & Creative, covering Podcasting, Transcription and Web scraping. It works with Substack. The repository describes itself as: AI Agent Skills built for Founders who hate Marketing. The licence is MIT.

When your agent uses it

  • Analyzing transcripts from Lennys Podcast
  • Dwarkesh Podcast
  • Asked to get transcript
  • Summarize podcast

Example prompts

  • “get transcript”
  • “find episode”
  • “summarize podcast”
  • “/podcast-transcript-fetcher”

Requirements

  • Python 3
  • A credential in GROQ_API_KEY
  • A credential in TADDY_API_KEY

Workflow steps

3 steps, taken from the step headings in SKILL.md.

  1. Install Dependencies
  2. Get a Transcript
  3. Analyze with AI

What it can do on your machine

Read from SKILL.md and the folder at commit 62e437a. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 2 files in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python
    • pip
    • winget
    • brew

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • console.groq.com
    • taddy.org

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • TADDY_API_KEY
    • GROQ_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Podcast Transcript Fetcher loads about 1.9k tokens when it runs, and up to ~3.4k if it reads all its reference files. Until then it costs about 106 tokens; SKILL.md has 616 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~106
When it runs · the whole SKILL.md, loaded when a task matches
~1.9k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~3.4k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NoteMentions a .env fileSKILL.md:190
    ys**: Set `GROQ_API_KEY` in your env or `.env` file

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from Varnan-Tech/opendirectory at commit 62e437a, republished under its MIT licence (© Varnan-Tech). 616 words, ~1,925 tokens.

Download SKILL.mdSave it as .claude/skills/podcast-transcript-fetcher/SKILL.md (or your agent's skills folder). This skill also uses 6 other files; get the full folder from GitHub.
name
podcast-transcript-fetcher
description
Use when fetching, searching, or analyzing transcripts from Lenny's Podcast, Dwarkesh Podcast, Cheeky Pint, 20VC, or A16z Podcast. Tier 2 (RSS+Groq Whisper) is the recommended approach -- fast, free, and most reliable. Also use when asked to "get transcript", "find episode", "summarize podcast", or "search podcast content". Do not use for general web scraping or non-podcast audio transcription.
author
farizanjum
version
1.1.0

Podcast Transcript Fetcher

Fetch transcripts from 5 supported podcasts. Tier 2 (RSS+Groq Whisper) is the recommended approach -- fast, free, and the most reliable across all podcasts. Tier 1 free sources are best-effort (limited availability). Tier 3 Taddy API is the premium/commercial option.

Quick Reference

bash
# Get latest episode transcript (auto-detects best method)
python scripts/get_transcript.py "Lenny's Podcast" --latest

# Search by episode title or number
python scripts/get_transcript.py 20vc --episode "Marc Andreessen"
python scripts/get_transcript.py dwarkesh --episode 15

# Force specific method
python scripts/get_transcript.py "cheeky pint" --latest --method whisper
python scripts/get_transcript.py a16z --latest --method taddy

# Save to file
python scripts/get_transcript.py lennys --latest --output transcript.md

# List all supported podcasts
python scripts/get_transcript.py --list-podcasts

Supported Podcasts

PodcastTier 1 (best-effort)Tier 2 RSS+Whisper [RECOMMENDED]Tier 3 Taddy (premium)
Lenny's PodcastGitHub archive (269 transcripts)✅ Substack RSS✅ Covered
Dwarkesh PodcastWebsite scrape + Substack PDF✅ Substack RSS✅ Covered
Cheeky Pint(none)✅ Transistor.fm RSS✅ Covered
20VCSubstack PDF✅ Libsyn RSS✅ Covered
A16z PodcastWebsite scrape✅ Simplecast RSS✅ Covered

Implementation

1. Install Dependencies
bash
# Core (always required)
pip install requests

# Cloud transcription (recommended — fast, free tier)
pip install groq
export GROQ_API_KEY="your-key"  # Get at https://console.groq.com

# Local transcription (free, needs ~5GB RAM)
pip install faster-whisper

# Audio compression (for Groq's 25 MB limit — Windows: winget/scoop)
#   winget install ffmpeg  or  scoop install ffmpeg

# Taddy API (commercial, optional)
export TADDY_API_KEY="your-key"  # Get at https://taddy.org
2. Get a Transcript

The script auto-selects the best method. Tier 2 is the default recommendation:

Tier 1 → Tier 2 (RECOMMENDED) → Tier 3
(best-effort)  (Whisper)  (Taddy API premium)

Tier 1: Free direct sources (best-effort, limited availability)

  • Lenny's: Clones ChatPRD/lennys-podcast-transcripts and searches by title
  • Dwarkesh: Substack PDF scrape
  • 20VC: Substack PDF scrape
  • A16z: Website scrape
  • Cheeky Pint: No Tier 1 sources available
  • Note: Tier 1 sources are best-effort and limited. Tier 2 (RSS+Whisper) is the recommended approach.

Tier 2: RSS + Whisper transcription [RECOMMENDED]

  • Downloads MP3 from podcast RSS feed
  • Compresses if >25 MB (ffmpeg)
  • Transcribes via Groq Whisper API (free tier, ~10s per hour of audio)
  • Fast, free, and works for every podcast in the registry
  • Default recommendation for all use cases

Tier 3: Taddy API (commercial/premium)

  • Requires TADDY_API_KEY ($75/mo+)
  • Use for large-scale or production transcript needs
  • Covers all 5 podcasts with auto-transcription
3. Analyze with AI

Once you have a transcript, pipe it to the agent for analysis:

markdown
I have this transcript from [podcast]. Can you:
1. Summarize the key arguments
2. Extract 3 actionable insights
3. Identify any controversial claims
4. Compare with [other podcast] on the same topic

Supported Workflows

Single Episode
ScenarioCommand
Latest episodeget_transcript.py "Lenny's Podcast" --latest
Specific episode by titleget_transcript.py 20vc --episode "Sam Altman"
Episode by numberget_transcript.py dwarkesh --episode 42
Force Whisper transcription (Tier 2, recommended)get_transcript.py a16z --latest --method whisper
Force Taddy API (premium)get_transcript.py lennys --latest --method taddy
Save to Markdownget_transcript.py cheeky-pint --latest --output episode.md
JSON outputget_transcript.py dwarkesh --latest --json
Cross-Podcast Search & Batch
ScenarioCommand
Search all podcasts by keywordget_transcript.py --search "Marc Andreessen"
Search by guest nameget_transcript.py --guest "Sam Altman"
Search within one podcastget_transcript.py "Lenny's Podcast" --search "vibe coding"
Batch-transcribe last N episodesget_transcript.py "Dwarkesh Podcast" --last 5
Search + transcribe top matchesget_transcript.py --search "AI safety" --transcribe
Pipeline with custom countget_transcript.py --search "scaling laws" --transcribe --transcribe-count 5
Filtered search pipelineget_transcript.py "A16z Podcast" --search "crypto" --transcribe
Show full SKILL.md (239 more words)Show less
Output Structure

Batch transcription saves to output/ with per-podcast subdirectories:

output/dwarkesh-podcast/Dwarkesh Podcast_2024-01-15_agi-is-still-30-years-away.md
output/20vc/20 Minutes VC (20VC)_2024-03-10_funding-round-analysis.md

Each file includes a YAML frontmatter header:

yaml
---
podcast: Dwarkesh Podcast
episode: AGI is still 30 years away
date: 2024-01-15
url: https://...
source: whisper
---

Podcast Registry

The registry at scripts/podcasts.json maps each podcast to its RSS feeds, transcript sources, and API endpoints. To add new podcasts:

json
{
  "id": "new-podcast",
  "name": "New Podcast",
  "rss": "https://example.com/feed.xml",
  "transcript_sources": {
    "primary": {"type": "website_scrape", "url": "https://example.com"}
  }
}

Troubleshooting

ProblemSolution
"No transcript found"Tier 2 (RSS+Whisper) is the recommended approach. If auto mode fails, try --method whisper to force it.
RSS fetch failsRSS feeds may change; check scripts/podcasts.json for current URLs
Audio download slowLarge MP3s can take minutes on slow connections
Groq rate limitedWait or switch to local faster-whisper
Taddy not returning transcriptsSome episodes lack transcripts; try --method whisper
Podcast not in registryAdd it to scripts/podcasts.json
Unicode error on WindowsFixed: script auto-reconfigures stdout to UTF-8; saved files use UTF-8 encoding
Audio > 25 MB for GroqInstall ffmpeg: winget install ffmpeg (Windows) or brew install ffmpeg (macOS)

RSS Feed Status (as of 2026-06)

PodcastOld Feed (broken)Current Feed
Cheeky Pintfeeds.transistor.fm/the-cheeky-pint (404)feeds.transistor.fm/cheeky-pint-with-john-collison
20VCfeeds.simplecast.com/3GxrMqOd (404)feeds.libsyn.com/61840/rss
A16zfeeds.simplecast.com/0cJfpoz2 (404)feeds.simplecast.com/JGE3yC0V

Common Mistakes

  • Forgetting API keys: Set GROQ_API_KEY in your env or .env file
  • Relying on Tier 1 free sources: Tier 1 is best-effort and limited. Always fall back to Tier 2 (RSS+Whisper) which is the recommended method.
  • Not cloning the Lenny's repo first: The GitHub archive must be cloned locally for Tier 1 to work
  • Using --method taddy without TADDY_API_KEY: Falls through silently; set the key or use auto mode

© Varnan-Tech, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 6 other files (scripts, references) in skills/podcast-transcript-fetcher of Varnan-Tech/opendirectory.

  • SKILL.md
  • .env.example
  • README.md
  • package.json
  • references/podcasts.md
  • scripts/get_transcript.py
  • scripts/podcasts.json

Open the folder on GitHubat commit 62e437a

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in Varnan-Tech/opendirectory, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Podcast Transcript Fetcher next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Podcast Transcript Fetcher compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Podcast Transcript Fetcher this skillVarnan-Tech/opendirectory6741 repos~1.9kAutomated safety check: NotesMIT
Podcast Publishing Assistantswyxio/skills175—~1.4kAutomated safety check: PassMIT
Voxtype Installpchalasani/claude-code-tools2k—~857Automated safety check: NotesMIT
Summarizetrpc-group/trpc-agent-go1.8k22 repos~552Automated safety check: PassApache-2.0
Speech To TextNoizAI/skills526—~917Automated safety check: PassNone
VideoiBigQiang/feedgrab614—~1.6kAutomated safety check: PassMIT

Similar skills

  • Transcribe long-form audio, YouTube videos, podcasts, interviews, or panels; summarize them; extract chapter markers; and draft publishing assets like titles, YouTube descriptions, show notes, and…

    175 GitHub stars~1.4k tokensUpdated 3 days ago
    Media & CreativeAuto-check passed
  • Voxtype Install

    pchalasani/claude-code-tools

    Guide the user through installing, configuring, and launching voxtype — local on-device voice dictation (speech-to-text that types wherever the cursor is).

    2k GitHub stars~857 tokensUpdated 3 days ago
    Media & CreativeAuto-check: notes
  • Summarize

    trpc-group/trpc-agent-go

    Summarize or extract text/transcripts from URLs, podcasts, and local files (great fallback for “transcribe this YouTube/video”).

    1.8k GitHub starsUsed in 22 repos~552 tokens
    Media & CreativeAuto-check passed
  • Speech To Text

    NoizAI/skills

    A skill your agent uses whenever the user wants to transcribe audio to text, convert speech to text, or get a transcript from an audio or video file.

    526 GitHub stars~917 tokensUpdated 10 days ago
    Media & CreativeAuto-check passed
  • Video

    iBigQiang/feedgrab

    Video & Podcast Digest — send a video/podcast link, get full transcript + structured summary.

    614 GitHub stars~1.6k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Spec Riffrec Feedback Analysis

    leo-kuang-ai/spec-first

    Analyze explicit Riffrec product-feedback captures, including riffrec-.zip, the Riffrec session.json + events.json + recording.webm + voice.webm bundle, or media/notes the user identifies as a…

    107 GitHub stars~1.4k tokensUpdated 14 days ago
    Media & CreativeAuto-check passed

More from Varnan-Tech/opendirectory

All 61 skills in this repo
  • Graphic Ebook

    Varnan-Tech/opendirectory

    Creates professionally designed B2B SaaS e-books in HTML + CSS, exported as print-ready PDF.

    674 GitHub stars~5k tokensUpdated 1 mo ago
    Auto-check passed
  • Where Your Customer Lives

    Varnan-Tech/opendirectory

    Given a product utility and ICP, researches the internet to find the specific channels.

    674 GitHub starsUsed in 1 repo~4.8k tokens
    Auto-check passed
  • Docs From Code

    Varnan-Tech/opendirectory

    Generates and updates README.md and API reference docs by reading your codebase's functions, routes, types, schemas, and architecture.

    674 GitHub stars~1.8k tokensUpdated 1 mo ago
    Auto-check passed
  • Graphic Chart

    Varnan-Tech/opendirectory

    Generates data visualization charts (bar, line, area, pie, doughnut, scatter, radar, treemap) as PNG using Apache ECharts v6.

    674 GitHub stars~2.9k tokensUpdated 1 mo ago
    Auto-check passed
  • Graphic Gif

    Varnan-Tech/opendirectory

    Creates animated looping GIFs from CSS animations (default) or AI image-to-video.

    674 GitHub stars~3k tokensUpdated 1 mo ago
    Auto-check passed
  • Map Your Market

    Varnan-Tech/opendirectory

    Given a product description, category keywords, or competitor names (any combination), searches Reddit, Hacker News, GitHub Issues, G2, and Google Trends for the real pains your market experiences…

    674 GitHub stars~4.3k tokensUpdated 1 mo ago
    Auto-check passed

Works with

Questions about Podcast Transcript Fetcher

What does Podcast Transcript Fetcher do?

A skill your agent uses when fetching, searching, or analyzing transcripts from Lenny's Podcast, Dwarkesh Podcast, Cheeky Pint, 20VC, or A16z Podcast. Podcast Transcript Fetcher is an agent skill from Varnan-Tech/opendirectory. Use when fetching, searching, or analyzing transcripts from Lenny's Podcast, Dwarkesh Podcast, Cheeky Pint, 20VC, or A16z Podcast.

When should I use Podcast Transcript Fetcher?

Podcast Transcript Fetcher fits situations like: analyzing transcripts from Lennys Podcast; dwarkesh Podcast; asked to get transcript; summarize podcast.

How do I install Podcast Transcript Fetcher in Claude Code?

Run `npx skills add Varnan-Tech/opendirectory --skill podcast-transcript-fetcher -a claude-code`. Or copy the skill folder (skills/podcast-transcript-fetcher in Varnan-Tech/opendirectory) into .claude/skills/podcast-transcript-fetcher in your project. Claude Code loads it when a task matches its description.

How do I install Podcast Transcript Fetcher in Codex?

Run `npx skills add Varnan-Tech/opendirectory --skill podcast-transcript-fetcher -a codex`. Or copy the skill folder (skills/podcast-transcript-fetcher in Varnan-Tech/opendirectory) into .agents/skills/podcast-transcript-fetcher in your project. Codex loads it when a task matches its description.

Can I use Podcast Transcript Fetcher in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Varnan-Tech/opendirectory --skill podcast-transcript-fetcher -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/podcast-transcript-fetcher, .gemini/skills/podcast-transcript-fetcher, .github/skills/podcast-transcript-fetcher and .opencode/skills/podcast-transcript-fetcher in your project.

What does Podcast Transcript Fetcher need to run?

Going by SKILL.md and its folder, Podcast Transcript Fetcher needs Python for the scripts in its folder, the command-line tools its instructions call (python, pip, winget and brew) and credentials named TADDY_API_KEY and GROQ_API_KEY. Our summary lists: Python 3; A credential in GROQ_API_KEY; A credential in TADDY_API_KEY.

Does Podcast Transcript Fetcher access the network?

SKILL.md names 2 domains. In commands or code: console.groq.com and taddy.org; the agent is likely to contact these when it follows the instructions. This is read from the text; nothing was executed.

Is Podcast Transcript Fetcher safe to install?

Our automated static check of SKILL.md found notes only (mentions a .env file), nothing it rates as a warning. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Podcast Transcript Fetcher use?

Podcast Transcript Fetcher is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Podcast Transcript Fetcher use?

About 1.9k tokens (SKILL.md is roughly 7.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 1.4k tokens, read only when the agent opens those files.

What are the alternatives to Podcast Transcript Fetcher?

Skills that share tags, products or a category with Podcast Transcript Fetcher: Podcast Publishing Assistant (swyxio/skills, 175 stars), Voxtype Install (pchalasani/claude-code-tools, 2k stars), Summarize (trpc-group/trpc-agent-go, 1.8k stars) and Speech To Text (NoizAI/skills, 526 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Podcast Transcript Fetcher?

Varnan-Tech (a GitHub organization) maintains it in Varnan-Tech/opendirectory, which has 674 GitHub stars. The repository holds 61 skills in this directory. The repository was last updated on August 16, 2026.

Source: Varnan-Tech/opendirectory on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.