Agent skill

Youtube Transcribe Skill

by feiskyer in feiskyer/claude-code-settings

Extract subtitles/transcripts from YouTube videos. An agent skill from feiskyer/claude-code-settings.

MITAuto-check passedMedia & Creative

Install Youtube Transcribe Skill

skills CLI
$ npx skills add feiskyer/claude-code-settings --skill youtube-transcribe-skill -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install feiskyer/claude-code-settings youtube-transcribe-skill --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/feiskyer/claude-code-settings.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/youtube-transcribe-skill .claude/skills/youtube-transcribe-skill && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
youtube-transcribe-skill
GitHub stars
1.7k
Token cost
~1.1k tokens
SKILL.md length
543 words
Files
1
Skills in repo
12
Repo updated
First seen
Licence
MIT

At a glance

Extract subtitles/transcripts from YouTube videos. An agent skill from feiskyer/claude-code-settings.

  • Works in 3 steps: Verify URL → CLI Quick Extraction (Priority Attempt) → Browser Automation (Fallback)
  • Tasks that involve Transcription
  • SKILL.md covers Step 1: Verify URL, Step 2: CLI Quick Extraction…, Step 3: Browser Automation… and Output Requirements
  • Calls yt-dlp

What it does

Youtube Transcribe Skill is an agent skill from feiskyer/claude-code-settings. Extract subtitles/transcripts from YouTube videos. Triggers: "youtube transcript", "extract subtitles", "video captions", "视频字幕", "字幕提取", "YouTube转文字", "提取字幕".

Its SKILL.md is about 1.1k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering Transcription. It works with YouTube, Chrome DevTools and Model Context Protocol. The repository describes itself as: Curated skills, sub-agents, and config templates that supercharge Claude Code — research, image gen, GitHub automation & more. The licence is MIT.

When your agent uses it

  • Tasks that involve Transcription

Example prompts

  • “youtube transcript”
  • “extract subtitles”
  • “video captions”
  • “/youtube-transcribe-skill”

Workflow steps

3 steps, taken from the step headings in SKILL.md.

  1. Verify URL
  2. CLI Quick Extraction (Priority Attempt)
  3. Browser Automation (Fallback)

What it can do on your machine

Read from SKILL.md and the folder at commit 95dab59. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • yt-dlp

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Youtube Transcribe Skill loads about 1.1k tokens when it runs. Until then it costs about 46 tokens; SKILL.md has 543 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~46
When it runs · the whole SKILL.md, loaded when a task matches
~1.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from feiskyer/claude-code-settings at commit 95dab59, republished under its MIT licence (© feiskyer). 543 words, ~1,140 tokens.

Download SKILL.mdSave it as .claude/skills/youtube-transcribe-skill/SKILL.md (or your agent's skills folder).
name
youtube-transcribe-skill
description
Extract subtitles/transcripts from YouTube videos. Triggers: "youtube transcript", "extract subtitles", "video captions", "视频字幕", "字幕提取", "YouTube转文字", "提取字幕".

YouTube Transcript Extraction

Extract subtitles/transcripts from a YouTube video URL and save them as a local file.

Input YouTube URL: $ARGUMENTS

Step 1: Verify URL

Confirm the input is a valid YouTube URL (supports youtube.com/watch?v=, youtu.be/, and youtube.com/shorts/ formats). If no URL is provided via arguments, check the conversation context for a YouTube link.

Step 2: CLI Quick Extraction (Priority Attempt)

Use command-line tools to quickly extract subtitles.

2.1 Check Tool Availability

Execute which yt-dlp.

  • If yt-dlp is found, proceed to 2.2.
  • If yt-dlp is not found, skip to Step 3.
2.2 Get Video Title
bash
yt-dlp --cookies-from-browser=chrome --get-title "[VIDEO_URL]"
  • Tip: Always add --cookies-from-browser to avoid sign-in restrictions. Default to chrome.
  • If it fails with a browser error (e.g., "Could not open Chrome"), ask the user to specify their available browser (e.g., firefox, safari, edge) and retry.
2.3 Download Subtitles
bash
yt-dlp --cookies-from-browser=chrome --write-auto-sub --write-sub --sub-lang zh-Hans,zh-Hant,en --skip-download --output "<Video Title>.%(ext)s" "[VIDEO_URL]"
2.4 Convert to Plain Text

yt-dlp saves subtitles as .vtt or .srt files. Convert the downloaded file to plain Timestamp Text format:

  1. Read the downloaded subtitle file (.vtt or .srt).
  2. Strip VTT/SRT headers, styling tags, and duplicate lines.
  3. Save as <Video Title>.txt with one Timestamp Text entry per line.
2.5 Verify Results
  • Exit code 0: Convert and save the subtitle file, then report completion.
  • Exit code non-0:
    • If error is related to browser/cookies, ask user for correct browser and retry.
    • If other errors (e.g., video unavailable), proceed to Step 3.

Step 3: Browser Automation (Fallback)

When the CLI method fails or yt-dlp is missing, use Chrome DevTools MCP to extract subtitles via browser UI automation.

3.1 Check Tool Availability

Check if Chrome DevTools MCP tools are available (look for tools matching chrome__new_page or similar).

If Chrome DevTools MCP is not available and yt-dlp was not found in Step 2, stop and notify the user: "Unable to proceed. Please either install yt-dlp (for fast CLI extraction) or configure Chrome DevTools MCP (for browser automation)."

3.2 Open Video Page

Use Chrome DevTools MCP new_page to open the video URL.

Show full SKILL.md (218 more words)Show less
3.3 Analyze Page State

Use Chrome DevTools MCP take_snapshot to read the page accessibility tree.

3.4 Expand Video Description

The "Show transcript" button is usually hidden within the collapsed description area.

  1. Search the snapshot for a button labeled "...more", "...更多", or "Show more" (in the description block below the video title).
  2. Use Chrome DevTools MCP click to click that button.
3.5 Open Transcript Panel
  1. Use Chrome DevTools MCP take_snapshot to get the updated UI.
  2. Search for a button labeled "Show transcript", "显示转录稿", or "内容转文字".
  3. Use Chrome DevTools MCP click to click that button.
  4. If the button is not found, the video may not have a transcript available — notify the user and stop.
3.6 Extract Content via DOM

Directly reading the accessibility tree for long transcript lists is slow and token-heavy. Use Chrome DevTools MCP evaluate_script to run this JavaScript instead:

javascript
() => {
  const segments = document.querySelectorAll("ytd-transcript-segment-renderer");
  if (!segments.length) return "BUFFERING";
  return Array.from(segments)
    .map((seg) => {
      const time = seg.querySelector(".segment-timestamp")?.innerText.trim();
      const text = seg.querySelector(".segment-text")?.innerText.trim();
      return `${time} ${text}`;
    })
    .join("\n");
};

If it returns "BUFFERING", wait a few seconds and retry (up to 3 attempts).

3.7 Save and Cleanup
  1. Save the extracted text as <Video Title>.txt.
  2. Use Chrome DevTools MCP close_page to release resources.

Output Requirements

  • Save the subtitle file to the current working directory.
  • Filename format: <Video Title>.txt
  • File content format: Each line should be Timestamp Subtitle Text.
  • Report upon completion: file path, subtitle language, and total number of lines.

© feiskyer, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/youtube-transcribe-skill of feiskyer/claude-code-settings.

Open the folder on GitHubat commit 95dab59

Compare with similar skills

Youtube Transcribe Skill next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Youtube Transcribe Skill compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Youtube Transcribe Skill this skillfeiskyer/claude-code-settings1.7k—~1.1kAutomated safety check: PassMIT
Scenario Caption Studioscenario-labs/skills946—~3.4kAutomated safety check: PassMIT
Summarizeswarmclawai/swarmclaw689—~531Automated safety check: PassMIT
Video Transcribewendy7756/AI-Video-Transcriber3.3k—~937Automated safety check: NotesApache-2.0
Ffmpeg Skillkajisho5/ffmpeg-skill1.9k—~7.4kAutomated safety check: PassMIT
Claude Real VideoHUANGCHIHHUNGLeo/claude-real-video2.2k—~639Automated safety check: PassMIT

Similar skills

  • Scenario Caption Studio

    scenario-labs/skills

    A skill your agent uses when a video needs its spoken words on screen through Scenario via MCP: burned-in styled captions for a TikTok, Reels, or Shorts cut, ad captions for sound-off feeds, YouTube…

    946 GitHub stars~3.4k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Summarize

    swarmclawai/swarmclaw

    Summarize or extract text/transcripts from URLs, podcasts, YouTube videos, and local files using the summarize CLI.

    689 GitHub stars~531 tokensUpdated 3 mo ago
    AI & LLM EngineeringAuto-check passed
  • Video Transcribe

    wendy7756/AI-Video-Transcriber

    Transcribe and summarize a video or podcast from a URL (YouTube, TikTok, Bilibili, Apple Podcasts, SoundCloud, 30+ platforms) or from a local media/.txt file.

    3.3k GitHub stars~937 tokensUpdated 26 days ago
    Media & CreativeAuto-check: notes
  • Ffmpeg Skill

    kajisho5/ffmpeg-skill

    Edit video and audio with local FFmpeg from natural-language requests: cut, trim, join, resize/reframe (9:16, 1:1), speed change, captions and subtitles (SRT/ASS, animated, karaoke), logos and text…

    1.9k GitHub stars~7.4k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Claude Real Video

    HUANGCHIHHUNGLeo/claude-real-video

    Watch a video for the user. An agent skill from HUANGCHIHHUNGLeo/claude-real-video.

    2.2k GitHub stars~639 tokensUpdated 2 days ago
    Media & CreativeAuto-check passed
  • Video Perception

    jordanrendric/claude-video-vision

    A skill your agent uses when the user mentions a video file (.mp4, .mov, .avi, .mkv, .webm), a YouTube URL, asks to watch/analyze/review a video, or references video content in conversation

    1.4k GitHub stars~1.4k tokensUpdated 3 days ago
    Media & CreativeAuto-check passed

More from feiskyer/claude-code-settings

All 12 skills in this repo
  • Skill Creator

    feiskyer/claude-code-settings

    Create, refine, and benchmark agent skills. An agent skill from feiskyer/claude-code-settings.

    1.7k GitHub stars~7.6k tokensUpdated 13 days ago
    Auto-check passed
  • Brainstorming

    feiskyer/claude-code-settings

    Explore user intent, requirements, and design options through collaborative dialogue before implementation.

    1.7k GitHub stars~985 tokensUpdated 13 days ago
    Auto-check passed
  • Deep Research

    feiskyer/claude-code-settings

    Multi-agent research orchestration: split a research goal into parallel sub-goals, run each via headless claude -p subprocesses, aggregate results into a polished report file.

    1.7k GitHub stars~2.6k tokensUpdated 13 days ago
    Auto-check: notes
  • Codex Skill

    feiskyer/claude-code-settings

    Leverage OpenAI Codex/GPT models for autonomous code implementation, code review, and plan review.

    1.7k GitHub stars~2.7k tokensUpdated 13 days ago
    Auto-check: warnings
  • GitHub Fix Issue

    feiskyer/claude-code-settings

    Fix GitHub issues end-to-end — analysis, branch creation, implementation, testing, and PR submission.

    1.7k GitHub stars~786 tokensUpdated 13 days ago
    Auto-check passed
  • GitHub Review PR

    feiskyer/claude-code-settings

    Review GitHub pull requests with detailed, multi-perspective code analysis using parallel subagents.

    1.7k GitHub stars~9.2k tokensUpdated 13 days ago
    Auto-check passed

Questions about Youtube Transcribe Skill

What does Youtube Transcribe Skill do?

Extract subtitles/transcripts from YouTube videos. An agent skill from feiskyer/claude-code-settings. Youtube Transcribe Skill is an agent skill from feiskyer/claude-code-settings. Extract subtitles/transcripts from YouTube videos.

When should I use Youtube Transcribe Skill?

Youtube Transcribe Skill fits situations like: tasks that involve Transcription.

How do I install Youtube Transcribe Skill in Claude Code?

Run `npx skills add feiskyer/claude-code-settings --skill youtube-transcribe-skill -a claude-code`. Or copy the skill folder (skills/youtube-transcribe-skill in feiskyer/claude-code-settings) into .claude/skills/youtube-transcribe-skill in your project. Claude Code loads it when a task matches its description.

How do I install Youtube Transcribe Skill in Codex?

Run `npx skills add feiskyer/claude-code-settings --skill youtube-transcribe-skill -a codex`. Or copy the skill folder (skills/youtube-transcribe-skill in feiskyer/claude-code-settings) into .agents/skills/youtube-transcribe-skill in your project. Codex loads it when a task matches its description.

Can I use Youtube Transcribe Skill in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add feiskyer/claude-code-settings --skill youtube-transcribe-skill -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/youtube-transcribe-skill, .gemini/skills/youtube-transcribe-skill, .github/skills/youtube-transcribe-skill and .opencode/skills/youtube-transcribe-skill in your project.

What does Youtube Transcribe Skill need to run?

Going by SKILL.md and its folder, Youtube Transcribe Skill needs the command-line tools its instructions call (yt-dlp).

Does Youtube Transcribe Skill access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Youtube Transcribe Skill safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Youtube Transcribe Skill use?

Youtube Transcribe Skill is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Youtube Transcribe Skill use?

About 1.1k tokens (SKILL.md is roughly 4.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Youtube Transcribe Skill?

Skills that share tags, products or a category with Youtube Transcribe Skill: Scenario Caption Studio (scenario-labs/skills, 946 stars), Summarize (swarmclawai/swarmclaw, 689 stars), Video Transcribe (wendy7756/AI-Video-Transcriber, 3.3k stars) and Ffmpeg Skill (kajisho5/ffmpeg-skill, 1.9k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Youtube Transcribe Skill?

feiskyer (a GitHub user) maintains it in feiskyer/claude-code-settings, which has 1,657 GitHub stars. The repository holds 12 skills in this directory. The repository was last updated on September 27, 2026.

Source: feiskyer/claude-code-settings on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.