Agent skill

Video Transcript

by notque in notque/vexjoy-agent

Extract video transcripts: yt-dlp subtitles to clean paragraphs.

MITAuto-check: notesMedia & Creative

Install Video Transcript

skills CLI
$ npx skills add notque/vexjoy-agent --skill video-transcript -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install notque/vexjoy-agent video-transcript --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/notque/vexjoy-agent.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/research/video-transcript .claude/skills/video-transcript && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
video-transcript
GitHub stars
435
Token cost
~576 tokens
SKILL.md length
194 words
Files
2 (incl. scripts)
Skills in repo
61
Repo updated
First seen
Licence
MIT

At a glance

Extract video transcripts: yt-dlp subtitles to clean paragraphs.

  • Tasks that involve Video and podcast notes
  • Runs Python scripts from its folder; calls yt-dlp and python3
  • Tasks that involve Transcription

What it does

Video Transcript is an agent skill from notque/vexjoy-agent. Extract video transcripts: yt-dlp subtitles to clean paragraphs.

Its SKILL.md is about 580 tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/vtt_to_paragraph.py`).

It sits in Media & Creative, covering Video and podcast notes and Transcription. The repository describes itself as: VexJoy AI Agent with Jev Intelligent Routing - /do routes plain-English requests to the right specialist agent and gates the work with reviews, tests, and a learning loop. The licence is MIT.

When your agent uses it

  • Tasks that involve Video and podcast notes
  • Tasks that involve Transcription

Example prompts

  • “/video-transcript”

Requirements

  • Python 3
  • Pre-approved tools (allowed-tools): Bash, Read

What it can do on your machine

Read from SKILL.md and the folder at commit 5218674. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Bash
    • Read

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • yt-dlp
    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Video Transcript loads about 576 tokens when it runs. Until then it costs about 20 tokens; SKILL.md has 194 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~20
When it runs · the whole SKILL.md, loaded when a task matches
~576

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NotePre-approves every shell command (allowed-tools: Bash)SKILL.md
    allowed-tools: Bash, Read

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from notque/vexjoy-agent at commit 5218674, republished under its MIT licence (© notque). 194 words, ~576 tokens.

Download SKILL.mdSave it as .claude/skills/video-transcript/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
video-transcript
description
Extract video transcripts: yt-dlp subtitles to clean paragraphs.
allowed-tools
Bash, Read
promoted_to
video-editing
user_invocable
false
agent
python-general-engineer
routing.triggers
video transcript, youtube transcript, extract transcript, download subtitles, what does this video say, get transcript, transcribe video, transcribe this…
routing.category
research
routing.pairs_with
research

Video Transcript

Pull a video's transcript as readable paragraphs. Two paths, in order:

Path 1 — uploader subtitles (accurate, prefer when present):

bash
yt-dlp --skip-download --write-subs --sub-langs en --sub-format vtt \
  -o '<work-dir>/%(id)s' '<URL>'

Path 2 — auto-generated captions (fallback when path 1 writes no file):

bash
yt-dlp --skip-download --write-auto-subs --sub-langs en --sub-format vtt \
  -o '<work-dir>/%(id)s' '<URL>'

Then clean the VTT into paragraphs:

bash
python3 skills/research/video-transcript/scripts/vtt_to_paragraph.py <work-dir>/<id>.en.vtt

Default output is plain paragraph text with [Music]-style cues stripped and the rolling duplicates of auto-captions deduplicated. Use --timestamps for [mm:ss] markers, --keep-brackets to keep cue tags, -o FILE to write to a file.

For other languages, change --sub-langs (e.g. de, en.*). List what a video offers with yt-dlp --list-subs '<URL>'.

Error handling

Both paths write no .vtt file

Cause: video has no subtitles or captions in the requested language. Solution: run yt-dlp --list-subs '<URL>' and pick an available language; if none exist, report that and offer audio transcription via the markdown-converter skill on a downloaded audio file.

HTTP 429 / "Sign in to confirm"

Cause: platform rate-limiting the host. Solution: wait and retry with --sleep-requests 2; keep request volume low.

Cleaner output repeats lines

Cause: VTT came from a third path (e.g. translated captions) with cue formats the dedupe misses. Solution: rerun the cleaner; if repeats remain, file the sample VTT alongside a fix to vtt_to_paragraph.py.

© notque, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (scripts) in skills/research/video-transcript of notque/vexjoy-agent.

  • SKILL.md
  • scripts/vtt_to_paragraph.py

Open the folder on GitHubat commit 5218674

Compare with similar skills

Video Transcript next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Video Transcript compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Video Transcript this skillnotque/vexjoy-agent435—~576Automated safety check: NotesMIT
Video TranscriptbozhouDev/video-skills-toolkit149—~208Automated safety check: PassMIT
Gemini Yt Video Transcriptsundial-org/awesome-openclaw-skills663—~293Automated safety check: PassNone
Video Link Transcript ExtractorSpaceZephyr/creator-buddy1.6k—~498Automated safety check: PassNone
Youtube FetcherJimmySadek/youtube-fetcher-to-markdown485—~1.8kAutomated safety check: PassMIT
Video To NotesKIRVO-REPORTING/video-to-notes105—~1.5kAutomated safety check: PassMIT

Similar skills

  • Video Transcript

    bozhouDev/video-skills-toolkit

    Deprecated compatibility package. An agent skill from bozhouDev/video-skills-toolkit.

    149 GitHub stars~208 tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed
  • Gemini Yt Video Transcript

    sundial-org/awesome-openclaw-skills

    Create a verbatim transcript for a YouTube URL using Google Gemini (speaker labels, paragraph breaks; no time codes).

    663 GitHub stars~293 tokensUpdated 7 mo ago
    Media & CreativeAuto-check passed
  • Video Link Transcript Extractor

    SpaceZephyr/creator-buddy

    Extracts subtitles or a full transcript from YouTube, Xiaoyuzhou, Bilibili, Douyin and Xiaohongshu links, using speech recognition when a video has no subtitles.

    1.6k GitHub stars~498 tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Youtube Fetcher

    JimmySadek/youtube-fetcher-to-markdown

    Retrieve YouTube transcripts and subtitles, summarize or analyze what was said, or save an Obsidian-ready Markdown knowledge-base note with captions, creator metadata, chapters, language, and source…

    485 GitHub stars~1.8k tokensUpdated 1 mo ago
    Knowledge ManagementAuto-check passed
  • Video To Notes

    KIRVO-REPORTING/video-to-notes

    Use immediately for any bare YouTube or YouTube Shorts URL, youtu.be link, Bilibili or b23.tv link, or other video URL; do not ask what the user wants.

    105 GitHub stars~1.5k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Video Transcript Extractor

    Backtthefuture/video-transcript

    Pulls text out of videos and podcasts by sending the media to a Quark cloud-drive skill for transcription, then proofreads the result lightly.

    115 GitHub stars~970 tokensUpdated 8 days ago
    Media & CreativeAuto-check: notes

More from notque/vexjoy-agent

All 61 skills in this repo
  • Game Asset Generator

    notque/vexjoy-agent

    Deterministic palette/matrix pixel art (not AI). An agent skill from notque/vexjoy-agent.

    435 GitHub stars~2.3k tokensUpdated 4 days ago
    Auto-check: notes
  • PR Workflow

    notque/vexjoy-agent

    Pull request lifecycle: commit, codex review, sync, review, fix, status, cleanup, and PR mining.

    435 GitHub stars~2.8k tokensUpdated 4 days ago
    Auto-check: notes
  • Architecture Deepening

    notque/vexjoy-agent

    Improve architecture across modules by deepening interfaces.

    435 GitHub stars~3.3k tokensUpdated 4 days ago
    Auto-check: notes
  • Code Quality

    notque/vexjoy-agent

    Code quality: cleanup, linting, formatting, quality gates. An agent skill from notque/vexjoy-agent.

    435 GitHub stars~1.5k tokensUpdated 4 days ago
    Auto-check: notes
  • Codebase Analyzer

    notque/vexjoy-agent

    Statistical rule discovery from Go codebase patterns. An agent skill from notque/vexjoy-agent.

    435 GitHub stars~2k tokensUpdated 4 days ago
    Auto-check: notes
  • Comment Quality

    notque/vexjoy-agent

    Review and fix temporal references in code comments. An agent skill from notque/vexjoy-agent.

    435 GitHub stars~2k tokensUpdated 4 days ago
    Auto-check: notes

Questions about Video Transcript

What does Video Transcript do?

Extract video transcripts: yt-dlp subtitles to clean paragraphs. Video Transcript is an agent skill from notque/vexjoy-agent. Extract video transcripts: yt-dlp subtitles to clean paragraphs.

When should I use Video Transcript?

Video Transcript fits situations like: tasks that involve Video and podcast notes; tasks that involve Transcription.

How do I install Video Transcript in Claude Code?

Run `npx skills add notque/vexjoy-agent --skill video-transcript -a claude-code`. Or copy the skill folder (skills/research/video-transcript in notque/vexjoy-agent) into .claude/skills/video-transcript in your project. Claude Code loads it when a task matches its description.

How do I install Video Transcript in Codex?

Run `npx skills add notque/vexjoy-agent --skill video-transcript -a codex`. Or copy the skill folder (skills/research/video-transcript in notque/vexjoy-agent) into .agents/skills/video-transcript in your project. Codex loads it when a task matches its description.

Can I use Video Transcript in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add notque/vexjoy-agent --skill video-transcript -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/video-transcript, .gemini/skills/video-transcript, .github/skills/video-transcript and .opencode/skills/video-transcript in your project.

What does Video Transcript need to run?

Going by SKILL.md and its folder, Video Transcript needs Python for the scripts in its folder and the command-line tools its instructions call (yt-dlp and python3). Our summary lists: Python 3. Its frontmatter pre-approves these tools: Bash, Read.

Does Video Transcript access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Video Transcript safe to install?

Our automated static check of SKILL.md found notes only (pre-approves every shell command (allowed-tools: bash)), nothing it rates as a warning. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Video Transcript use?

Video Transcript is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Video Transcript use?

About 576 tokens (SKILL.md is roughly 2.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Video Transcript?

Skills that share tags, products or a category with Video Transcript: Video Transcript (bozhouDev/video-skills-toolkit, 149 stars), Gemini Yt Video Transcript (sundial-org/awesome-openclaw-skills, 663 stars), Video Link Transcript Extractor (SpaceZephyr/creator-buddy, 1.6k stars) and Youtube Fetcher (JimmySadek/youtube-fetcher-to-markdown, 485 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Video Transcript?

notque (a GitHub user) maintains it in notque/vexjoy-agent, which has 435 GitHub stars. The repository holds 61 skills in this directory. The repository was last updated on October 3, 2026.

Source: notque/vexjoy-agent on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.