Agent skill

Youtube Transcript Extractor API Skill

by browser-act in browser-act/skills

This skill helps users automatically extract YouTube video transcripts and metadata via the BrowserAct API.

MITAuto-check passedKnowledge Management

Install Youtube Transcript Extractor API Skill

skills CLI
$ npx skills add browser-act/skills --skill youtube-transcript-extractor-api-skill -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install browser-act/skills youtube-transcript-extractor-api-skill --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/browser-act/skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/solutions/video-platforms/youtube-transcript-extractor-api-skill .claude/skills/youtube-transcript-extractor-api-skill && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
youtube-transcript-extractor-api-skill
GitHub stars
6.1k
Used in
1 other repo
Token cost
~1.2k tokens
SKILL.md length
546 words
Files
2 (incl. scripts)
Skills in repo
87
Repo updated
First seen
Licence
MIT

At a glance

This skill helps users automatically extract YouTube video transcripts and metadata via the BrowserAct API.

  • Works in 5 steps: No hallucinations, ensuring stable and… → No CAPTCHA issues: No need to handle… → No IP access restrictions or geofencing:… → …
  • Tasks that involve Video and podcast notes
  • SKILL.md covers 📖 Introduction, ✨ Features, 🔑 API Key Setup and 🛠️ Input Parameters, plus 3 more sections
  • Runs Python scripts from its folder; calls python; reaches youtube.com; needs BROWSERACT_API_KEY

What it does

Youtube Transcript Extractor API Skill is an agent skill from browser-act/skills. This skill helps users automatically extract YouTube video transcripts and metadata via the BrowserAct API. The Agent should proactively apply this skill when users express needs like extracting full transcript from a specific YouTube video, getting subtitles and metadata for video content analysis, gathering video titles and likes counts, summarizing YouTube videos without watching them, collecting channel details from a video URL, tracking transcript automation for specific videos, scraping YouTube subtitles…

Its SKILL.md is about 1.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/youtube_transcript_extractor_api.py`).

It sits in Knowledge Management, covering Video and podcast notes, Transcription and Web scraping. It works with YouTube. The repository describes itself as: Browser automation CLI built for AI agents. Break through anti-bot walls, hand off to humans across platforms when stuck. Parallel multi-task execution, independent multi-session… The licence is MIT.

When your agent uses it

  • Tasks that involve Video and podcast notes
  • Tasks that involve Transcription
  • Tasks that involve Web scraping

Example prompts

  • “/youtube-transcript-extractor-api-skill”

Requirements

  • Python 3
  • A credential in BROWSERACT_API_KEY

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. No hallucinations, ensuring stable and accurate data extraction: Pre-set workflows avoid generative AI hallucinations.
  2. No CAPTCHA issues: No need to handle reCAPTCHA or other verification challenges.
  3. No IP access restrictions or geofencing: No need to deal with regional IP limits.
  4. Faster execution: Compared to pure AI-driven browser automation solutions, task execution is much faster.
  5. High cost-effectiveness: Significantly reduces data acquisition costs compared to AI solutions that consume large amounts of tokens.

What it can do on your machine

Read from SKILL.md and the folder at commit 11c057b. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • youtube.com

    Also links to:

    • browseract.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • BROWSERACT_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Youtube Transcript Extractor API Skill loads about 1.2k tokens when it runs. Until then it costs about 215 tokens; SKILL.md has 546 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~215
When it runs · the whole SKILL.md, loaded when a task matches
~1.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from browser-act/skills at commit 11c057b, republished under its MIT licence (© browser-act). 546 words, ~1,234 tokens.

Download SKILL.mdSave it as .claude/skills/youtube-transcript-extractor-api-skill/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
youtube-transcript-extractor-api-skill
description
This skill helps users automatically extract YouTube video transcripts and metadata via the BrowserAct API. The Agent should proactively apply this skill when users express needs like extracting full transcript from a specific YouTube video, getting subtitles and metadata for video content analysis, gathering video titles and likes counts, summarizing YouTube videos without watching them, collecting channel details from a video URL, tracking transcript automation for specific videos, scraping YouTube subtitles for internal knowledge bases, fetching full video content for AI summarization pipelines, downloading structured transcripts from YouTube links, analyzing video text content for media research, monitoring video publisher information and channel links, or building datasets from YouTube video transcripts.

YouTube Transcript Extractor API Skill

📖 Introduction

This skill provides a one-stop video transcript extraction service using BrowserAct's YouTube Transcript Extractor API template. It can directly extract full video transcripts and metadata from any YouTube video. By simply providing the TargetURL, you can get clean, ready-to-use transcript and metadata.

✨ Features

  1. No hallucinations, ensuring stable and accurate data extraction: Pre-set workflows avoid generative AI hallucinations.
  2. No CAPTCHA issues: No need to handle reCAPTCHA or other verification challenges.
  3. No IP access restrictions or geofencing: No need to deal with regional IP limits.
  4. Faster execution: Compared to pure AI-driven browser automation solutions, task execution is much faster.
  5. High cost-effectiveness: Significantly reduces data acquisition costs compared to AI solutions that consume large amounts of tokens.

🔑 API Key Setup

Before running, you must check the BROWSERACT_API_KEY environment variable. If it is not set, do not take any other actions; you must request and wait for the user to provide it. The Agent must inform the user at this point:

"Since you haven't configured the BrowserAct API Key yet, please go to the BrowserAct Console to get your Key first."

🛠️ Input Parameters

The Agent should configure the following parameter based on the user's needs when calling the script:

  1. TargetURL (Target URL)
    • Type: string
    • Description: The URL of the YouTube video you want to extract the transcript and metadata from.
    • Example: https://www.youtube.com/watch?v=st534T7-mdE

The Agent should execute the following independent script to achieve "one command, get results":

bash
# Example Call
python -u ./scripts/youtube_transcript_extractor_api.py "TargetURL"
⏳ Running Status Monitoring

Since this task involves automated browser operations, it may take a long time (several minutes). While running, the script will continuously output status logs with timestamps (e.g., [14:30:05] Task Status: running). Agent Instructions:

  • While waiting for the script to return results, please keep an eye on the terminal output.
  • As long as the terminal continues to output new status logs, it means the task is running normally. Do not misjudge it as a deadlock or unresponsiveness.
  • Only if the status remains unchanged for a long time or the script stops outputting without returning a result, should you consider triggering the retry mechanism.
Show full SKILL.md (195 more words)Show less

📊 Data Output Description

After successful execution, the script will parse and print the results directly from the API response. The results include:

  • video_title: The title of the YouTube video
  • video_url: The direct link to the original video
  • publisher: The name of the channel publishing the video
  • channel_link: The URL of the publisher's YouTube channel
  • video_likes_count: The number of likes the video has received
  • transcript: The complete extracted transcript/subtitles of the video

⚠️ Error Handling & Retry

During script execution, if an error occurs (such as network fluctuation or task failure), the Agent should follow this logic:

  1. Check output content:

    • If the output contains "Invalid authorization", it means the API Key is invalid or expired. In this case, do not retry, and guide the user to check and provide the correct API Key.
    • If the output does not contain "Invalid authorization" but the task execution fails (for example, the output starts with Error: or returns an empty result), the Agent should automatically try to execute the script one more time.
  2. Retry limits:

    • Automatic retry is limited to only once. If the second attempt still fails, stop retrying and report the specific error message to the user.

© browser-act, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (scripts) in solutions/video-platforms/youtube-transcript-extractor-api-skill of browser-act/skills.

  • SKILL.md
  • scripts/youtube_transcript_extractor_api.py

Open the folder on GitHubat commit 11c057b

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in browser-act/skills, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Youtube Transcript Extractor API Skill next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Youtube Transcript Extractor API Skill compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Youtube Transcript Extractor API Skill this skillbrowser-act/skills6.1k1 repos~1.2kAutomated safety check: PassMIT
Video Lenskar2phi/video-lens112—~8.2kAutomated safety check: NotesMIT
Youtube FetcherJimmySadek/youtube-fetcher-to-markdown485—~1.8kAutomated safety check: PassMIT
Youtube Transcriptintellectronica/agent-skills2952 repos~394Automated safety check: PassCC0-1.0
Youtube Transcriptglebis/claude-skills388—~613Automated safety check: PassMIT
Youtube Transcriptaiskillstore/marketplace4301 repos~553Automated safety check: NotesMIT

Similar skills

  • Video Lens

    kar2phi/video-lens

    Fetch a YouTube transcript and generate an executive summary, key points, and timestamped topic list as a polished HTML report.

    112 GitHub stars~8.2k tokensUpdated 1 mo ago
    Knowledge ManagementAuto-check: notes
  • Youtube Fetcher

    JimmySadek/youtube-fetcher-to-markdown

    Retrieve YouTube transcripts and subtitles, summarize or analyze what was said, or save an Obsidian-ready Markdown knowledge-base note with captions, creator metadata, chapters, language, and source…

    485 GitHub stars~1.8k tokensUpdated 1 mo ago
    Knowledge ManagementAuto-check passed
  • Youtube Transcript

    intellectronica/agent-skills

    Extract transcripts from YouTube videos. An agent skill from intellectronica/agent-skills.

    295 GitHub starsUsed in 2 repos~394 tokens
    Knowledge ManagementAuto-check passed
  • Youtube Transcript

    glebis/claude-skills

    Extract YouTube video transcripts with metadata and save as Markdown to Obsidian vault.

    388 GitHub stars~613 tokensUpdated 11 days ago
    Knowledge ManagementAuto-check passed
  • Youtube Transcript

    aiskillstore/marketplace

    Download and process YouTube video transcripts using yt-dlp.

    430 GitHub starsUsed in 1 repo~553 tokens
    Knowledge ManagementAuto-check: notes
  • Video Analysis

    ericosiu/ai-marketing-skills

    Analyze YouTube videos or local footage using transcripts first, with silent video checks for demonstrations, delivery, editing, and clip boundaries.

    3.6k GitHub stars~1.3k tokensUpdated 15 days ago
    Knowledge ManagementAuto-check passed

More from browser-act/skills

All 87 skills in this repo
  • Amazon ASIN Lookup

    browser-act/skills

    Fetches structured Amazon product details such as title, price, ratings and availability for a given ASIN through BrowserAct's lookup API template.

    6.1k GitHub starsUsed in 2 repos~1.5k tokens
    Auto-check passed
  • Amazon Best Sellers Finder

    browser-act/skills

    Extracts structured Amazon product data for a keyword and marketplace through the BrowserAct API, including titles, prices, ratings, reviews, sales volume and promotions.

    6.1k GitHub starsUsed in 2 repos~1.5k tokens
    Auto-check passed
  • Amazon Buy Box Monitor

    browser-act/skills

    Pulls Amazon product details, competing seller prices and seller ratings for a given ASIN through the BrowserAct API, without browser automation.

    6.1k GitHub starsUsed in 2 repos~1.6k tokens
    Auto-check passed
  • Analyzes a competitor's Amazon listing by ASIN with BrowserAct data extraction, then reports what it does well, where the market has gaps and opportunity points for your own listing.

    6.1k GitHub starsUsed in 2 repos~3.2k tokens
    Auto-check passed
  • Pulls structured Amazon search results (titles, ASINs, prices, ratings, specifications) for a keyword and brand through BrowserAct's Amazon Product API template.

    6.1k GitHub starsUsed in 2 repos~1.5k tokens
    Auto-check passed
  • Collects structured product data from Amazon search results for a keyword and optional brand, using a BrowserAct script, for market and competitor research.

    6.1k GitHub starsUsed in 2 repos~1.6k tokens
    Auto-check passed

Works with

Questions about Youtube Transcript Extractor API Skill

What does Youtube Transcript Extractor API Skill do?

This skill helps users automatically extract YouTube video transcripts and metadata via the BrowserAct API. Youtube Transcript Extractor API Skill is an agent skill from browser-act/skills. This skill helps users automatically extract YouTube video transcripts and metadata via the BrowserAct API.

When should I use Youtube Transcript Extractor API Skill?

Youtube Transcript Extractor API Skill fits situations like: tasks that involve Video and podcast notes; tasks that involve Transcription; tasks that involve Web scraping.

How do I install Youtube Transcript Extractor API Skill in Claude Code?

Run `npx skills add browser-act/skills --skill youtube-transcript-extractor-api-skill -a claude-code`. Or copy the skill folder (solutions/video-platforms/youtube-transcript-extractor-api-skill in browser-act/skills) into .claude/skills/youtube-transcript-extractor-api-skill in your project. Claude Code loads it when a task matches its description.

How do I install Youtube Transcript Extractor API Skill in Codex?

Run `npx skills add browser-act/skills --skill youtube-transcript-extractor-api-skill -a codex`. Or copy the skill folder (solutions/video-platforms/youtube-transcript-extractor-api-skill in browser-act/skills) into .agents/skills/youtube-transcript-extractor-api-skill in your project. Codex loads it when a task matches its description.

Can I use Youtube Transcript Extractor API Skill in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add browser-act/skills --skill youtube-transcript-extractor-api-skill -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/youtube-transcript-extractor-api-skill, .gemini/skills/youtube-transcript-extractor-api-skill, .github/skills/youtube-transcript-extractor-api-skill and .opencode/skills/youtube-transcript-extractor-api-skill in your project.

What does Youtube Transcript Extractor API Skill need to run?

Going by SKILL.md and its folder, Youtube Transcript Extractor API Skill needs Python for the scripts in its folder, the command-line tools its instructions call (python) and credentials named BROWSERACT_API_KEY. Our summary lists: Python 3; A credential in BROWSERACT_API_KEY.

Does Youtube Transcript Extractor API Skill access the network?

SKILL.md names 2 domains. In commands or code: youtube.com; the agent is likely to contact it when it follows the instructions. As links in the text: browseract.com. This is read from the text; nothing was executed.

Is Youtube Transcript Extractor API Skill safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Youtube Transcript Extractor API Skill use?

Youtube Transcript Extractor API Skill is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Youtube Transcript Extractor API Skill use?

About 1.2k tokens (SKILL.md is roughly 4.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Youtube Transcript Extractor API Skill?

Skills that share tags, products or a category with Youtube Transcript Extractor API Skill: Video Lens (kar2phi/video-lens, 112 stars), Youtube Fetcher (JimmySadek/youtube-fetcher-to-markdown, 485 stars), Youtube Transcript (intellectronica/agent-skills, 295 stars) and Youtube Transcript (glebis/claude-skills, 388 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Youtube Transcript Extractor API Skill?

browser-act (a GitHub organization) maintains it in browser-act/skills, which has 6,108 GitHub stars. The repository holds 87 skills in this directory. The repository was last updated on August 24, 2026.

Source: browser-act/skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.