Agent skill

Video Content Extractor

by sickn33 in sickn33/agentic-awesome-skills

Extract key frames from MP4 videos at configurable intervals, run Tesseract OCR, and generate structured Markdown reports with video metadata and timestamped text transcripts.

MITAuto-check passedDocuments & Office

Install Video Content Extractor

skills CLI
$ npx skills add sickn33/agentic-awesome-skills --skill video-content-extractor -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install sickn33/agentic-awesome-skills video-content-extractor --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/sickn33/agentic-awesome-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/video-content-extractor .claude/skills/video-content-extractor && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
video-content-extractor
GitHub stars
47k
Used in
1 other repo
Token cost
~1k tokens
SKILL.md length
506 words
Files
1
Skills in repo
1,493
Repo updated
First seen
Licence
MIT

At a glance

Extract key frames from MP4 videos at configurable intervals, run Tesseract OCR, and generate structured Markdown reports with video metadata and timestamped text transcripts.

  • Works in 4 steps: Analyze Video Metadata → Extract Key Frames → OCR Text Recognition → …
  • Tasks that involve Markdown
  • SKILL.md covers Overview, When to Use This Skill, How It Works and Examples, plus 5 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Video Content Extractor is an agent skill from sickn33/agentic-awesome-skills. Extract key frames from MP4 videos at configurable intervals, run Tesseract OCR, and generate structured Markdown reports with video metadata and timestamped text transcripts.

Its SKILL.md is about 1k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Documents & Office, covering Markdown. It works with FFmpeg. The repository describes itself as: AAS Core is the local, agent-first control plane for complete catalog discovery, agent-owned selection, stack validation, and planning, backed by 2,400+ agentic skills. Includes… The licence is MIT.

When your agent uses it

  • Tasks that involve Markdown

Example prompts

  • “/video-content-extractor”

Requirements

  • Python 3

Workflow steps

4 steps, taken from the step headings in SKILL.md.

  1. Analyze Video Metadata
  2. Extract Key Frames
  3. OCR Text Recognition
  4. Generate Markdown Report

What it can do on your machine

Read from SKILL.md and the folder at commit 680176d. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Video Content Extractor loads about 1k tokens when it runs. Until then it costs about 50 tokens; SKILL.md has 506 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~50
When it runs · the whole SKILL.md, loaded when a task matches
~1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from sickn33/agentic-awesome-skills at commit 680176d, republished under its MIT licence (© sickn33). 506 words, ~1,038 tokens.

Download SKILL.mdSave it as .claude/skills/video-content-extractor/SKILL.md (or your agent's skills folder).
name
video-content-extractor
description
Extract key frames from MP4 videos at configurable intervals, run Tesseract OCR, and generate structured Markdown reports with video metadata and timestamped text transcripts.
category
media-processing
risk
safe
source
community
source_repo
274326424/video-content-extractor
source_type
community
date_added
2026-06-06
author
274326424
tags
video, ocr, ffmpeg, tesseract, frame-extraction, media
tools
codex

Video Content Extractor

Overview

Automatically extracts key frames from MP4 video files at configurable time intervals, performs OCR text recognition on each frame, and generates a structured Markdown report. The report includes video metadata (duration, resolution, codecs) and frame-by-frame OCR transcripts with timestamp references.

This skill is designed for Codex CLI and requires FFmpeg and Tesseract OCR installed on the local machine.

When to Use This Skill

  • Use when you need to extract text content from video presentations, lectures, or screencasts.
  • Use when you want to create searchable transcripts from video files without embedded subtitles.
  • Use when you need to analyze video content programmatically and generate structured summaries.
  • Use when the user asks to "read what is on screen" or "extract the content from this video."

How It Works

Step 1: Analyze Video Metadata

The skill uses ffprobe to extract video metadata: duration, resolution, frame rate, codec information, and file size.

Step 2: Extract Key Frames

Using FFmpeg, the skill captures frames at the configured interval (default: every 30 seconds). Each frame is saved as a timestamped JPEG image.

Step 3: OCR Text Recognition

Each extracted frame is processed by Tesseract OCR. If the default PSM mode returns no meaningful text, it falls back to fully automatic page segmentation.

Step 4: Generate Markdown Report

All extracted data is assembled into a structured Markdown document.

Examples

Example 1: Basic Extraction

Agent prompt: Use the video-content-extractor skill to extract content from lecture.mp4

Output generates lecture.md and lecture_frames/ directory.

Example 2: Custom Interval

Parameters: video_path, output_dir, interval(seconds), lang Extract every 60 seconds with English-only OCR: python scripts/extract_video.py recording.mp4 ./output 60 eng

Example 3: Bilingual Content

Extract with default Chinese + English OCR: python scripts/extract_video.py lecture.mp4 . 15 chi_sim+eng

Best Practices

  • Use shorter intervals (10-15s) for fast-paced content with frequent text changes.
  • Use longer intervals (30-60s) for presentation slides or slow lectures to reduce duplicate frames.
  • For Chinese content, ensure Tesseract Chinese language pack is installed (chi_sim).
Show full SKILL.md (185 more words)Show less

Limitations

  • Requires FFmpeg and Tesseract OCR to be installed and accessible via PATH.
  • Tesseract OCR accuracy depends on video quality, text size, and font clarity.
  • Does not extract audio or perform speech-to-text transcription.
  • Frame extraction is time-based (not scene-change-based), which may produce near-duplicate frames.
  • Large videos with short intervals can generate many frames - ensure sufficient disk space.

Security and Safety Notes

  • This skill only reads video files and writes extracted frames and Markdown reports.
  • It does NOT send any data over the network - all processing is local.
  • FFmpeg and Tesseract are invoked with fixed, pre-vetted arguments.
  • The skill does not modify or delete the original video file.

Common Pitfalls

  • Problem: Tesseract returns garbled text Solution: Ensure the correct language pack is installed. Run tesseract --list-langs to verify.

  • Problem: FFmpeg fails with "not found" Solution: Make sure FFmpeg is on PATH. Run ffmpeg -version to verify.

  • Problem: OCR is slow on large videos Solution: Increase the interval parameter to reduce frames processed.

  • @media-summarizer - For summarizing video content using visual and audio cues.
  • @document-ocr - For OCR on static images or scanned documents without video processing.

© sickn33, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/video-content-extractor of sickn33/agentic-awesome-skills.

Open the folder on GitHubat commit 680176d

Used in 1 other repository

We found 5 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in sickn33/agentic-awesome-skills, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Video Content Extractor next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Video Content Extractor compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Video Content Extractor this skillsickn33/agentic-awesome-skills47k1 repos~1kAutomated safety check: PassMIT
dedao-dl Download Assistantyann0917/dedao-dl896—~772Automated safety check: PassMIT
PPT Visual Enhancement Generatoranbeime/skill7.7k—~1.2kAutomated safety check: NotesNone
Docs Page Frontmatteranam-org/metaxy124—~948Automated safety check: PassApache-2.0
Mira Slide To Videosandeco/mira-animator218—~2kAutomated safety check: PassCustom licence
Save Mdmblode/agent-skills144—~537Automated safety check: PassMIT

Similar skills

  • Operates the dedao-dl command-line downloader for Dedao courses, audiobooks and ebooks, with login, search, download formats and troubleshooting.

    896 GitHub stars~772 tokensUpdated yesterday
    Documents & OfficeAuto-check passed
  • Turns a PPT content outline into styled slide images, an interactive HTML viewer and an optional stitched video from a content-planning partner skill.

    7.7k GitHub stars~1.2k tokensUpdated yesterday
    Documents & OfficeAuto-check: notes
  • Docs Page Frontmatter

    anam-org/metaxy

    Write YAML front matter for documentation pages with appropriate titles and descriptions for social cards.

    124 GitHub stars~948 tokensUpdated 8 days ago
    Documents & OfficeAuto-check passed
  • Mira Slide To Video

    sandeco/mira-animator

    Gera um .mp4 de um ou mais slides de um deck do Mira gravando a animação real (Chrome headless + ffmpeg): cada slide dispara do zero, sem vazar o anterior, preenchendo o frame.

    218 GitHub stars~2k tokensUpdated 25 days ago
    Documents & OfficeAuto-check passed
  • Save Md

    mblode/agent-skills

    Saves a named source to Markdown with provenance and faithful extraction through direct export endpoints.

    144 GitHub stars~537 tokensUpdated yesterday
    Documents & OfficeAuto-check passed
  • Ppt Compressor

    LeoYeAI/openclaw-master-skills

    This skill should be used when the user wants to compress a PowerPoint (.pptx) file by reducing the size of embedded videos and large images.

    2.2k GitHub stars~3.9k tokensUpdated 2 mo ago
    Documents & OfficeAuto-check passed

More from sickn33/agentic-awesome-skills

All 1,493 skills in this repo
  • Liuguang Banlan UI

    sickn33/agentic-awesome-skills

    Implements an interface in one of two named color modes, iridescent white or colorful black, from a parameterized starter that reports measured color intensity.

    47k GitHub starsUsed in 1 repo~2.5k tokens
    Auto-check passed
  • User Thoughts Memory

    sickn33/agentic-awesome-skills

    Saves a user's project decisions, rules and preferences into a project-local mdbase so later sessions and other agents can recover the intent.

    47k GitHub starsUsed in 1 repo~2.5k tokens
    Auto-check passed
  • Using LWC Memory and Graphs

    sickn33/agentic-awesome-skills

    Keeps project decisions, research and verified results available across coding-agent sessions through LWC memory, a document Wiki graph and a CodeGraph code index.

    47k GitHub starsUsed in 1 repo~2k tokens
    Auto-check passed
  • Find Complementary Founders

    sickn33/agentic-awesome-skills

    Guides an agent through assessing its own owner for cofounder fit, publishing an approved profile, and ranking complementary profiles other agents published for their owners.

    47k GitHub starsUsed in 1 repo~4.8k tokens
    Auto-check passed
  • Whatsapp Cloud API

    sickn33/agentic-awesome-skills

    Integracao com WhatsApp Business Cloud API (Meta). An agent skill from sickn33/agentic-awesome-skills.

    47k GitHub starsUsed in 2 repos~4.5k tokens
    Auto-check passed
  • Cline Pilot

    sickn33/agentic-awesome-skills

    Acts as a proxy for the Cline CLI, dispatching coding tasks one at a time, monitoring runs by hard evidence, relaying decisions to you and learning per-project preferences.

    47k GitHub starsUsed in 1 repo~4.6k tokens
    Auto-check passed

Works with

Questions about Video Content Extractor

What does Video Content Extractor do?

Extract key frames from MP4 videos at configurable intervals, run Tesseract OCR, and generate structured Markdown reports with video metadata and timestamped text transcripts. Video Content Extractor is an agent skill from sickn33/agentic-awesome-skills. Extract key frames from MP4 videos at configurable intervals, run Tesseract OCR, and generate structured Markdown reports with video metadata and timestamped text transcripts.

When should I use Video Content Extractor?

Video Content Extractor fits situations like: tasks that involve Markdown.

How do I install Video Content Extractor in Claude Code?

Run `npx skills add sickn33/agentic-awesome-skills --skill video-content-extractor -a claude-code`. Or copy the skill folder (skills/video-content-extractor in sickn33/agentic-awesome-skills) into .claude/skills/video-content-extractor in your project. Claude Code loads it when a task matches its description.

How do I install Video Content Extractor in Codex?

Run `npx skills add sickn33/agentic-awesome-skills --skill video-content-extractor -a codex`. Or copy the skill folder (skills/video-content-extractor in sickn33/agentic-awesome-skills) into .agents/skills/video-content-extractor in your project. Codex loads it when a task matches its description.

Can I use Video Content Extractor in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add sickn33/agentic-awesome-skills --skill video-content-extractor -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/video-content-extractor, .gemini/skills/video-content-extractor, .github/skills/video-content-extractor and .opencode/skills/video-content-extractor in your project.

What does Video Content Extractor need to run?

SKILL.md names no scripts, command-line tools or credentials: Video Content Extractor is instructions for the agent only. Our summary lists: Python 3.

Does Video Content Extractor access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Video Content Extractor safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Video Content Extractor use?

Video Content Extractor is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Video Content Extractor use?

About 1k tokens (SKILL.md is roughly 4.2k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Video Content Extractor?

Skills that share tags, products or a category with Video Content Extractor: dedao-dl Download Assistant (yann0917/dedao-dl, 896 stars), PPT Visual Enhancement Generator (anbeime/skill, 7.7k stars), Docs Page Frontmatter (anam-org/metaxy, 124 stars) and Mira Slide To Video (sandeco/mira-animator, 218 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Video Content Extractor?

sickn33 (a GitHub user) maintains it in sickn33/agentic-awesome-skills, which has 47,379 GitHub stars. The repository holds 1,493 skills in this directory. The repository was last updated on October 9, 2026.

Source: sickn33/agentic-awesome-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.