Youtube Clipper
op7418/Youtube-clipper-skill
YouTube 视频智能剪辑工具。下载视频和字幕,AI 分析生成精细章节(几分钟级别), 用户选择片段后自动剪辑、翻译字幕为中英双语、烧录字幕到视频,并生成总结文案。
Capture frames from a YouTube video at a regular interval, produce a manifest mapping filenames to timestamps, and describe each frame with an LLM.
$ npx skills add pamelafox/presentation-skills --skill capture-video-frames -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install pamelafox/presentation-skills capture-video-frames --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/pamelafox/presentation-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/capture-video-frames .claude/skills/capture-video-frames && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "capture-video-frames" agent skill from https://github.com/pamelafox/presentation-skills/tree/main/.agents/skills/capture-video-frames into .claude/skills/capture-video-frames/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "capture-video-frames", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/pamelafox/presentation-skills/tree/main/.agents/skills/capture-video-framesType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add pamelafox/presentation-skills --skill capture-video-frames -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install pamelafox/presentation-skills capture-video-frames --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/pamelafox/presentation-skills.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.agents/skills/capture-video-frames .agents/skills/capture-video-frames && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "capture-video-frames" agent skill from https://github.com/pamelafox/presentation-skills/tree/main/.agents/skills/capture-video-frames into .agents/skills/capture-video-frames/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "capture-video-frames", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add pamelafox/presentation-skills --skill capture-video-frames -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install pamelafox/presentation-skills capture-video-frames --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/pamelafox/presentation-skills.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.agents/skills/capture-video-frames .cursor/skills/capture-video-frames && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "capture-video-frames" agent skill from https://github.com/pamelafox/presentation-skills/tree/main/.agents/skills/capture-video-frames into .cursor/skills/capture-video-frames/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "capture-video-frames", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/pamelafox/presentation-skills.git --path .agents/skills/capture-video-frames--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add pamelafox/presentation-skills --skill capture-video-frames -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install pamelafox/presentation-skills capture-video-frames --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/pamelafox/presentation-skills.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.agents/skills/capture-video-frames .gemini/skills/capture-video-frames && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "capture-video-frames" agent skill from https://github.com/pamelafox/presentation-skills/tree/main/.agents/skills/capture-video-frames into .gemini/skills/capture-video-frames/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "capture-video-frames", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install pamelafox/presentation-skills capture-video-framesInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add pamelafox/presentation-skills --skill capture-video-frames -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/pamelafox/presentation-skills.git skills-src && mkdir -p .github/skills && cp -r skills-src/.agents/skills/capture-video-frames .github/skills/capture-video-frames && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "capture-video-frames" agent skill from https://github.com/pamelafox/presentation-skills/tree/main/.agents/skills/capture-video-frames into .github/skills/capture-video-frames/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "capture-video-frames", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add pamelafox/presentation-skills --skill capture-video-frames -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install pamelafox/presentation-skills capture-video-frames --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/pamelafox/presentation-skills.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.agents/skills/capture-video-frames .opencode/skills/capture-video-frames && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "capture-video-frames" agent skill from https://github.com/pamelafox/presentation-skills/tree/main/.agents/skills/capture-video-frames into .opencode/skills/capture-video-frames/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "capture-video-frames", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
capture-video-framesCapture frames from a YouTube video at a regular interval, produce a manifest mapping filenames to timestamps, and describe each frame with an LLM.
Capture Video Frames is an agent skill from pamelafox/presentation-skills. Capture frames from a YouTube video at a regular interval, produce a manifest mapping filenames to timestamps, and describe each frame with an LLM. USE FOR: extract video frames, capture screenshots from YouTube, describe video frames, video frame analysis, frame-by-frame summary.
Its SKILL.md is about 1.7k tokens, which your agent loads only when the skill is triggered. The skill folder holds 1 other file (for example `capture_video_frames.py`).
It sits in Agent Workflows. It works with YouTube and FFmpeg. The repository describes itself as: Skills for AI agents to process presentations - helpful for teachers and speakers. The licence is MIT.
3 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 2b809b3. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships script files (Python), which the agent can run.
Shell commands in SKILL.md call:
brewuvyt-dlpffmpegpipapt-getFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use uv and pip, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Capture Video Frames loads about 1.7k tokens when it runs. Until then it costs about 76 tokens; SKILL.md has 663 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from pamelafox/presentation-skills at commit 2b809b3, republished under its MIT licence (© pamelafox). 663 words, ~1,689 tokens.
.claude/skills/capture-video-frames/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.Run the capture_video_frames.py script:
uv run .agents/skills/capture-video-frames/capture_video_frames.py <youtube_url> <output_dir> [--interval SECONDS]youtube_url (required): YouTube video URL (same formats accepted by the extract-transcript skill).output_dir (required): Directory to save frames and the manifest file. Created if it doesn't exist.--interval (optional): Seconds between captured frames. Defaults to 30.Example frames_manifest.md:
| File | Timestamp | Description |
|------|-----------|-------------|
| frame_0000.png | [00:00] | |
| frame_0030.png | [00:30] | |
| frame_0060.png | [01:00] | |brew install yt-dlp or pip install yt-dlpbrew install ffmpeg or apt-get install ffmpegdescribe-frame subagentAfter capturing frames, describe each frame by running the describe-frame custom agent as a subagent. Each subagent invocation gets an isolated context, so frame images won't accumulate and exhaust the context window.
The describe-frame agent is defined in .github/agents/describe-frame.md.
describe-frame agent as a subagent with a prompt that includes:(same as previous) if the frame is essentially identical to the previous one).Use this as the prompt when invoking the describe-frame subagent (fill in the bracketed values):
Describe the current frame image at: [CURRENT_FRAME_ABSOLUTE_PATH]
[If previous frame exists, include these two lines:]
The previous frame image is at: [PREVIOUS_FRAME_ABSOLUTE_PATH]
The previous frame was described as: "[PREVIOUS_DESCRIPTION]"After describing all frames, frames_manifest.md should look like:
| File | Timestamp | Description |
|------|-----------|-------------|
| frame_0000.png | [00:00] | Title slide introducing "Building RAG apps with Python" |
| frame_0030.png | [00:30] | Speaker showing the agenda with four main topics |
| frame_0060.png | [01:00] | (same as previous) |
| frame_0090.png | [01:30] | Architecture diagram of a retrieval-augmented generation pipeline |After all frames are described, groups of consecutive (same as previous) rows represent the same visual content captured at different moments. Within each group, speaker faces may differ — eyes open vs closed, mouth open vs closed, facing camera vs turned away.
For each group of duplicate frames, keep only one frame — the one with the best speaker face quality — and remove the rest.
(same as previous). Each group starts with the "anchor" frame (the one with an actual description) followed by one or more (same as previous) rows.describe-frame subagent to compare faces across the anchor frame and each duplicate. Use this prompt template:Compare these two frames focusing ONLY on the speaker faces visible in webcam feeds. Which frame has better speaker faces — eyes open, facing camera, mouth open (mid-speech), not mid-blink or turned away?
Frame A: [ANCHOR_FRAME_ABSOLUTE_PATH]
Frame B: [DUPLICATE_FRAME_ABSOLUTE_PATH]
Reply with ONLY one of:
- "A BETTER" if the anchor frame has better speaker faces
- "B BETTER" if the duplicate frame has better speaker faces
- "EQUAL" if both are equivalent
- "NO SPEAKERS" if no speaker faces are visible in either frame
Then add a brief reason.(better speaker faces than frame_XXXX: eyes open, facing camera)(same as previous) rows from the manifest.After deduplication, if the best frame in a group still has the speaking person's mouth closed (both speakers have mouths closed), try recapturing at nearby timestamps:
yt-dlp -f "bestvideo[height<=720]" --no-playlist -o "<output_dir>/video.%(ext)s" "<youtube_url>"ffmpeg -ss <SECONDS> -i <output_dir>/video.mp4 -frames:v 1 -q:v 2 <output_dir>/alt_<FRAME>_<SECONDS>.png -ydescribe-frame subagent to check if the speaker's mouth is open in any alternative, AND that the slide/demo content is still the same.cp alt_XXXX.png frame_XXXX.png).rm -f <output_dir>/alt_*.png <output_dir>/video.mp4© pamelafox, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 1 other file in .agents/skills/capture-video-frames of pamelafox/presentation-skills.
Open the folder on GitHubat commit 2b809b3
Capture Video Frames next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Capture Video Frames this skillpamelafox/presentation-skills | 125 | — | ~1.7k | Automated safety check: Pass | MIT | |
| Youtube Clipperop7418/Youtube-clipper-skill | 2.2k | — | ~1.6k | Automated safety check: Notes | MIT | |
| Video Transcribewendy7756/AI-Video-Transcriber | 3.3k | — | ~937 | Automated safety check: Notes | Apache-2.0 | |
| Claude Real VideoHUANGCHIHHUNGLeo/claude-real-video | 2.2k | — | ~639 | Automated safety check: Pass | MIT | |
| Ffmpeg Skillkajisho5/ffmpeg-skill | 1.9k | — | ~7.4k | Automated safety check: Pass | MIT | |
| Video Perceptionjordanrendric/claude-video-vision | 1.4k | — | ~1.4k | Automated safety check: Pass | MIT |
op7418/Youtube-clipper-skill
YouTube 视频智能剪辑工具。下载视频和字幕,AI 分析生成精细章节(几分钟级别), 用户选择片段后自动剪辑、翻译字幕为中英双语、烧录字幕到视频,并生成总结文案。
wendy7756/AI-Video-Transcriber
Transcribe and summarize a video or podcast from a URL (YouTube, TikTok, Bilibili, Apple Podcasts, SoundCloud, 30+ platforms) or from a local media/.txt file.
HUANGCHIHHUNGLeo/claude-real-video
Watch a video for the user. An agent skill from HUANGCHIHHUNGLeo/claude-real-video.
kajisho5/ffmpeg-skill
Edit video and audio with local FFmpeg from natural-language requests: cut, trim, join, resize/reframe (9:16, 1:1), speed change, captions and subtitles (SRT/ASS, animated, karaoke), logos and text…
jordanrendric/claude-video-vision
A skill your agent uses when the user mentions a video file (.mp4, .mov, .avi, .mkv, .webm), a YouTube URL, asks to watch/analyze/review a video, or references video content in conversation
vincentsch/explainroo
Make an explainer video (MP4) with a voice-over using explainroo.
pamelafox/presentation-skills
Generate or edit bitmap images with Microsoft MAI-Image-2.5 through the Azure AI image APIs.
pamelafox/presentation-skills
Create or update a RevealJS HTML presentation using the repository's bundled slide template.
pamelafox/presentation-skills
Fetch presentation slides from a URL and convert them to PDF.
pamelafox/presentation-skills
Generate an annotated blog-style write-up from a presentation's slides and video recording.
pamelafox/presentation-skills
Generate a numbered outline of presentation slides with one-sentence summaries.
pamelafox/presentation-skills
Capture a thumbnail image of a slide from a OneDrive/Office presentation link.
Categories
Capture frames from a YouTube video at a regular interval, produce a manifest mapping filenames to timestamps, and describe each frame with an LLM. Capture Video Frames is an agent skill from pamelafox/presentation-skills. Capture frames from a YouTube video at a regular interval, produce a manifest mapping filenames to timestamps, and describe each frame with an LLM.
Capture Video Frames fits situations like: : extract video frames; capture screenshots from YouTube; describe video frames; video frame analysis.
Run `npx skills add pamelafox/presentation-skills --skill capture-video-frames -a claude-code`. Or copy the skill folder (.agents/skills/capture-video-frames in pamelafox/presentation-skills) into .claude/skills/capture-video-frames in your project. Claude Code loads it when a task matches its description.
Run `npx skills add pamelafox/presentation-skills --skill capture-video-frames -a codex`. Or copy the skill folder (.agents/skills/capture-video-frames in pamelafox/presentation-skills) into .agents/skills/capture-video-frames in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add pamelafox/presentation-skills --skill capture-video-frames -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/capture-video-frames, .gemini/skills/capture-video-frames, .github/skills/capture-video-frames and .opencode/skills/capture-video-frames in your project.
Going by SKILL.md and its folder, Capture Video Frames needs Python for the scripts in its folder and the command-line tools its instructions call (brew, uv, yt-dlp, ffmpeg, pip and apt-get). Our summary lists: Python 3.
SKILL.md contains no URLs. Its commands use uv and pip, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Capture Video Frames is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.7k tokens (SKILL.md is roughly 6.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Capture Video Frames: Youtube Clipper (op7418/Youtube-clipper-skill, 2.2k stars), Video Transcribe (wendy7756/AI-Video-Transcriber, 3.3k stars), Claude Real Video (HUANGCHIHHUNGLeo/claude-real-video, 2.2k stars) and Ffmpeg Skill (kajisho5/ffmpeg-skill, 1.9k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
pamelafox (a GitHub user) maintains it in pamelafox/presentation-skills, which has 125 GitHub stars. The repository holds 14 skills in this directory. The repository was last updated on September 2, 2026.
Source: pamelafox/presentation-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.