Embedded Video Captions
heygen-com/hyperframes
Adds captions to a single-subject talking-head video without editing the footage, from plain subtitles to cinematic text placed behind the speaker.
Turns raw or rough-cut talking-head footage into a captioned, animated final video through a staged, resumable production pipeline.
$ npx skills add naive-kun/naive-video-skill --skill talking-head-video-pipeline -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install naive-kun/naive-video-skill talking-head-video-pipeline --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
Claude Code skills documentation · loads skills from .claude/skills/
Install the "talking-head-video-pipeline" agent skill from https://github.com/naive-kun/naive-video-skill/tree/main into .claude/skills/talking-head-video-pipeline/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "talking-head-video-pipeline", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add naive-kun/naive-video-skill --skill talking-head-video-pipeline -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install naive-kun/naive-video-skill talking-head-video-pipeline --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "talking-head-video-pipeline" agent skill from https://github.com/naive-kun/naive-video-skill/tree/main into .agents/skills/talking-head-video-pipeline/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "talking-head-video-pipeline", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add naive-kun/naive-video-skill --skill talking-head-video-pipeline -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install naive-kun/naive-video-skill talking-head-video-pipeline --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "talking-head-video-pipeline" agent skill from https://github.com/naive-kun/naive-video-skill/tree/main into .cursor/skills/talking-head-video-pipeline/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "talking-head-video-pipeline", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add naive-kun/naive-video-skill --skill talking-head-video-pipeline -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install naive-kun/naive-video-skill talking-head-video-pipeline --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "talking-head-video-pipeline" agent skill from https://github.com/naive-kun/naive-video-skill/tree/main into .gemini/skills/talking-head-video-pipeline/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "talking-head-video-pipeline", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install naive-kun/naive-video-skill talking-head-video-pipelineInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add naive-kun/naive-video-skill --skill talking-head-video-pipeline -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "talking-head-video-pipeline" agent skill from https://github.com/naive-kun/naive-video-skill/tree/main into .github/skills/talking-head-video-pipeline/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "talking-head-video-pipeline", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add naive-kun/naive-video-skill --skill talking-head-video-pipeline -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install naive-kun/naive-video-skill talking-head-video-pipeline --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "talking-head-video-pipeline" agent skill from https://github.com/naive-kun/naive-video-skill/tree/main into .opencode/skills/talking-head-video-pipeline/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "talking-head-video-pipeline", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
talking-head-video-pipelineTurns raw or rough-cut talking-head footage into a captioned, animated final video through a staged, resumable production pipeline.
This skill runs talking-head video production as a staged pipeline for someone who may never have edited video with an agent before: it asks one question at a time only when it can't infer the answer, defaults to a working choice over a menu of technical options, and explains the next visible result rather than renderer internals.
It treats the main audio track as the immovable clock and the original source footage as untouchable, optionally helps with a rough cut, transcription and caption revision, groups content for the viewer, places screenshots and demos, builds a staged keyframe and HyperFrames or GSAP preview, and only produces a synchronized final export after that preview is approved. Work can be interrupted and resumed, diagnosed, and revised within a limited scope rather than restarted from scratch, and it can optionally pull in raw footage through a separate video-use integration.
It keeps project paths, media names, brand rules, screenshots and API keys inside the project rather than inside the skill, and it only learns a lasting style preference from explicit feedback, never from silence or a single unconfirmed draft.
12 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 3e9c7c5. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships 1 file in scripts/ (Shell, from the files we listed), which the agent can run.
Shell commands in SKILL.md call:
python3bashFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Talking-Head Video Pipeline loads about 3.3k tokens when it runs, and up to ~28k if it reads all its reference files. Until then it costs about 161 tokens; SKILL.md has 1,485 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
The full file from naive-kun/naive-video-skill at commit 3e9c7c5, republished under its MIT licence (© naive-kun). 1,485 words, ~3,320 tokens.
.claude/skills/talking-head-video-pipeline/SKILL.md (or your agent's skills folder). This skill also uses 76 other files; get the full folder from GitHub.Turn raw or rough-cut talking-head footage into an approved working cut, editable captions, a motion-packaged preview, and a verified final export. Treat this file as the router and shared contract. Read only the routed internal workflow and references needed for the current stage.
Assume the user has never edited with Codex.
working_video becomes the downstream clock while the original remains immutable.references/workflows/; do not add nested SKILL.md files that package managers may register separately.Read references/quality-gates.md before preview or export work.
| User intent or phrase | Route | Typical output |
|---|---|---|
初始化视频项目, 第一次用, start a video project | references/workflows/init.md | State, edit plan, design profile |
原片还没剪, 删口误和停顿, 帮我粗剪, raw takes, rough cut | references/workflows/rough-cut.md | Approved non-destructive rough cut and timing handoff |
抽字幕轴, 只做字幕, transcribe, 改字幕 | references/workflows/captions.md | SRT, CSV, transcript JSON |
帮我设计, 换颜色风格, 规划弹窗, design | references/workflows/design.md | DESIGN and complete edit plan |
参考这张截图设计, 按这张图的视觉语言, style reference | references/workflows/design.md reference branch | Project-local STYLE_REFERENCE |
按语义匹配动效, GSAP 动效丰富一点, semantic motion | references/workflows/design.md, then preview | Validated MOTION_PLAN and preview |
参考 ShotCraft, 推荐电影感镜头, cinematic shot reference | references/workflows/design.md ShotCraft branch | Optional mapped shot references in MOTION_PLAN |
做预览, 加动效, build preview | references/workflows/preview.md | Official preview URL |
出成片, 导出 4K, export final | references/workflows/export.md | Verified final video |
改这个, 恢复上一版, revise | references/workflows/revise.md | Scoped revision and new preview/final |
进度, 到哪了, status | references/workflows/status.md | Current stage and next action |
体检, 为什么失败, doctor | references/workflows/doctor.md | Read-only diagnosis |
以后都这样, 记住这个风格, learn this | references/workflows/learn.md | Confirmed project-local lessons |
复盘成片, 下次怎么改进, retro | references/workflows/retro.md | Structured delivery retrospective |
升级状态, 迁移, migrate | references/workflows/migrate.md | State schema upgrade |
Before routing:
.naive-video-state.json.EDIT_PLAN.md, DESIGN.md, and VIDEO_LESSONS.md only when relevant. Use working_video when present; otherwise use main_video.python3 tools/video_doctor.py --project <project_dir> when state and files disagree.Use these stages:
initialized -> captions_ready -> design_ready -> preview_ready
-> approved -> rendering -> final_readyNever mark a later stage until its quality gate passes.
When the user asks for the complete workflow:
ffprobe.working_video when present, otherwise main_video.skip. Preserve existing approved projects, the native recipe, and offline fallback.The init workflow creates this minimal project memory without touching source media:
<project_dir>/
├── .naive-video-state.json
├── EDIT_PLAN.md
├── DESIGN.md
├── CONTENT_LOGIC.json # viewer-facing reasoning groups and timed beats
├── STYLE_REFERENCE.md # created only when a reference image is used
├── MOTION_PLAN.json # created when semantic motion is planned
├── VIDEO_LESSONS.md
├── VIDEO_RETRO.md
├── edit/
│ ├── rough-cut.mp4 # optional; original media is never overwritten
│ ├── rough-cut-edl.json # optional decision record
│ ├── script-aligned.srt
│ ├── caption-table.csv
│ └── transcripts/
├── preview/
├── final/
└── qa/
└── KEYFRAME_REVIEW.md # static-composition review before dynamic previewState is operational metadata, not a media database. Keep it small and never store transcript bodies, private screenshots, or secrets in it. See references/state-management.md.
If the user wants speed and gives no reference:
Ask for an accent color only if brand consistency matters. Otherwise use the neutral preset and make it easy to change later. Beginners may optionally provide a screenshot and choose low, medium, or high reference strength; explain that the skill copies visual language, never the source brand or content. Motion density is restrained, balanced, or energetic. See references/style-onboarding.md.
Detailed GSAP recipes and semantic mappings live in references/motion-recipes.md. Runtime and plugin selection live in references/gsap-runtime.md. Screenshot extraction and anti-copy rules live in references/style-reference-workflow.md. Asset timing choices live in references/asset-onboarding.md. Load them only for the matching design branch.
ShotCraft discovery and adaptation rules live in references/shotcraft-integration.md; the native beginner pack lives in references/shotcraft-default-pack.md. Load them for new automatic projects, named cards, or cinematic shot requests. Never install a provider silently.
Word-level timing, viewer-facing logic groups, and accumulation/exit behavior live in references/content-logic-workflow.md. Load it after captions and before semantic design.
Typography, component geometry, glass notifications, and seek-safe focus/type/split adaptation live in references/visual-quality-rules.md. Load it for every design or preview task.
Self-iteration means improving the current user's workflow without leaking it into public defaults.
Classify feedback into one scope:
project: applies only to this video.profile: applies to this user's future videos; write it only after explicit confirmation.product: a privacy-safe, general reliability improvement; propose it to the skill maintainer separately.Record profile feedback in VIDEO_LESSONS.md with the user's words, the confirmed rule, and the affected stage. Never copy media paths or private evidence into this repository. See references/self-iteration.md.
Use the retro workflow after delivery to separate failures, environment issues, and taste feedback before promoting any rule. New projects import active private-profile rules into VIDEO_LESSONS.md so the editor can actually apply them.
Run:
bash scripts/doctor.sh --privacy-scan .
python3 tools/validate_skill.py .Block publication if either command reports personal absolute paths, secrets, private media names, invalid frontmatter, multiple skill manifests, missing internal workflows, unsafe install commands, or broken templates.
© naive-kun, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 76 other files (scripts, references) in the repository root of naive-kun/naive-video-skill.
Open the folder on GitHubat commit 3e9c7c5
Talking-Head Video Pipeline next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Talking-Head Video Pipeline this skillnaive-kun/naive-video-skill | 132 | — | ~3.3k | Automated safety check: Pass | MIT | |
| Embedded Video Captionsheygen-com/hyperframes | 60k | 3 repos | ~8.6k | Automated safety check: Pass | Apache-2.0 | |
| Yuv Viral Videohoodini/ai-agents-skills | 282 | — | ~7.5k | Automated safety check: Notes | None | |
| Noti Tiktok Full Textnotivn/AIEV | 127 | — | ~4.8k | Automated safety check: Pass | MIT | |
| Noti Tiktok Vnnotivn/AIEV | 127 | — | ~5.4k | Automated safety check: Pass | MIT | |
| HyperFrames Animationheygen-com/hyperframes | 60k | 3 repos | ~2.1k | Automated safety check: Pass | Apache-2.0 |
heygen-com/hyperframes
Adds captions to a single-subject talking-head video without editing the footage, from plain subtitles to cinematic text placed behind the speaker.
hoodini/ai-agents-skills
Edit any selfie or screen-share footage into a viral short-form video in YUV.AI's signature style — Apple-style liquid-glass cards (real CSS backdrop-filter), dark-mode polish, MrBeast-paced cuts…
notivn/AIEV
Build a Vietnamese vertical TikTok explainer in the "MỔ XẺ PAPER AI" (AI paper dissection) format with HyperFrames (HTML/CSS/GSAP → MP4), Noti.vn style.
notivn/AIEV
Edit a Vietnamese vertical TikTok video (9:16) with HyperFrames following the Noti.vn/GĐT standard - talking-head + kinetic typography + karaoke captions + zoom/punch-in camera + timestamp-synced…
heygen-com/hyperframes
Collects motion rules, scene blueprints, transitions and runtime adapters for HyperFrames video compositions, with GSAP as the default animation runtime.
heygen-com/hyperframes
Imports Figma assets, brand tokens, components and motion into a HyperFrames video composition, using the Figma REST API with a connector or native export for shaders.
Categories
Turns raw or rough-cut talking-head footage into a captioned, animated final video through a staged, resumable production pipeline. This skill runs talking-head video production as a staged pipeline for someone who may never have edited video with an agent before: it asks one question at a time only when it can't infer the answer, defaults to a working choice over a menu of technical options, and explains the next visible result rather than renderer internals.
Talking-Head Video Pipeline fits situations like: turning rough talking-head footage into a captioned final video; resuming a talking-head edit that was interrupted partway through; reviewing and revising captions or keyframes before a final export.
Run `npx skills add naive-kun/naive-video-skill --skill talking-head-video-pipeline -a claude-code`. Or copy the skill folder (the naive-kun/naive-video-skill repository) into .claude/skills/talking-head-video-pipeline in your project. Claude Code loads it when a task matches its description.
Run `npx skills add naive-kun/naive-video-skill --skill talking-head-video-pipeline -a codex`. Or copy the skill folder (the naive-kun/naive-video-skill repository) into .agents/skills/talking-head-video-pipeline in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add naive-kun/naive-video-skill --skill talking-head-video-pipeline -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/talking-head-video-pipeline, .gemini/skills/talking-head-video-pipeline, .github/skills/talking-head-video-pipeline and .opencode/skills/talking-head-video-pipeline in your project.
Going by SKILL.md and its folder, Talking-Head Video Pipeline needs a shell for the scripts in its folder and the command-line tools its instructions call (python3 and bash).
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
Talking-Head Video Pipeline is published under the MIT licence (from the LICENSE file in the skill folder). It allows redistribution, so the full SKILL.md is shown on this page.
About 3.3k tokens (SKILL.md is roughly 13k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 25k tokens, read only when the agent opens those files.
Skills that share tags, products or a category with Talking-Head Video Pipeline: Embedded Video Captions (heygen-com/hyperframes, 60k stars), Yuv Viral Video (hoodini/ai-agents-skills, 282 stars), Noti Tiktok Full Text (notivn/AIEV, 127 stars) and Noti Tiktok Vn (notivn/AIEV, 127 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
naive-kun (a GitHub user) maintains it in naive-kun/naive-video-skill, which has 132 GitHub stars. The repository was last updated on August 11, 2026.
Source: naive-kun/naive-video-skill on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.