Agent skill

Explainer Video

by 0xsline in 0xsline/OpenChatCut

Create finished explainer videos from a topic, script, outline, voiceover, product logic, data, technical concept, course material, or reference assets.

AGPL-3.0Auto-check passedMedia & Creative

Install Explainer Video

skills CLI
$ npx skills add 0xsline/OpenChatCut --skill explainer-video -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install 0xsline/OpenChatCut explainer-video --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/0xsline/OpenChatCut.git skills-src && mkdir -p .claude/skills && cp -r skills-src/src/agent/skills/explainer-video .claude/skills/explainer-video && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
explainer-video
GitHub stars
2.2k
Token cost
~2.8k tokens
SKILL.md length
1,429 words
Files
2 (incl. references)
Skills in repo
31
Repo updated
First seen
Licence
AGPL-3.0

At a glance

Create finished explainer videos from a topic, script, outline, voiceover, product logic, data, technical concept, course material, or reference assets.

  • Works in 12 steps: Read the project state, prompt, attached… → Identify the working labels → Respect source structure. If the user… → …
  • The user wants narration
  • SKILL.md covers Workflow, Rules and Plan Format
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Explainer Video is an agent skill from 0xsline/OpenChatCut. Create finished explainer videos from a topic, script, outline, voiceover, product logic, data, technical concept, course material, or reference assets. Use when the user wants narration, motion graphics, stock footage, generated visuals, or mixed visuals to explain an idea.

Its SKILL.md is about 2.8k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including reference files (for example `references/explainer-animation.md`).

It sits in Media & Creative, covering Video production, Text to speech and voice and Motion graphics. The repository describes itself as: Open-source, local-first conversational AI video editor with a professional multi-track timeline, Agent Skills, MCP integration, and Remotion rendering. The licence is AGPL-3.0.

When your agent uses it

  • The user wants narration
  • Motion graphics
  • Generated visuals
  • Mixed visuals to explain an idea

Example prompts

  • “/explainer-video”

Workflow steps

12 steps, taken from the first numbered list in SKILL.md.

  1. Read the project state, prompt, attached files, assets, transcript, and timeline.
  2. Identify the working labels
  3. Respect source structure. If the user provides structured material such as timestamps, numbered sections, slides/pages, scene labels…
  4. Ask only for missing details that change the result: topic or script, target length, audience, platform/aspect ratio, language/voice…
  5. If more than one detail is missing, load widget-forms and ask in one . Use text fields for topic/script/context and single-choice fields…
  6. Complete the preflight before writing visual treatments. The plan must have values for
  7. Animation Reference Gate. If any section may use motion graphics, animation, animated diagrams, data animation, mechanism visualization…
  8. Motion Graphic Direction Gate. If any section will generate MG/animation, load create-motion-graphics before asking the user to choose…
  9. Build a compact explainer plan only after the relevant gates above are complete
  10. For topic_only, write a short outline before drafting or generating. For script_or_outline, preserve the user's claims and meaning while…
  11. Voice Gate. For generated_tts, load voice before recommending voices, choosing a preset, or submitting TTS. If the user has not confirmed…
  12. Create or align narration only when needed. For generated_tts, draft or tighten section-level narration lines first; estimate whether they…

What it can do on your machine

Read from SKILL.md and the folder at commit 2e6f4a2. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Explainer Video loads about 2.8k tokens when it runs, and up to ~5.1k if it reads all its reference files. Until then it costs about 73 tokens; SKILL.md has 1,429 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~73
When it runs · the whole SKILL.md, loaded when a task matches
~2.8k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~5.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from 0xsline/OpenChatCut at commit 2e6f4a2, republished under its AGPL-3.0 licence (© 0xsline). 1,429 words, ~2,767 tokens.

Download SKILL.mdSave it as .claude/skills/explainer-video/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
explainer-video
description
Create finished explainer videos from a topic, script, outline, voiceover, product logic, data, technical concept, course material, or reference assets. Use when the user wants narration, motion graphics, stock footage, generated visuals, or mixed visuals to explain an idea.

Explainer Video

Use this workflow to turn information into a clear finished video. The information is the product: topic, script, logic, data, product mechanism, or voiceover. Visuals support understanding. Explainer Video owns the section plan, narration mode, timing order, assembly, and QA; create-motion-graphics is the helper workflow for direct Motion Graphic authoring and placement.

Workflow

  1. Read the project state, prompt, attached files, assets, transcript, and timeline.

  2. Identify the working labels:

    • explainer_start: topic_only, script_or_outline, voiceover_or_transcript, product_or_data, reference_assets, or direct_mg_animation_brief.
    • source_structure: free_topic, script_sections, timestamped_sections, slides_or_pages, existing_voiceover, uploaded_assets, product_or_data, or mixed.
    • narration_mode: generated_tts, existing_voiceover, transcript_only, or none.
    • visual_mode: motion_graphics, stock_or_uploaded_footage, generated_video_or_images, or mixed.
  3. Respect source structure. If the user provides structured material such as timestamps, numbered sections, slides/pages, scene labels, chapters, bullet outline, product points, transcript ranges, or voiceover sections, use that as the default planning scaffold. Merge, split, reorder, or relabel only when there is a clear production reason; explain the change and get user acceptance before treating it as the plan.

  4. Ask only for missing details that change the result: topic or script, target length, audience, platform/aspect ratio, language/voice, visual mode, tone, brand/style constraints, and whether to plan first or create directly.

  5. If more than one detail is missing, load widget-forms and ask in one <widget>. Use text fields for topic/script/context and single-choice fields for duration, platform, language/voice, and visual mode.

  6. Complete the preflight before writing visual treatments. The plan must have values for:

    • source_structure
    • narration_mode
    • visual_mode
    • animation_reference: read or not_needed
    • visual_direction_source: active Design Style, chosen preset, concrete user style/reference, accepted role anchor, explicit proceed-without-alignment, or not_needed
    • voice_selection: confirmed concrete preset, audition needed, or not_needed
    • timing_source: actual voiceover/transcript ranges, generated TTS duration, user timestamps, planned duration, or not_needed
  7. Animation Reference Gate. If any section may use motion graphics, animation, animated diagrams, data animation, mechanism visualization, abstract concept visualization, or MG overlays, read references/explainer-animation.md before writing those visual treatments. If the visual plan uses only stock footage, uploaded footage, generated live-action/video clips, or still images, mark animation_reference: not_needed and continue without loading it.

  8. Motion Graphic Direction Gate. If any section will generate MG/animation, load create-motion-graphics before asking the user to choose visual style. Use it to read the existing project visual language, align or confirm the Design Style, and directly author and place the Motion Graphic. Explainer Video still owns narration mode, section order, timing, assembly, and QA. Before final MG authoring, confirm visual direction through one of: active Design Style, catalog Design Style preset chosen from visual cards, concrete user style/reference, accepted role anchor, or explicit proceed-without-alignment. Treat broad hints such as "clean", "modern", "technical", "cinematic", or "tech style" as filters for preset selection, not as enough to generate final MGs. Do not invent text-only style choices before checking presets; assistant-written style options are fallback alignment, not a catalog preset.

  9. Build a compact explainer plan only after the relevant gates above are complete:

    • viewer promise or thesis
    • preserved or proposed sections
    • narration source and timing source
    • narration-to-visual map per section: narration text or time range, visual goal, visual treatment, source assets, and sync risk
    • assumptions and claims that need grounding
    • first visible result to create before batching
  10. For topic_only, write a short outline before drafting or generating. For script_or_outline, preserve the user's claims and meaning while tightening structure. For product_or_data, explain the mechanism or value without inventing unsupported claims. For direct_mg_animation_brief, do not force a broad explainer outline; inspect the provided script, assets, references, transcript, or style target, then create the requested MG section, intro, diagram, or overlay inside the same gates.

  11. Voice Gate. For generated_tts, load voice before recommending voices, choosing a preset, or submitting TTS. If the user has not confirmed a concrete voice preset, follow voice to read the curated voice list and show an audition widget first. Do not infer a voiceId from the content topic, language, gender, or broad style words.

  12. Create or align narration only when needed. For generated_tts, draft or tighten section-level narration lines first; estimate whether they fit target timing before submission, rewrite obvious mismatches, generate/place TTS by section only after the Voice Gate is complete, then read actual audio duration before any matching narration-backed MG/animation generation. Do not submit TTS and its matching MG in the same parallel batch. For existing_voiceover, do not regenerate narration; transcribe or read the audio and split it into section time ranges before generating matching visuals. For transcript_only, confirm whether the transcript should become TTS, captions, or only structure if ambiguous. For none, skip narration sync and plan visuals from the information structure and output rhythm.

  13. Produce visuals section by section. For MG/animation, verify the animation reference has been read, visual direction is confirmed, and create-motion-graphics has been used for direct authoring and placement before final generation or batching. For generated-TTS sections, even the first representative section MG must wait until that section's actual TTS duration is known. For narration-backed MG/animation, duration must come from the matching narration's actual audio duration when available, not from script estimates. For stock, uploaded, generated-video, or mixed visual sections, inspect/select the visual source first and use it only when it supports the section. When style or correctness is uncertain, create the first representative section or shot before batching only after required narration timing exists; a pre-audio style proof requires explicit user approval and must be labeled style-only, not treated as a section MG or placed as final timeline content.

  14. Assemble the timeline with narration, visuals, captions when useful, background music, and section pacing. For narration-backed MG/animation sections, align narration and matching visuals to the same start time and cover the full narration section unless the visual plan intentionally changes shots within that section.

  15. Run Narration-Visual Sync QA before done. For each narration-backed section, check whether spoken content matches the visual, whether visual duration covers narration, whether visual information density supports the spoken point, and whether transitions happen too early or too late. Fix failures before delivery by tightening narration, splitting the section, extending/regenerating visuals, adjusting timing, or asking the user to choose a tradeoff.

  16. Final QA before done: topic clarity, factual grounding, visual-mode fit, narration coverage, timing, caption readability, audio mix, timeline continuity, narration-visual sync, and export readiness.

Show full SKILL.md (396 more words)Show less

Rules

  • Explain the idea; do not merely decorate narration with icons or subtitles.
  • Do not invent facts, prices, medical claims, performance claims, legal claims, or product guarantees.
  • Do not use existing footage as filler when it is unrelated to the explanation.
  • Do not average every uploaded asset into the video. Use assets only when they support a beat.
  • Do not keep asking after enough information exists to make the first visible result.
  • Do not put full spoken sentences on screen. Use labels, numbers, short questions, or section titles.
  • Do not rewrite structured user inputs into a different section plan without explaining why and getting user acceptance.
  • Do not output MG/animation visual treatments before the animation reference gate has been resolved.
  • Do not generate final MG/animation before the visual-direction gate is resolved.
  • Do not satisfy the visual-direction gate with ad hoc style choices you invented before loading create-motion-graphics and checking its project visual-language intake. A preset means a catalog Design Style preset shown through visual cards, not a text label.
  • Do not force generated TTS when the user already has a usable voiceover or does not want narration.
  • Do not call submit_voice before loading voice and confirming a concrete voice preset.
  • Do not submit a narration-backed MG/animation in parallel with the TTS that should time it.
  • Do not submit final narration-backed MG from estimated script duration. Narration-backed MG must use the matching narration text or time range and the real audio duration when available.
  • Do not let create-motion-graphics override the Explainer Video source structure, narration mode, section order, timing, assembly, or QA workflow.

Plan Format

Use this compact format when planning:

  • explainer_start: starting point label
  • source_structure: input scaffold and whether it is preserved or changed
  • narration_mode: generated_tts, existing_voiceover, transcript_only, or none
  • output: platform, aspect ratio, target length, language, and voice when relevant
  • viewer_promise: what the viewer will understand by the end
  • preflight: animation reference status, visual-direction source, voice selection, and timing source
  • sections: beat, narration text or time range, visual goal, visual treatment, source assets, sync risk
  • first_visible_result: the first section or shot to create before batching; mark it as non-final if real narration timing is not known
  • sync_check: for narration-backed MG, note the timing source, narration duration, MG duration, match result, and any fix applied

When reporting execution, include created timeline names, narration mode, visual modes used, assumptions, sync fixes applied, and what to review first.

© 0xsline, AGPL-3.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (references) in src/agent/skills/explainer-video of 0xsline/OpenChatCut.

  • SKILL.md
  • references/explainer-animation.md

Open the folder on GitHubat commit 2e6f4a2

Compare with similar skills

Explainer Video next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Explainer Video compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Explainer Video this skill0xsline/OpenChatCut2.2k—~2.8kAutomated safety check: PassAGPL-3.0
Paw Cra Agent Video Producerpawbytes/skill-suites110—~2.3kAutomated safety check: PassMIT
Hyperframes Ad DirectorOneWave-AI/claude-skills323—~988Automated safety check: PassMIT
JianYing Editor Automationluoluoluo22/jianying-editor-skill3.8k—~2.2kAutomated safety check: PassApache-2.0
Video TalkcraftVincentwei1021/video-talkcraft1.4k—~8.7kAutomated safety check: NotesCustom licence
Anything2explainerVincentwei1021/anything2explainer2.3k—~2.7kAutomated safety check: PassCustom licence

Similar skills

  • Paw Cra Agent Video Producer

    pawbytes/skill-suites

    Video production specialist for short-form, long-form, episodic, and motion graphics video.

    110 GitHub stars~2.3k tokensUpdated 4 days ago
    Media & CreativeAuto-check passed
  • Hyperframes Ad Director

    OneWave-AI/claude-skills

    Turn a marketing brief or product offer into a finished short-form video ad built in HyperFrames — hook, script, shot-by-shot storyboard, scene-by-scene HTML composition, captions, and voiceover.

    323 GitHub stars~988 tokensUpdated 6 days ago
    Media & CreativeAuto-check passed
  • JianYing Editor Automation

    luoluoluo22/jianying-editor-skill

    Automates JianYing Pro video editing through a Python wrapper, JyWrapper, covering drafts, media import, subtitles, screen recording, voiceover and export.

    3.8k GitHub stars~2.2k tokensUpdated 26 days ago
    Media & CreativeAuto-check passed
  • Video Talkcraft

    Vincentwei1021/video-talkcraft

    终极口播视频 skill:中文口播稿 + 成品配音 → CPU 字级时间戳 → SHOTBOOK 层矩阵分镜 → Remotion 电影感成片(横屏默认/竖屏)。当用户要"做口播视频"、"解说/科普视频"、"把文案变成视频"、"给配音配画面动效"时使用。默认使用成品配音,可选 Fish Audio 从稿子合成配音与时间戳;数字人生成技术不在本 skill…

    1.4k GitHub stars~8.7k tokensUpdated 6 days ago
    Media & CreativeAuto-check: notes
  • Anything2explainer

    Vincentwei1021/anything2explainer

    给一个主题,产出一条黑底 MG 风格(幕底可选星点或点阵波)、有配音字幕章节进度条的科普讲解视频(中文或英文;Remotion 代码动画;时长由用户定,常用 3–5 分钟)。内含可编译模板、图元库、配音/分镜/渲染工具、风格与动效规范、多 agent 分工协议与 QC 判据,以及一条完整样片(《RAG 与知识库》)作为质量标尺。Turn any topic into a narrated…

    2.3k GitHub stars~2.7k tokensUpdated 19 days ago
    Media & CreativeAuto-check passed
  • Qiaomu Cut

    joeseesun/qiaomu-cut-skill

    把一句话需求转成可复现、可验收视频工程的乔木智能剪辑导演。Use when the user asks to create, plan, edit, remix, explain, narrate, subtitle, animate, composite, or render a video—including one-line requests such as “制作一个科普视频:介绍…

    369 GitHub stars~6.8k tokensUpdated 10 days ago
    Media & CreativeAuto-check: notes

More from 0xsline/OpenChatCut

All 31 skills in this repo
  • OpenChatCut Video Editing

    0xsline/OpenChatCut

    Connects an MCP-capable agent to the local OpenChatCut video editor to inspect and edit projects through draft edit sessions, with manual approval by default.

    2.2k GitHub starsUsed in 1 repo~655 tokens
    Auto-check passed
  • Video Shader Generator

    0xsline/OpenChatCut

    Generates WebGL shaders for video effects, transitions, masks and color grades in the OpenChatCut editor, trying built-in catalog effects such as zoom before making anything new.

    2.2k GitHub starsUsed in 1 repo~3.2k tokens
    Auto-check passed
  • AI Image Generation

    0xsline/OpenChatCut

    Generates still images through the submit_image tool, choosing among Fal.ai, gpt-image-2, nano-banana, MiniMax image-01 and Grok Imagine by configured keys.

    2.2k GitHub stars~1.3k tokensUpdated yesterday
    Auto-check passed
  • Livestream to Clips

    0xsline/OpenChatCut

    Cuts a livestream recording into evidence-backed, platform-ready clips by combining transcript, visual, audio and genre-specific signals.

    2.2k GitHub stars~2.7k tokensUpdated yesterday
    Auto-check passed
  • Music Generation

    0xsline/OpenChatCut

    Generates instrumentals, songs, soundtracks and covers through Mureka, MiniMax, Atlas Cloud or Sonilo using the `submit_music` tool.

    2.2k GitHub stars~1.1k tokensUpdated yesterday
    Auto-check passed
  • AI Video Generation

    0xsline/OpenChatCut

    Submits AI video generation jobs to Fal.ai, Seedance, Kling, MiniMax Hailuo, xAI Grok Imagine or OFox for text-to-video, image-to-video, transitions and clip extension.

    2.2k GitHub stars~4.3k tokensUpdated yesterday
    Auto-check passed

Questions about Explainer Video

What does Explainer Video do?

Create finished explainer videos from a topic, script, outline, voiceover, product logic, data, technical concept, course material, or reference assets. Explainer Video is an agent skill from 0xsline/OpenChatCut. Create finished explainer videos from a topic, script, outline, voiceover, product logic, data, technical concept, course material, or reference assets.

When should I use Explainer Video?

Explainer Video fits situations like: the user wants narration; motion graphics; generated visuals; mixed visuals to explain an idea.

How do I install Explainer Video in Claude Code?

Run `npx skills add 0xsline/OpenChatCut --skill explainer-video -a claude-code`. Or copy the skill folder (src/agent/skills/explainer-video in 0xsline/OpenChatCut) into .claude/skills/explainer-video in your project. Claude Code loads it when a task matches its description.

How do I install Explainer Video in Codex?

Run `npx skills add 0xsline/OpenChatCut --skill explainer-video -a codex`. Or copy the skill folder (src/agent/skills/explainer-video in 0xsline/OpenChatCut) into .agents/skills/explainer-video in your project. Codex loads it when a task matches its description.

Can I use Explainer Video in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add 0xsline/OpenChatCut --skill explainer-video -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/explainer-video, .gemini/skills/explainer-video, .github/skills/explainer-video and .opencode/skills/explainer-video in your project.

What does Explainer Video need to run?

SKILL.md names no scripts, command-line tools or credentials: Explainer Video is instructions for the agent only.

Does Explainer Video access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Explainer Video safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Explainer Video use?

Explainer Video is published under the AGPL-3.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Explainer Video use?

About 2.8k tokens (SKILL.md is roughly 11k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 2.4k tokens, read only when the agent opens those files.

What are the alternatives to Explainer Video?

Skills that share tags, products or a category with Explainer Video: Paw Cra Agent Video Producer (pawbytes/skill-suites, 110 stars), Hyperframes Ad Director (OneWave-AI/claude-skills, 323 stars), JianYing Editor Automation (luoluoluo22/jianying-editor-skill, 3.8k stars) and Video Talkcraft (Vincentwei1021/video-talkcraft, 1.4k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Explainer Video?

0xsline (a GitHub user) maintains it in 0xsline/OpenChatCut, which has 2,178 GitHub stars. The repository holds 31 skills in this directory. The repository was last updated on October 7, 2026.

Source: 0xsline/OpenChatCut on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.