Agent skill

Video Compose

by CatCatUncle in CatCatUncle/openworkbuddy

图文成片——脚本→分镜卡片图→TTS配音→ffmpeg 拼装成带字幕的竖版/横版视频(口播、知识分享、带货讲解). An agent skill from CatCatUncle/openworkbuddy.

Custom licenceAuto-check passedMedia & Creative

Install Video Compose

skills CLI
$ npx skills add CatCatUncle/openworkbuddy --skill video-compose -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install CatCatUncle/openworkbuddy video-compose --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/CatCatUncle/openworkbuddy.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/video-compose .claude/skills/video-compose && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
video-compose
GitHub stars
277
Token cost
~869 tokens
SKILL.md length
252 words
Files
1
Skills in repo
22
Repo updated
First seen
Licence
Custom licence

At a glance

图文成片——脚本→分镜卡片图→TTS配音→ffmpeg 拼装成带字幕的竖版/横版视频(口播、知识分享、带货讲解). An agent skill from CatCatUncle/openworkbuddy.

  • Works in 3 steps: run_shell 执行 ffmpeg -version:没装就直接告诉用户… → 语音合成(text_to_speech)没配渠道时:照常出片但改为「无声+大字幕」… → 分镜超过 6…
  • Tasks that involve Video production
  • SKILL.md covers 适用场景, 前置检查(先做,缺了就早说), 流程 and 硬约束
  • Calls ffmpeg, brew and winget

What it does

Video Compose is an agent skill from CatCatUncle/openworkbuddy. 图文成片——脚本→分镜卡片图→TTS配音→ffmpeg 拼装成带字幕的竖版/横版视频(口播、知识分享、带货讲解)

Its SKILL.md is about 870 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering Video production and Text to speech and voice. It works with FFmpeg. The repository describes itself as: Open-source Claude Cowork / Codex alternative — a local-first AI office agent that turns one request into real PPTX, DOCX, XLSX and HTML files. Runs DeepSeek, Qwen, Doubao…

When your agent uses it

  • Tasks that involve Video production
  • Tasks that involve Text to speech and voice

Example prompts

  • “/video-compose”

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. run_shell 执行 ffmpeg -version:没装就直接告诉用户 brew install ffmpeg(mac)/ winget install ffmpeg(win),并先把分镜图做完,配音和拼装留到装好后再跑(按句配音要靠 ffmpeg…
  2. 语音合成(text_to_speech)没配渠道时:照常出片但改为「无声+大字幕」样式,并明说配音跳过的原因。
  3. 分镜超过 6 段就先让用户把额度调上去:一段要走出图、拼片两步(配音一次调用整批出),6 段也有十几步,而单独用本技能时默认上限是 25 步 / 30 分钟(走 promo-video 配方时,开头表单一交上限就按配方放宽了,跳过这一条)。撞上限断在最后一步 concat…

What it can do on your machine

Read from SKILL.md and the folder at commit 3e4365b. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • ffmpeg
    • brew
    • winget

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Video Compose loads about 869 tokens when it runs. Until then it costs about 18 tokens; SKILL.md has 252 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~18
When it runs · the whole SKILL.md, loaded when a task matches
~869

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 252 words (~869 tokens).

“「把这篇文章做成视频」「做一条口播/知识分享视频」。产出 = 一条 mp4(分镜卡片轮播 + 配音 + 字幕)+ 发布文案。走的是「图文成片」路线(信息密度高、成本为零、可控性强),不是 AI 生成实拍画面——需要实拍感画面时才用 generate_video 生成个别镜头素材。”

— opening of SKILL.md by CatCatUncle, Custom licence
name
video-compose

Read the full SKILL.md on GitHub

Files

Just SKILL.md in skills/video-compose of CatCatUncle/openworkbuddy.

Open the folder on GitHubat commit 3e4365b

Compare with similar skills

Video Compose next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Video Compose compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Video Compose this skillCatCatUncle/openworkbuddy277—~869Automated safety check: PassCustom licence
Stage EditOrkas-AI/Orkas-VideoStudio498—~2.4kAutomated safety check: PassMIT
Proof Videoopenclaw/openclaw392k—~2.4kAutomated safety check: PassMIT
Stage AssembleOrkas-AI/Orkas-VideoStudio498—~3.4kAutomated safety check: PassMIT
Cadence Videopa001024/dna-builder136—~703Automated safety check: PassMIT
Video Podcast Makerdtsola/xiaoyaosearch1k1 repos~3.4kAutomated safety check: PassMIT

Similar skills

  • Stage Edit

    Orkas-AI/Orkas-VideoStudio

    Intelligent editing of real user-supplied footage—understand it with transcript/inspected-frame/scene/silence/quality evidence, then choose deterministic timeline operations or a constrained…

    498 GitHub stars~2.4k tokensUpdated 16 days ago
    Media & CreativeAuto-check passed
  • Proof Video

    openclaw/openclaw

    Add subtitles, captions, narration cues, or zoom to a proof video or PR recording using repo-local capture helpers and a system ffmpeg renderer.

    392k GitHub stars~2.4k tokensUpdated today
    Media & CreativeAuto-check passed
  • Stage Assemble

    Orkas-AI/Orkas-VideoStudio

    Deterministically assemble an approved cross-modal EDL (plan.json) into a finished video — produce each segment (edit/compose/generate/provided, delegated to its line), then assemble in ffmpeg…

    498 GitHub stars~3.4k tokensUpdated 16 days ago
    Media & CreativeAuto-check passed
  • Cadence Video

    pa001024/dna-builder

    用 cadence 框架 做"代码渲染、音画同步"的视频:两条并列入口(脚本+TTS 实测时长 / 歌曲分析对齐)写同一份 project.json,可混合;带网页端时间线编辑器(拖拽剪辑、改字重生成语音、场景与帧特效标注)和无头 Chrome+ffmpeg…

    136 GitHub stars~703 tokensUpdated today
    Media & CreativeAuto-check passed
  • Video Podcast Maker

    dtsola/xiaoyaosearch

    Turns a topic into a 4K horizontal video podcast through research, scripting, text-to-speech, Remotion rendering and background music, and can learn styles from references.

    1k GitHub starsUsed in 1 repo~3.4k tokens
    Media & CreativeAuto-check passed
  • KrillinAI Render Vertical

    krillinai/OpenCreator

    Renders a source video as a portrait video with the KrillinAI CLI, with short bilingual subtitles or a dubbed audio track, then checks the result.

    13k GitHub starsUsed in 1 repo~527 tokens
    Media & CreativeAuto-check passed

More from CatCatUncle/openworkbuddy

All 22 skills in this repo
  • Short Drama

    CatCatUncle/openworkbuddy

    AI 真生成画面的竖版短剧:分镜表→定妆照→首帧→图生视频→配音→拼片配乐,改一镜只重算一镜. An agent skill from CatCatUncle/openworkbuddy.

    277 GitHub stars~1.8k tokensUpdated today
    Auto-check passed
  • Character Photo Studio

    CatCatUncle/openworkbuddy

    批量生成「角色扮演写真/剧照」类 AI 图片——把中国古典名著/神话角色的造型要素拆成 prompt 模板,套上指定影调预设(CCD数码、复古胶片、黑白电影、电视剧照、vlog自拍等)与画幅组合,一次交付多张成品图。适用于用户说「给我生成某角色的写真/剧照/拼贴照」的场景。

    277 GitHub stars~725 tokensUpdated today
    Auto-check passed
  • Data Viz

    CatCatUncle/openworkbuddy

    数据可视化与画图——流程图/架构图/时序图/数据图表,用 gendiagram 工具一键渲染成 SVG+PNG,或生成交互式 ECharts 页面

    277 GitHub stars~494 tokensUpdated today
    Auto-check passed
  • DOCX

    CatCatUncle/openworkbuddy

    用 docx 库生成排版规整的 Word 文档(报告、方案、公文、合同草稿),以及改写已有 .docx 里的文字. An agent skill from CatCatUncle/openworkbuddy.

    277 GitHub stars~1.1k tokensUpdated today
    Auto-check passed
  • Feishu Doc

    CatCatUncle/openworkbuddy

    创建/写入飞书云文档的完整 API 姿势(含插图)。要"发到飞书/建飞书文档"优先用 feishudoccreate 工具;需要手写 API(插图片/表格/自定义凭证)或排查权限、块结构报错时加载本技能。

    277 GitHub stars~750 tokensUpdated today
    Auto-check passed
  • Lark CLI

    CatCatUncle/openworkbuddy

    用命令行操作飞书/Lark——发消息、建群、读写云文档与多维表格、日历日程、任务、审批、邮件、妙搭应用、白板、知识库。凡是"发到飞书/在飞书里建/查我的飞书"的任务都用它

    277 GitHub stars~733 tokensUpdated today
    Auto-check passed

Works with

Questions about Video Compose

What does Video Compose do?

图文成片——脚本→分镜卡片图→TTS配音→ffmpeg 拼装成带字幕的竖版/横版视频(口播、知识分享、带货讲解). An agent skill from CatCatUncle/openworkbuddy. Video Compose is an agent skill from CatCatUncle/openworkbuddy.

When should I use Video Compose?

Video Compose fits situations like: tasks that involve Video production; tasks that involve Text to speech and voice.

How do I install Video Compose in Claude Code?

Run `npx skills add CatCatUncle/openworkbuddy --skill video-compose -a claude-code`. Or copy the skill folder (skills/video-compose in CatCatUncle/openworkbuddy) into .claude/skills/video-compose in your project. Claude Code loads it when a task matches its description.

How do I install Video Compose in Codex?

Run `npx skills add CatCatUncle/openworkbuddy --skill video-compose -a codex`. Or copy the skill folder (skills/video-compose in CatCatUncle/openworkbuddy) into .agents/skills/video-compose in your project. Codex loads it when a task matches its description.

Can I use Video Compose in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add CatCatUncle/openworkbuddy --skill video-compose -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/video-compose, .gemini/skills/video-compose, .github/skills/video-compose and .opencode/skills/video-compose in your project.

What does Video Compose need to run?

Going by SKILL.md and its folder, Video Compose needs the command-line tools its instructions call (ffmpeg, brew and winget).

Does Video Compose access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Video Compose safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Video Compose use?

Video Compose has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Video Compose use?

About 869 tokens (SKILL.md is roughly 3.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Video Compose?

Skills that share tags, products or a category with Video Compose: Stage Edit (Orkas-AI/Orkas-VideoStudio, 498 stars), Proof Video (openclaw/openclaw, 392k stars), Stage Assemble (Orkas-AI/Orkas-VideoStudio, 498 stars) and Cadence Video (pa001024/dna-builder, 136 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Video Compose?

CatCatUncle (a GitHub user) maintains it in CatCatUncle/openworkbuddy, which has 277 GitHub stars. The repository holds 22 skills in this directory. The repository was last updated on October 7, 2026.

Source: CatCatUncle/openworkbuddy on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.