Agent skill

Auto Short Video

by ZJU-REAL in ZJU-REAL/Easel

一句话主题 → 成品短视频:自动串联 文案→配图/AI视频→配音→字幕→BGM→合成,把 Easel 制作层零件编排成一条'一键出片'流水线。单条视频、口播/资讯向,画面默认逐句配图 + Ken Burns 缓动,需要动态时才逐段图生视频。当用户说 一键生成视频、自动做短视频、主题生成视频、帮我做条视频、口播视频一条龙、自动出片、短视频一键生成 时使用。有剧情/角色/对白/反转/多集的短剧改用…

Apache-2.0Auto-check passedMedia & Creative

Install Auto Short Video

skills CLI
$ npx skills add ZJU-REAL/Easel --skill auto-short-video -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install ZJU-REAL/Easel auto-short-video --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/ZJU-REAL/Easel.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/openclaw/auto-short-video .claude/skills/auto-short-video && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
auto-short-video
GitHub stars
3.4k
Token cost
~748 tokens
SKILL.md length
134 words
Files
3 (incl. scripts)
Skills in repo
114
Repo updated
First seen
Licence
Apache-2.0

At a glance

一句话主题 → 成品短视频:自动串联 文案→配图/AI视频→配音→字幕→BGM→合成,把 Easel 制作层零件编排成一条'一键出片'流水线。单条视频、口播/资讯向,画面默认逐句配图 + Ken Burns 缓动,需要动态时才逐段图生视频。当用户说 一键生成视频、自动做短视频、主题生成视频、帮我做条视频、口播视频一条龙、自动出片、短视频一键生成 时使用。有剧情/角色/对白/反转/多集的短剧改用…

  • Works in 7 steps: 写脚本分镜:用 video-script 把主题写成口播文案,拆成 N… → 生成画面(每个分镜一张图/一段片) → 配音:tts-voiceover 把文案合成口播(同时出 SRT)。配了… → …
  • Tasks that involve AI video generation
  • SKILL.md covers 输入, 输出, 执行步骤(按需裁剪,缺 API key 的环节自动降级或询问) and 编排原则, plus 1 more section
  • Runs Python scripts from its folder; calls python

What it does

Auto Short Video is an agent skill from ZJU-REAL/Easel. 一句话主题 → 成品短视频:自动串联 文案→配图/AI视频→配音→字幕→BGM→合成,把 Easel 制作层零件编排成一条'一键出片'流水线。单条视频、口播/资讯向,画面默认逐句配图 + Ken Burns 缓动,需要动态时才逐段图生视频。当用户说 一键生成视频、自动做短视频、主题生成视频、帮我做条视频、口播视频一条龙、自动出片、短视频一键生成 时使用。有剧情/角色/对白/反转/多集的短剧改用 short-drama(每镜强制图生视频、不用静态图冒充);只写脚本用 video-script;只生成单个片段用 ai-video-gen。

Its SKILL.md is about 750 tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files, including scripts (for example `EASEL-META.md` and `scripts/assemble.py`).

It sits in Media & Creative, covering AI video generation, Video scripts and shorts and Text to speech and voice. The repository describes itself as: An open-source AI agent for social media — discover trends, create content, publish everywhere, and learn what works across Xiaohongshu, Douyin, Zhihu, Bilibili, and more.🎨一个开源的… The licence is Apache-2.0.

When your agent uses it

  • Tasks that involve AI video generation
  • Tasks that involve Video scripts and shorts
  • Tasks that involve Text to speech and voice

Example prompts

  • “/auto-short-video”

Requirements

  • Python 3

Workflow steps

7 steps, taken from the first numbered list in SKILL.md.

  1. 写脚本分镜:用 video-script 把主题写成口播文案,拆成 N 句(每句一个分镜),每句配一个画面描述。
  2. 生成画面(每个分镜一张图/一段片)
  3. 配音:tts-voiceover 把文案合成口播(同时出 SRT)。配了 VOICE_PROVIDER 默认走闭源好嗓子(有情感、像真人),没 key 才退 edge(机械)——想要口播不"生硬"务必配闭源 key。不需要配音才跳过。
  4. 字幕:用 TTS 附带的 SRT,或对配音跑 auto-subtitle;也可让 assemble 用各分镜 caption 自动生成。
  5. BGM:ai-music 生成,或用用户提供的音乐。可选。
  6. 合成成片:把上面的素材写成 storyboard JSON,调合成器
  7. 交付:产出 final.mp4,附一句制作说明(用了哪些环节、哪些降级了)。

What it can do on your machine

Read from SKILL.md and the folder at commit 278f420. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Auto Short Video loads about 748 tokens when it runs. Until then it costs about 74 tokens; SKILL.md has 134 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~74
When it runs · the whole SKILL.md, loaded when a task matches
~748

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from ZJU-REAL/Easel at commit 278f420, republished under its Apache-2.0 licence (© ZJU-REAL). 134 words, ~748 tokens.

Download SKILL.mdSave it as .claude/skills/auto-short-video/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.
name
auto-short-video
description
一句话主题 → 成品短视频:自动串联 文案→配图/AI视频→配音→字幕→BGM→合成,把 Easel 制作层零件编排成一条'一键出片'流水线。**单条视频、口播/资讯向,画面默认逐句配图 + Ken Burns 缓动,需要动态时才逐段图生视频**。当用户说 一键生成视频、自动做短视频、主题生成视频、帮我做条视频、口播视频一条龙、自动出片、短视频一键生成 时使用。**有剧情/角色/对白/反转/多集的短剧改用 short-drama(每镜强制图生视频、不用静态图冒充);只写脚本用 video-script;只生成单个片段用 ai-video-gen。**
layer
produce

一键短视频(端到端编排)

输入一个主题,自动产出一条短视频。本 SKILL 是编排层:把已有制作零件串成流水线—— 文案(video-script) → 逐句配图(ai-image-gen)或片段(ai-video-gen) → 配音(tts-voiceover) → 字幕(auto-subtitle) → BGM(ai-music) → 合成(scripts/assemble.py)。 沉淀自 Pixelle-Video / MoneyPrinterTurbo 的自动短视频引擎思路。

输入

  • 主题 / 文案(必填)
  • 可选:目标时长、风格、是否要配音/字幕/BGM、配图用 AI 生图还是用户素材
  • 画幅确认硬门:用户或上游任务未明确横版/竖版(或 16:9/9:16/具体分辨率)时,制作/付费调用前必须追问并等确认;不得从平台、Profile 或默认值静默推断,已明确则不重复问

输出

成品短视频写入 outputs/主题名/final.mp4;分镜图、配音、字幕和 storyboard 写入 outputs/主题名/assets/。

执行步骤(按需裁剪,缺 API key 的环节自动降级或询问)

  1. 写脚本分镜:用 video-script 把主题写成口播文案,拆成 N 句(每句一个分镜),每句配一个画面描述。

  2. 生成画面(每个分镜一张图/一段片):

    • 有图像 API key → ai-image-gen 逐句 text2img(按已确认画幅)
    • 要动态 → ai-video-gen text2video/image2video
    • 用户自带素材 → 用 image-editing pad 到已确认画幅
    • 都没有 → 按已确认画幅选图卡(竖版用 card-xiaohongshu/poster-hero,横版用 card-quote)再 pad,避免画幅错配。
  3. 配音:tts-voiceover 把文案合成口播(同时出 SRT)。配了 VOICE_PROVIDER 默认走闭源好嗓子(有情感、像真人),没 key 才退 edge(机械)——想要口播不"生硬"务必配闭源 key。不需要配音才跳过。

  4. 字幕:用 TTS 附带的 SRT,或对配音跑 auto-subtitle;也可让 assemble 用各分镜 caption 自动生成。

  5. BGM:ai-music 生成,或用用户提供的音乐。可选。

  6. 合成成片:把上面的素材写成 storyboard JSON,调合成器:

    bash
    python skills/openclaw/auto-short-video/scripts/assemble.py assemble \
      --storyboard outputs/主题名/assets/storyboard.json \
      -o outputs/主题名/final.mp4

    storyboard 结构(图/片二选一,narration/bgm/subtitle 可选,缺 duration 时按配音均分):

    json
    {
      "size": "<已确认尺寸,如1080x1920或1920x1080>",
      "image_motion": "ken-burns",
      "shots": [
        {"image": "outputs/主题名/assets/shot1.png", "duration": 3, "caption": "第一句", "motion": "static"},
        {"video": "outputs/主题名/assets/clip2.mp4", "caption": "第二句"}
      ],
      "narration": "outputs/主题名/assets/voice.mp3",
      "bgm": "outputs/主题名/assets/bgm.mp3",
      "subtitle": "outputs/主题名/assets/voice.srt"
    }

    image_motion 设整条图片默认运动,单镜 motion 可覆写:照片用 ken-burns,含文字的 slide/图表/界面必须用 static(等比缩放 + 补边,不裁切、不平移)。 合成器自动做:按 image_motion 生成静帧或 Ken Burns、补边到画幅、拼接、配音+BGM 混音(BGM 自动压低)、烧录字幕。

  7. 交付:产出 final.mp4,附一句制作说明(用了哪些环节、哪些降级了)。

编排原则

  • 零件可缺:缺图像/视频/TTS API key 的环节自动降级(图卡兜底 / 跳过配音),不阻断整体,并如实告知用户降级了什么。
  • 先出 Plan:涉及多个付费 API(生图/生视频/生乐)时,先向用户说明将调用哪些、大致耗时/花费,确认后再跑。
  • 中间产物留档:分镜图、配音、字幕和 storyboard 都写进 outputs/主题名/assets/,方便单独替换后重新合成。

Profile 感知

  • 有 Profile:从 style.md 取调性/视觉风格贯穿文案与配图 prompt;platforms.md 只用于给出画幅/时长建议,画幅仍须确认;preferences.md 红线过滤。
  • 无 Profile:先确认横版/竖版,再用通用口播风格。

© ZJU-REAL, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 2 other files (scripts) in skills/openclaw/auto-short-video of ZJU-REAL/Easel.

  • SKILL.md
  • EASEL-META.md
  • scripts/assemble.py

Open the folder on GitHubat commit 278f420

Compare with similar skills

Auto Short Video next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Auto Short Video compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Auto Short Video this skillZJU-REAL/Easel3.4k—~748Automated safety check: PassApache-2.0
Video Cover Imageitwanger/toBeBetterJavaer18k—~3.3kAutomated safety check: PassNone
Narrator AI CLINarratorAI-Studio/narrator-ai-cli-skill3k—~4.5kAutomated safety check: PassMIT
Ergo Remotion Videoitwanger/toBeBetterJavaer18k—~1.1kAutomated safety check: PassNone
Wedding Video Guided Wizardaaronyi97/wedding-video-guided-wizard310—~1kAutomated safety check: PassMIT
RunninghubHM-RunningHub/OpenClaw_RH_Skills142—~1.6kAutomated safety check: PassApache-2.0

Similar skills

  • Video Cover Image

    itwanger/toBeBetterJavaer

    Generate matched 3:4, 16:9, and 4:3 short-video cover images from toBeBetterJavaer video scripts or AI/Java technical topics.

    18k GitHub stars~3.3k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Narrator AI CLI

    NarratorAI-Studio/narrator-ai-cli-skill

    AI 电影/短剧解说视频自动生成(AI 解说大师 CLI Skill)。当用户需要创建电影解说视频、短剧解说、影视二创、AI 配音旁白视频、film commentary、video narration、drama dubbing、movie narration 时触发。内置电影素材库、BGM、多语种配音、解说模板。通过 narrator-ai-cli 命令行实现:搜片→选模板→选…

    3k GitHub stars~4.5k tokensUpdated 3 mo ago
    Media & CreativeAuto-check passed
  • Ergo Remotion Video

    itwanger/toBeBetterJavaer

    把口播稿做成二哥风格的 Remotion 视频,包括整理视频用稿、火山 TTS 配音、音画对齐、逐章动画预览和导出带配音的 MP4。用户说“做视频”“口播稿转视频”“Remotion”“继续做下一章”“出片”“渲染”“改读音”“配音读错了”,或给出 docs/src/ai/video/ 下的稿子要做成视频时使用。共享工具、配置和素材在…

    18k GitHub stars~1.1k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Wedding Video Guided Wizard

    aaronyi97/wedding-video-guided-wizard

    Guide a creator through a real couple's custom wedding video, from a shareable story intake card and Kimi writing pack through narration, external GPT image prompts, image-to-video packs, music and…

    310 GitHub stars~1k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Runninghub

    HM-RunningHub/OpenClaw_RH_Skills

    Generate images, videos, audio, and 3D models via RunningHub API (420 endpoints) and run any RunningHub AI Application (custom ComfyUI workflow) by webappId.

    142 GitHub stars~1.6k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • AI Video Production Assistant

    wanghui2323/ai-video-maker

    Turns an idea, article, outline or audio file into a sourced, reviewable AI video, tracking whether narration uses a human, synthetic or cloned voice.

    101 GitHub stars~924 tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed

More from ZJU-REAL/Easel

All 114 skills in this repo
  • Gzh Design

    ZJU-REAL/Easel

    微信公众号文章排版引擎:把 Markdown / Word(.docx) / PDF / 纯文本转成可直接粘贴进公众号编辑器的 HTML,自动章节编号、关键词标记、引言卡、目录、代码块、图片/GIF、作者签名;主题从 references/theme-index.md…

    3.4k GitHub stars~2.1k tokensUpdated today
    Auto-check passed
  • 微信公众号文章自动创作与发布工具。给定参考文章、文字或文档,自动搜索整理全网相关信息、生成图文并茂的公众号文章,并发布到微信公众号草稿箱。特别强调反 AI 检测写作。

    3.4k GitHub stars~1.8k tokensUpdated today
    Auto-check passed
  • Card Design

    ZJU-REAL/Easel

    社媒卡片视觉设计系统:提供配色、中文字体层级、满画幅布局、品类骨架和死空白/密度质检,避免模板化 PPT 与廉价 AI 感。

    3.4k GitHub stars~657 tokensUpdated today
    Auto-check passed
  • Ecom Details Image

    ZJU-REAL/Easel

    生成电商商品视觉方案:主图概念、场景图、详情页视觉方向和 AI 生图 Prompt. An agent skill from ZJU-REAL/Easel.

    3.4k GitHub stars~1.1k tokensUpdated today
    Auto-check: notes
  • Infographic

    ZJU-REAL/Easel

    将数据或文字内容转化为可视化信息图,支持静态(AntV)和动画 GIF 两种模式。当用户需要制作信息图、数据可视化、流程图、对比图、动画图表、GIF 图表、思维导图、SWOT 分析图时调用。本地渲染信息图/GIF 动画;要单张静态图片 URL 用 chart-visualization,要 CSV/JSON→整页报告用 data-report

    3.4k GitHub stars~643 tokensUpdated today
    Auto-check passed
  • Novel Writer

    ZJU-REAL/Easel

    长篇小说/网文连载创作:从世界观、人设和三级大纲写到逐章正文,并用文件化状态维护伏笔、前情和跨章一致性. An agent skill from ZJU-REAL/Easel.

    3.4k GitHub stars~1k tokensUpdated today
    Auto-check passed

Questions about Auto Short Video

What does Auto Short Video do?

一句话主题 → 成品短视频:自动串联 文案→配图/AI视频→配音→字幕→BGM→合成,把 Easel 制作层零件编排成一条'一键出片'流水线。单条视频、口播/资讯向,画面默认逐句配图 + Ken Burns 缓动,需要动态时才逐段图生视频。当用户说 一键生成视频、自动做短视频、主题生成视频、帮我做条视频、口播视频一条龙、自动出片、短视频一键生成 时使用。有剧情/角色/对白/反转/多集的短剧改用…. Auto Short Video is an agent skill from ZJU-REAL/Easel.

When should I use Auto Short Video?

Auto Short Video fits situations like: tasks that involve AI video generation; tasks that involve Video scripts and shorts; tasks that involve Text to speech and voice.

How do I install Auto Short Video in Claude Code?

Run `npx skills add ZJU-REAL/Easel --skill auto-short-video -a claude-code`. Or copy the skill folder (skills/openclaw/auto-short-video in ZJU-REAL/Easel) into .claude/skills/auto-short-video in your project. Claude Code loads it when a task matches its description.

How do I install Auto Short Video in Codex?

Run `npx skills add ZJU-REAL/Easel --skill auto-short-video -a codex`. Or copy the skill folder (skills/openclaw/auto-short-video in ZJU-REAL/Easel) into .agents/skills/auto-short-video in your project. Codex loads it when a task matches its description.

Can I use Auto Short Video in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add ZJU-REAL/Easel --skill auto-short-video -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/auto-short-video, .gemini/skills/auto-short-video, .github/skills/auto-short-video and .opencode/skills/auto-short-video in your project.

What does Auto Short Video need to run?

Going by SKILL.md and its folder, Auto Short Video needs Python for the scripts in its folder and the command-line tools its instructions call (python). Our summary lists: Python 3.

Does Auto Short Video access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Auto Short Video safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Auto Short Video use?

Auto Short Video is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Auto Short Video use?

About 748 tokens (SKILL.md is roughly 3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Auto Short Video?

Skills that share tags, products or a category with Auto Short Video: Video Cover Image (itwanger/toBeBetterJavaer, 18k stars), Narrator AI CLI (NarratorAI-Studio/narrator-ai-cli-skill, 3k stars), Ergo Remotion Video (itwanger/toBeBetterJavaer, 18k stars) and Wedding Video Guided Wizard (aaronyi97/wedding-video-guided-wizard, 310 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Auto Short Video?

ZJU-REAL (a GitHub organization) maintains it in ZJU-REAL/Easel, which has 3,376 GitHub stars. The repository holds 114 skills in this directory. The repository was last updated on October 9, 2026.

Source: ZJU-REAL/Easel on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.