Agent skill

Video To Article

by ZJU-REAL in ZJU-REAL/Easel

把口播、讲座、直播或 Vlog 转录并改写成小红书笔记、公众号文章或知乎内容,同时抽帧配图。当用户说“视频转图文/文章/笔记、视频扒文案、口播转文章、视频内容复用”时使用。只生成字幕文件用 auto-subtitle;翻译已有字幕用 subtitle-translate。

Apache-2.0Auto-check passedMedia & Creative

Install Video To Article

skills CLI
$ npx skills add ZJU-REAL/Easel --skill video-to-article -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install ZJU-REAL/Easel video-to-article --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/ZJU-REAL/Easel.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/openclaw/video-to-article .claude/skills/video-to-article && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
video-to-article
GitHub stars
3.4k
Token cost
~585 tokens
SKILL.md length
126 words
Files
1
Skills in repo
114
Repo updated
First seen
Licence
Apache-2.0

At a glance

把口播、讲座、直播或 Vlog 转录并改写成小红书笔记、公众号文章或知乎内容,同时抽帧配图。当用户说“视频转图文/文章/笔记、视频扒文案、口播转文章、视频内容复用”时使用。只生成字幕文件用 auto-subtitle;翻译已有字幕用 subtitle-translate。

  • Works in 3 steps: 语音转录(带时间轴) → 结构化成图文(你来做) → 抽取配图
  • Tasks that involve Transcription
  • SKILL.md covers 输入, 输出(outputs/主题名/), 执行步骤 and Profile 感知, plus 2 more sections
  • Calls python

What it does

Video To Article is an agent skill from ZJU-REAL/Easel. 把口播、讲座、直播或 Vlog 转录并改写成小红书笔记、公众号文章或知乎内容,同时抽帧配图。当用户说“视频转图文/文章/笔记、视频扒文案、口播转文章、视频内容复用”时使用。只生成字幕文件用 auto-subtitle;翻译已有字幕用 subtitle-translate。

Its SKILL.md is about 590 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering Transcription. It works with Xiaohongshu. The repository describes itself as: An open-source AI agent for social media — discover trends, create content, publish everywhere, and learn what works across Xiaohongshu, Douyin, Zhihu, Bilibili, and more.🎨一个开源的… The licence is Apache-2.0.

When your agent uses it

  • Tasks that involve Transcription

Example prompts

  • “视频转图文/文章/笔记、视频扒文案、口播转文章、视频内容复用”
  • “/video-to-article”

Requirements

  • Python 3

Workflow steps

3 steps, taken from the step headings in SKILL.md.

  1. 语音转录(带时间轴)
  2. 结构化成图文(你来做)
  3. 抽取配图

What it can do on your machine

Read from SKILL.md and the folder at commit 278f420. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Video To Article loads about 585 tokens when it runs. Until then it costs about 38 tokens; SKILL.md has 126 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~38
When it runs · the whole SKILL.md, loaded when a task matches
~585

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from ZJU-REAL/Easel at commit 278f420, republished under its Apache-2.0 licence (© ZJU-REAL). 126 words, ~585 tokens.

Download SKILL.mdSave it as .claude/skills/video-to-article/SKILL.md (or your agent's skills folder).
name
video-to-article
description
把口播、讲座、直播或 Vlog 转录并改写成小红书笔记、公众号文章或知乎内容,同时抽帧配图。当用户说“视频转图文/文章/笔记、视频扒文案、口播转文章、视频内容复用”时使用。只生成字幕文件用 auto-subtitle;翻译已有字幕用 subtitle-translate。
layer
produce

视频转图文(视频 → 笔记/文章)

把视频复用成图文内容:转录 → 结构化成篇 → 抽帧配图。转录与抽帧走确定性脚本 (asr.py / video_ops.py),结构化成文由你(LLM)完成——这是本 SKILL 的核心价值。

只出字幕文件见 auto-subtitle;翻译字幕见 subtitle-translate; 出成套小红书卡片见 xhs-note-creator;纯文案润色见 text-polisher。

输入

字段必填说明
视频文件是口播/讲座/直播/Vlog(没给就问)
目标形态否小红书笔记(默认)/ 公众号文章 / 知乎回答 / 通用图文
配图数量否从视频抽几张配图(默认 3-6,按内容节点)

输出(outputs/主题名/)

  • article.md — 成篇图文(标题 + 正文 + 小标题/要点 + 金句 + 话题标签)
  • assets/frame-*.jpg — 抽取的配图
  • assets/transcript.txt / assets/transcript.json — 转录原文与时间轴(备查)

执行步骤

脚本路径(相对项目根):skills/shared/scripts/asr.py、skills/shared/scripts/video_ops.py。

1. 语音转录(带时间轴)
bash
python skills/shared/scripts/asr.py transcribe -i input.mp4 --format json \
  -o outputs/主题名/assets/transcript.json
python skills/shared/scripts/asr.py transcribe -i input.mp4 --format txt \
  -o outputs/主题名/assets/transcript.txt

(首次跑 ASR 需外网代理下模型,见 auto-subtitle 前置说明。)

2. 结构化成图文(你来做)

读转录,按目标形态改写成图文,不是照抄口语:

  • 提炼结构:口语流水账 → 清晰的标题 + 3-6 个小标题/要点段落。
  • 去口水:删"然后、就是、那个"等口头禅,书面化但保留个人风格。
  • 抓金句:把视频里最有价值的观点提成金句/加粗句。
  • 按形态适配:小红书(emoji、短段、闺蜜语气、话题标签)/ 公众号(成文、有起承转合)/ 知乎(专业、有逻辑链)。字数与排版参考 post-formatter / social-content 规范。
  • 写入 article.md,并在文中标注"【配图1:xx画面 @ 02:15】"指明每张配图对应的视频时间点。
3. 抽取配图

按第 2 步标注的时间点,逐个抽帧:

bash
python skills/shared/scripts/video_ops.py frame -i input.mp4 \
  -o outputs/主题名/assets/frame-01.jpg --time 00:02:15 --width 1080

挑画面清晰、有信息量的时间点(避免糊帧/转场帧)。

4.(可选)成套卡片

需要做成小红书卡片组时,把 article.md 交给 xhs-note-creator 或 card-xiaohongshu。

Profile 感知

  • 有 Profile:目标形态默认按 platforms.md 主平台;语气/称呼/emoji 尺度贴合 style.md; 话题标签贴合账号垂类;合规底线遵守 preferences.md。
  • 无 Profile:默认小红书笔记形态 + 中性口语风,末尾提示可提供 Profile 定制语气。

规则

  1. 是改写不是照搬转录——口语要书面化、结构化,去口水词。
  2. 配图从视频真实画面抽取,时间点由内容决定,避免糊帧。
  3. 不编造视频里没有的信息;转录不清处标注"[听不清]"而非臆测。
  4. 保留说话人的核心观点与个人风格,别改成千篇一律的 AI 腔(可再过 text-polisher)。
  5. 最终 article.md 放 outputs/主题名/,转录和抽帧等中间件放 outputs/主题名/assets/。

参考来源

视频→图文是创作者复用内容的高频需求(一鱼多吃)。转录用 faster-whisper(asr.py),配图用 ffmpeg 抽帧(video_ops.py frame),成文结构化交给 LLM——把确定性 IO 与创意改写分层。

© ZJU-REAL, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/openclaw/video-to-article of ZJU-REAL/Easel.

Open the folder on GitHubat commit 278f420

Compare with similar skills

Video To Article next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Video To Article compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Video To Article this skillZJU-REAL/Easel3.4k—~585Automated safety check: PassApache-2.0
Lingzaoatian-create/lingzao-skill2961 repos~8.8kAutomated safety check: PassMIT
Member Skill DistillerJamailar/Beav1.8k—~296Automated safety check: PassCustom licence
Rednote ResearchLeoYeAI/openclaw-master-skills2.2k—~4.3kAutomated safety check: PassMIT
FeedgrabiBigQiang/feedgrab614—~2kAutomated safety check: PassMIT
Media To TranscriptbozhouDev/video-skills-toolkit150—~1.8kAutomated safety check: NotesMIT

Similar skills

  • Lingzao

    atian-create/lingzao-skill

    Use Lingzao creator-content tools for Xiaohongshu/XHS, Douyin, and WeChat official-account public content.

    296 GitHub starsUsed in 1 repo~8.8k tokens
    Media & CreativeAuto-check passed
  • Distill team members from profile, files, and YouTube subtitles into session-activated member skills.

    1.8k GitHub stars~296 tokensUpdated 2 days ago
    Media & CreativeAuto-check passed
  • Rednote Research

    LeoYeAI/openclaw-master-skills

    Research a topic through RedNote/Xiaohongshu discussion signals using either public-web mode (no login) or optional login-enhanced browser review when the user explicitly chooses deeper access.

    2.2k GitHub stars~4.3k tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed
  • Feedgrab

    iBigQiang/feedgrab

    Universal content grabber — fetch any URL and return structured Markdown.

    614 GitHub stars~2k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Media To Transcript

    bozhouDev/video-skills-toolkit

    Convert audio/video URLs or local media into corrected Markdown transcripts through Volcengine recording-file ASR 2.0.

    150 GitHub stars~1.8k tokensUpdated 2 mo ago
    Media & CreativeAuto-check: notes
  • Ra Video Download

    Pluviobyte/rnskill

    Download source video or audio from Douyin, YouTube, Bilibili, Twitter/X, Xiaohongshu, and other yt-dlp-supported URLs into the content-creation workspace.

    1.6k GitHub stars~861 tokensUpdated 20 days ago
    Media & CreativeAuto-check: notes

More from ZJU-REAL/Easel

All 114 skills in this repo
  • Gzh Design

    ZJU-REAL/Easel

    微信公众号文章排版引擎:把 Markdown / Word(.docx) / PDF / 纯文本转成可直接粘贴进公众号编辑器的 HTML,自动章节编号、关键词标记、引言卡、目录、代码块、图片/GIF、作者签名;主题从 references/theme-index.md…

    3.4k GitHub stars~2.1k tokensUpdated today
    Auto-check passed
  • 微信公众号文章自动创作与发布工具。给定参考文章、文字或文档,自动搜索整理全网相关信息、生成图文并茂的公众号文章,并发布到微信公众号草稿箱。特别强调反 AI 检测写作。

    3.4k GitHub stars~1.8k tokensUpdated today
    Auto-check passed
  • Card Design

    ZJU-REAL/Easel

    社媒卡片视觉设计系统:提供配色、中文字体层级、满画幅布局、品类骨架和死空白/密度质检,避免模板化 PPT 与廉价 AI 感。

    3.4k GitHub stars~657 tokensUpdated today
    Auto-check passed
  • Ecom Details Image

    ZJU-REAL/Easel

    生成电商商品视觉方案:主图概念、场景图、详情页视觉方向和 AI 生图 Prompt. An agent skill from ZJU-REAL/Easel.

    3.4k GitHub stars~1.1k tokensUpdated today
    Auto-check: notes
  • Infographic

    ZJU-REAL/Easel

    将数据或文字内容转化为可视化信息图,支持静态(AntV)和动画 GIF 两种模式。当用户需要制作信息图、数据可视化、流程图、对比图、动画图表、GIF 图表、思维导图、SWOT 分析图时调用。本地渲染信息图/GIF 动画;要单张静态图片 URL 用 chart-visualization,要 CSV/JSON→整页报告用 data-report

    3.4k GitHub stars~643 tokensUpdated today
    Auto-check passed
  • Novel Writer

    ZJU-REAL/Easel

    长篇小说/网文连载创作:从世界观、人设和三级大纲写到逐章正文,并用文件化状态维护伏笔、前情和跨章一致性. An agent skill from ZJU-REAL/Easel.

    3.4k GitHub stars~1k tokensUpdated today
    Auto-check passed

Works with

Questions about Video To Article

What does Video To Article do?

把口播、讲座、直播或 Vlog 转录并改写成小红书笔记、公众号文章或知乎内容,同时抽帧配图。当用户说“视频转图文/文章/笔记、视频扒文案、口播转文章、视频内容复用”时使用。只生成字幕文件用 auto-subtitle;翻译已有字幕用 subtitle-translate。. Video To Article is an agent skill from ZJU-REAL/Easel.

When should I use Video To Article?

Video To Article fits situations like: tasks that involve Transcription.

How do I install Video To Article in Claude Code?

Run `npx skills add ZJU-REAL/Easel --skill video-to-article -a claude-code`. Or copy the skill folder (skills/openclaw/video-to-article in ZJU-REAL/Easel) into .claude/skills/video-to-article in your project. Claude Code loads it when a task matches its description.

How do I install Video To Article in Codex?

Run `npx skills add ZJU-REAL/Easel --skill video-to-article -a codex`. Or copy the skill folder (skills/openclaw/video-to-article in ZJU-REAL/Easel) into .agents/skills/video-to-article in your project. Codex loads it when a task matches its description.

Can I use Video To Article in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add ZJU-REAL/Easel --skill video-to-article -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/video-to-article, .gemini/skills/video-to-article, .github/skills/video-to-article and .opencode/skills/video-to-article in your project.

What does Video To Article need to run?

Going by SKILL.md and its folder, Video To Article needs the command-line tools its instructions call (python). Our summary lists: Python 3.

Does Video To Article access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Video To Article safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Video To Article use?

Video To Article is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Video To Article use?

About 585 tokens (SKILL.md is roughly 2.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Video To Article?

Skills that share tags, products or a category with Video To Article: Lingzao (atian-create/lingzao-skill, 296 stars), Member Skill Distiller (Jamailar/Beav, 1.8k stars), Rednote Research (LeoYeAI/openclaw-master-skills, 2.2k stars) and Feedgrab (iBigQiang/feedgrab, 614 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Video To Article?

ZJU-REAL (a GitHub organization) maintains it in ZJU-REAL/Easel, which has 3,376 GitHub stars. The repository holds 114 skills in this directory. The repository was last updated on October 9, 2026.

Source: ZJU-REAL/Easel on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.