Agent skill

Voice To Video

by wwwzhouhui in wwwzhouhui/skills_collection

口播文字稿一键成片:TTS 配音(edge-tts,带逐句/逐词时间戳)→ HTML 动画合成(场景由时间戳驱动,画面跟着声音走)→ Playwright 逐帧确定性渲染合成 MP4。Use whenever the user wants to 把口播稿/文字稿/文案/文章做成视频、配音视频、解说视频、知识类短视频、口播视频、文字转视频、TTS video、voice-over…

No licenceAuto-check passedMedia & Creative

Install Voice To Video

skills CLI
$ npx skills add wwwzhouhui/skills_collection --skill voice-to-video -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install wwwzhouhui/skills_collection voice-to-video --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/wwwzhouhui/skills_collection.git skills-src && mkdir -p .claude/skills && cp -r skills-src/voice-to-video .claude/skills/voice-to-video && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
voice-to-video
GitHub stars
283
Token cost
~1.1k tokens
SKILL.md length
260 words
Files
21 (incl. scripts, references, assets)
Skills in repo
25
Repo updated
First seen
Licence
None found

At a glance

口播文字稿一键成片:TTS 配音(edge-tts,带逐句/逐词时间戳)→ HTML 动画合成(场景由时间戳驱动,画面跟着声音走)→ Playwright 逐帧确定性渲染合成 MP4。Use whenever the user wants to 把口播稿/文字稿/文案/文章做成视频、配音视频、解说视频、知识类短视频、口播视频、文字转视频、TTS video、voice-over…

  • Works in 7 steps: 环境自检 → 口播稿 script.txt → 生成配音与时间轴 → …
  • Tasks that involve Text to speech and voice
  • SKILL.md covers 工作目录约定, Step 0 · 环境自检, Step 1 · 口播稿 script.txt and Step 2 · 生成配音与时间轴, plus 5 more sections
  • Runs JavaScript scripts from its folder; calls python and ffmpeg

What it does

Voice To Video is an agent skill from wwwzhouhui/skills_collection. 口播文字稿一键成片:TTS 配音(edge-tts,带逐句/逐词时间戳)→ HTML 动画合成(场景由时间戳驱动,画面跟着声音走)→ Playwright 逐帧确定性渲染合成 MP4。Use whenever the user wants to 把口播稿/文字稿/文案/文章做成视频、配音视频、解说视频、知识类短视频、口播视频、文字转视频、TTS video、voice-over video、画音同步、字幕驱动画面,or mentions HyperFrames 式的 HTML 渲视频流程 — even if they just say "帮我做个视频" 并附了一段稿子.

Its SKILL.md is about 1.1k tokens, which your agent loads only when the skill is triggered. The skill folder holds 23 other files, including scripts, reference files and assets (for example `README.md`, `assets/engine.js` and `references/composition-guide.md`).

It sits in Media & Creative, covering Text to speech and voice, Motion graphics and Browser testing. It works with Playwright and HeyGen. The repository describes itself as: 本项目是个人开发的 Claude Code Skills 集合,提供实用的技能工具,助力提升开发效率和内容创作。 分享一些好用的 Claude Code Skills,自用、学习两相宜,适用于 Claude Code v2.0 及以上版本。

When your agent uses it

  • Tasks that involve Text to speech and voice
  • Tasks that involve Motion graphics
  • Tasks that involve Browser testing

Example prompts

  • “帮我做个视频”
  • “/voice-to-video”

Requirements

  • Python 3
  • Node.js

Workflow steps

7 steps, taken from the step headings in SKILL.md.

  1. 环境自检
  2. 口播稿 script.txt
  3. 生成配音与时间轴
  4. 编写合成页 composition.html
  5. 预览(合成后必做)
  6. 渲染成片
  7. 交付

What it can do on your machine

Read from SKILL.md and the folder at commit b98f166. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (JavaScript, from the files we listed), which the agent can run.

    Shell commands in SKILL.md call:

    • python
    • ffmpeg

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Voice To Video loads about 1.1k tokens when it runs, and up to ~5.2k if it reads all its reference files. Until then it costs about 76 tokens; SKILL.md has 260 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~76
When it runs · the whole SKILL.md, loaded when a task matches
~1.1k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~5.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 260 words (~1,141 tokens).

“把一段口播文字稿变成一条 文字稿 ↔ 配音 ↔ 画面逐句对应 的 MP4:”

— opening of SKILL.md by wwwzhouhui
name
voice-to-video

Read the full SKILL.md on GitHub

Files

SKILL.md and 20 other files (scripts, references, assets) in voice-to-video of wwwzhouhui/skills_collection.

  • SKILL.md
  • README.md
  • assets/engine.js
  • assets/kit.css
  • assets/kits/blueprint.css
  • assets/kits/clay.css
  • assets/kits/dark-hud.css
  • assets/kits/editorial.css
  • assets/kits/glass.css
  • assets/kits/ink.css
  • assets/kits/pixel.css
  • assets/kits/stage.css
  • assets/kits/swiss.css
  • assets/kits/terminal.css
  • assets/kits/whiteboard.css
  • assets/template.html
  • references/composition-guide.md
  • references/styles.md
  • … and 3 more

Open the folder on GitHubat commit b98f166

Compare with similar skills

Voice To Video next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Voice To Video compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Voice To Video this skillwwwzhouhui/skills_collection283—~1.1kAutomated safety check: PassNone
Motion Adfabricioctelles/skills106—~4.1kAutomated safety check: PassApache-2.0
Render Myth Vs Factgooseworks-ai/goose-skills1.2k—~2.2kAutomated safety check: PassMIT
Hyperframes Mediachmonitor/chmonitor3011 repos~2.8kAutomated safety check: NotesGPL-3.0
Hyperframes CLInateherkai/hyperframes-student-kit1.3k3 repos~1.2kAutomated safety check: PassCustom licence
Create VideoCuongyd196/auto-compare-video211—~2.8kAutomated safety check: NotesCustom licence

Similar skills

  • Motion Ad

    fabricioctelles/skills

    Produce a short motion-graphics video ad — a 15s Facebook/Instagram/TikTok spot — as a rendered MP4.

    106 GitHub stars~4.1k tokensUpdated today
    Media & CreativeAuto-check passed
  • Render Myth Vs Fact

    gooseworks-ai/goose-skills

    Assemble a myth-vs-fact kinetic-typography explainer video ad (≈29.5s, 9:16) from N myth/fact pairs + hook / turn / punch copy + palette + a brand end-card PNG + a VO track — a hook, 3 red-strike…

    1.2k GitHub stars~2.2k tokensUpdated 2 days ago
    Media & CreativeAuto-check passed
  • Hyperframes Media

    chmonitor/chmonitor

    Audio and media assets for HyperFrames compositions, produced by one shared audio engine (scripts/audio.mjs) — multi-provider TTS (HeyGen / ElevenLabs / Kokoro local), background music + sound…

    301 GitHub starsUsed in 1 repo~2.8k tokens
    Media & CreativeAuto-check: notes
  • Hyperframes CLI

    nateherkai/hyperframes-student-kit

    HyperFrames CLI tool — hyperframes init, lint, preview, render, transcribe, tts, doctor, browser, info, upgrade, compositions, docs, benchmark.

    1.3k GitHub starsUsed in 3 repos~1.2k tokens
    Media & CreativeAuto-check passed
  • Create Video

    Cuongyd196/auto-compare-video

    Tạo một video MỚI cho series "so sánh / phân biệt kiến thức" của repo này — clip dọc TikTok/Reels/Shorts 30-40s, layout 3-zone cố định theo DESIGN.md, voiceover tiếng Việt sinh bằng VieNeu TTS…

    211 GitHub stars~2.8k tokensUpdated 17 days ago
    Media & CreativeAuto-check: notes
  • Super Video Maker

    Bomx/super-video-maker-skill

    End-to-end AI video production skill for agentic frameworks.

    310 GitHub stars~11k tokensUpdated 2 mo ago
    Media & CreativeAuto-check: notes

More from wwwzhouhui/skills_collection

All 25 skills in this repo
  • Knowledge Absorber

    wwwzhouhui/skills_collection

    深度解析链接/文档/代码,生成导师级教学笔记 + Wan 2.7 知识海报. An agent skill from wwwzhouhui/skills_collection.

    283 GitHub stars~2k tokensUpdated 5 days ago
    Auto-check passed
  • Remotion Video Factory

    wwwzhouhui/skills_collection

    视觉从代码生长的技术讲解视频工厂:给一个主题,产出一条"图解动画 + AI 配音"的 MP4——结构图/矩阵/连线拓扑/图表/数字滚动全部用 Remotion + React/SVG 代码精确绘制(不用 AI 视频模型,杜绝公式乱码与连线漂移,可参数化、改数据自动重排),旁白由 TTS 逐段生成 + ffprobe 实测时长(默认 edge-tts 神经音色;--engine clone…

    283 GitHub stars~990 tokensUpdated 5 days ago
    Auto-check passed
  • Article Explainer Video

    wwwzhouhui/skills_collection

    把一篇技术长文/论文解读自动做成章节式解说视频(1080p, 5-8 分钟)。双主题:warm(奶油底+珊瑚红+cozy-handdrawn 透明插图,亲和感)和 midnight(深蓝黑底+琥珀金+宋体标题+executive-tech 插图,AI 科技感),storyboard 一个 theme 字段切换。每章三种 layout 混排:illustration(左文右图+Ken…

    283 GitHub stars~2.2k tokensUpdated 5 days ago
    Auto-check passed
  • Edu Teaching Animation

    wwwzhouhui/skills_collection

    把单个学科概念(中文/英文都行,如"声现象""杠杆原理""光合作用")自动做成动态教学内容。两种产出都由 HyperFrames 渲染、共用同一个 index.html(mode 变量切换):① 配音教学视频 — 分镜 + Minimax / Edge TTS 中文配音 + 字幕 + 完整 MP4(1080p, ~90s, 发视频号/给孩子看);② 无声循环动图 — 同一内容的紧凑无声版…

    283 GitHub stars~1.5k tokensUpdated 5 days ago
    Auto-check passed
  • Env Setup

    wwwzhouhui/skills_collection

    Checking and provisioning the machine's environment for the video-agent-kit plugin — probing for ffmpeg/ffprobe that actually carry the encoders and filters we render with (libx264/aac/libmp3lame…

    283 GitHub stars~3.5k tokensUpdated 5 days ago
    Auto-check: notes
  • GitHub Trending Wan Skill

    wwwzhouhui/skills_collection

    抓取 GitHub Trending 当前前 5 个开源项目,先把摘要字段翻译成中文,再生成低信息密度中文简报和 Wan 2.7 海报 prompt。支持 10 种视觉风格选择。Use when user asks for GitHub trending、Top 5、开源日报、中文热门项目海报、Wan 2.7 poster、今天有什么热门项目、热门开源、做个开源海报、trending…

    283 GitHub stars~1.7k tokensUpdated 5 days ago
    Auto-check passed

Questions about Voice To Video

What does Voice To Video do?

口播文字稿一键成片:TTS 配音(edge-tts,带逐句/逐词时间戳)→ HTML 动画合成(场景由时间戳驱动,画面跟着声音走)→ Playwright 逐帧确定性渲染合成 MP4。Use whenever the user wants to 把口播稿/文字稿/文案/文章做成视频、配音视频、解说视频、知识类短视频、口播视频、文字转视频、TTS video、voice-over…. Voice To Video is an agent skill from wwwzhouhui/skills_collection. 口播文字稿一键成片:TTS 配音(edge-tts,带逐句/逐词时间戳)→ HTML 动画合成(场景由时间戳驱动,画面跟着声音走)→ Playwright 逐帧确定性渲染合成 MP4。Use whenever the user wants to 把口播稿/文字稿/文案/文章做成视频、配音视频、解说视频、知识类短视频、口播视频、文字转视频、TTS video、voice-over video、画音同步、字幕驱动画面,or mentions HyperFrames 式的 HTML 渲视频流程 — even if they just say "帮我做个视频" 并附了一段稿子.

When should I use Voice To Video?

Voice To Video fits situations like: tasks that involve Text to speech and voice; tasks that involve Motion graphics; tasks that involve Browser testing.

How do I install Voice To Video in Claude Code?

Run `npx skills add wwwzhouhui/skills_collection --skill voice-to-video -a claude-code`. Or copy the skill folder (voice-to-video in wwwzhouhui/skills_collection) into .claude/skills/voice-to-video in your project. Claude Code loads it when a task matches its description.

How do I install Voice To Video in Codex?

Run `npx skills add wwwzhouhui/skills_collection --skill voice-to-video -a codex`. Or copy the skill folder (voice-to-video in wwwzhouhui/skills_collection) into .agents/skills/voice-to-video in your project. Codex loads it when a task matches its description.

Can I use Voice To Video in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add wwwzhouhui/skills_collection --skill voice-to-video -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/voice-to-video, .gemini/skills/voice-to-video, .github/skills/voice-to-video and .opencode/skills/voice-to-video in your project.

What does Voice To Video need to run?

Going by SKILL.md and its folder, Voice To Video needs JavaScript for the scripts in its folder and the command-line tools its instructions call (python and ffmpeg). Our summary lists: Python 3; Node.js.

Does Voice To Video access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Voice To Video safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Voice To Video use?

No licence was found for Voice To Video or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Voice To Video use?

About 1.1k tokens (SKILL.md is roughly 4.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 4k tokens, read only when the agent opens those files.

What are the alternatives to Voice To Video?

Skills that share tags, products or a category with Voice To Video: Motion Ad (fabricioctelles/skills, 106 stars), Render Myth Vs Fact (gooseworks-ai/goose-skills, 1.2k stars), Hyperframes Media (chmonitor/chmonitor, 301 stars) and Hyperframes CLI (nateherkai/hyperframes-student-kit, 1.3k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Voice To Video?

wwwzhouhui (a GitHub user) maintains it in wwwzhouhui/skills_collection, which has 283 GitHub stars. The repository holds 25 skills in this directory. The repository was last updated on October 6, 2026.

Source: wwwzhouhui/skills_collection on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.