Agent skill

Tts Voiceover

by liancheng-zcy in liancheng-zcy/remotion-com-skills

Generate voiceover audio using local VoxCPM2 TTS model with voice cloning.

MITAuto-check passedMedia & Creative

Install Tts Voiceover

skills CLI
$ npx skills add liancheng-zcy/remotion-com-skills --skill tts-voiceover -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install liancheng-zcy/remotion-com-skills tts-voiceover --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/liancheng-zcy/remotion-com-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/tts-voiceover .claude/skills/tts-voiceover && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
tts-voiceover
GitHub stars
106
Token cost
~569 tokens
SKILL.md length
87 words
Files
1
Skills in repo
1
Repo updated
First seen
Licence
MIT

At a glance

Generate voiceover audio using local VoxCPM2 TTS model with voice cloning.

  • Works in 6 steps: 首次加载模型约需 5-15 秒,请耐心等待 → 音质受参考音频质量影响:建议用 10 秒以上的干净录音,不要有背景噪音 → cfg_value:默认 2.0,值越高越接近参考音频的语调风格(范围… → …
  • The user says anything about: 配音
  • SKILL.md covers 概述, 工作流程, 工具调用方式 and 注意事项
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Tts Voiceover is an agent skill from liancheng-zcy/remotion-com-skills. Generate voiceover audio using local VoxCPM2 TTS model with voice cloning. TRIGGER when the user says anything about: 配音, 配音, 语音合成, TTS, voice clone, voiceover, 克隆声音, 音色设计, 用我的声音, 朗读文案, 生成语音, 音频克隆, 语音生成, or asks to turn text into speech / narration / voiceover. This skill handles the entire workflow: locating the TTS integration package, configuring reference audio, generating voice clone or voice design audio, and saving output files. Use this whenever the user wants to generate spoken audio from text, even if…

Its SKILL.md is about 570 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering Text to speech and voice. The repository describes itself as: remotion 常用组件库 + 自定义skills. The licence is MIT.

When your agent uses it

  • The user says anything about: 配音
  • Asks to turn text into speech / narration / voiceover
  • Wants to generate spoken audio from text
  • Even if they dont explicitly mention TTS — if they want narration

Example prompts

  • “t explicitly mention”
  • “/tts-voiceover”

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. 首次加载模型约需 5-15 秒,请耐心等待
  2. 音质受参考音频质量影响:建议用 10 秒以上的干净录音,不要有背景噪音
  3. cfg_value:默认 2.0,值越高越接近参考音频的语调风格(范围 1.0-3.0)
  4. 声音设计的文本格式:(音色描述)要说的内容 — 括号内是音色描述,括号外是台词
  5. 批量处理支持每行单独指定参考音频,用 ||| 分隔
  6. 所有处理在本地完成,不上传数据

What it can do on your machine

Read from SKILL.md and the folder at commit 22e2a7c. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are bash).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Tts Voiceover loads about 569 tokens when it runs. Until then it costs about 160 tokens; SKILL.md has 87 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~160
When it runs · the whole SKILL.md, loaded when a task matches
~569

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from liancheng-zcy/remotion-com-skills at commit 22e2a7c, republished under its MIT licence (© liancheng-zcy). 87 words, ~569 tokens.

Download SKILL.mdSave it as .claude/skills/tts-voiceover/SKILL.md (or your agent's skills folder).
name
tts-voiceover
description
Generate voiceover audio using local VoxCPM2 TTS model with voice cloning. TRIGGER when the user says anything about: 配音, 配音, 语音合成, TTS, voice clone, voiceover, 克隆声音, 音色设计, 用我的声音, 朗读文案, 生成语音, 音频克隆, 语音生成, or asks to turn text into speech / narration / voiceover. This skill handles the entire workflow: locating the TTS integration package, configuring reference audio, generating voice clone or voice design audio, and saving output files. Use this whenever the user wants to generate spoken audio from text, even if they don't explicitly mention "TTS" — if they want narration, voiceover, or cloned voice, this skill applies.

TTS Voiceover Skill

概述

本 skill 封装了本地 VoxCPM2 TTS 模型的配音流程。首次使用时,需要先确认用户的环境配置。

工作流程

第一步:询问整合包路径

用户提出配音需求后,主动询问整合包位置:

"请告诉我你的 VoxCPM2 TTS 整合包放在哪个目录?比如 D:\Tools\jinxi-voxcpm-tts\"

如果用户不清楚,引导用户:

  • 整合包包含 tts-cli.bat、python/、models/ 等文件
  • 通常是下载后解压的文件夹
  • 可以提供几个常见路径示例让用户选择
第二步:检查参考音频配置

询问用户:

"是否已配置参考音频(用于声音克隆)?如果没有,请告诉我你的参考音频文件路径,我来帮你配置。"

根据回答:

  • 有参考音频 → 获取路径后执行 tts-cli config reference_audio "路径"
  • 没有但想用声音克隆 → 引导用户准备一段 10 秒以上的干净录音(.wav),然后配置
  • 不需要声音克隆 → 使用 voice design 模式,无需参考音频
第三步:生成配音

根据用户需求选择模式:

声音克隆(clone) — 用参考音频的音色朗读指定文案

bash
tts-cli clone "文案内容" [输出路径.wav]
tts-cli clone "文案内容" --ref "其他参考音频.wav"  # 临时换参考音频

声音设计(design) — 通过文字描述生成音色,不需要参考音频

bash
tts-cli design "描述文字(文案内容)" [输出路径.wav]
# 示例: tts-cli design "暴躁中年男声,语速快,愤怒(踩离合!你往哪开呢!)" output.wav
# 示例: tts-cli design "温柔女声,舒缓柔和(你好,欢迎光临)" output.wav

批量处理(batch) — 每行一条文案,批量生成

bash
tts-cli batch 文案列表.txt [输出目录]
# 每行可以单独指定参考音频:
# 文案内容 ||| 自定义参考音频.wav
第四步:输出文件
  • 默认输出到整合包目录下的 outputs/ 文件夹
  • 也可以指定任意输出路径
  • 格式:48kHz mono WAV
  • 如果用在 Remotion 项目中,建议输出到 public/voiceover/

工具调用方式

确认整合包路径后,用完整路径调用:

bash
# PowerShell / CMD
& "D:\path\to\tts-cli.bat" clone "文案内容" output.wav

如果用户在整合包目录下操作,也可以直接用相对路径:

bash
cd /d D:\path\to\jinxi-voxcpm-tts
tts-cli clone "文案内容" output.wav

注意事项

  1. 首次加载模型约需 5-15 秒,请耐心等待
  2. 音质受参考音频质量影响:建议用 10 秒以上的干净录音,不要有背景噪音
  3. cfg_value:默认 2.0,值越高越接近参考音频的语调风格(范围 1.0-3.0)
  4. 声音设计的文本格式:(音色描述)要说的内容 — 括号内是音色描述,括号外是台词
  5. 批量处理支持每行单独指定参考音频,用 ||| 分隔
  6. 所有处理在本地完成,不上传数据

© liancheng-zcy, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .claude/skills/tts-voiceover of liancheng-zcy/remotion-com-skills.

Open the folder on GitHubat commit 22e2a7c

Compare with similar skills

Tts Voiceover next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Tts Voiceover compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Tts Voiceover this skillliancheng-zcy/remotion-com-skills106—~569Automated safety check: PassMIT
MoneyPrinterTurbo Video Generatorharry0703/MoneyPrinterTurbo130k—~2.1kAutomated safety check: WarnMIT
HyperFrames Media Useheygen-com/hyperframes60k—~2.4kAutomated safety check: PassApache-2.0
Openspec OnboardSAP/e-mobility-charging-stations-simulator22725 repos~3.5kAutomated safety check: PassMIT
Blog AudioAgriciDaniel/claude-blog2.3k1 repos~2.2kAutomated safety check: NotesMIT
Musictadaspetra/loop2962 repos~827Automated safety check: PassMIT

Similar skills

  • MoneyPrinterTurbo Video Generator

    harry0703/MoneyPrinterTurbo

    Installs and runs MoneyPrinterTurbo to turn a topic or script into a finished short video with voice-over, subtitles, stock footage and music.

    130k GitHub stars~2.1k tokensUpdated yesterday
    Media & CreativeAuto-check: warnings
  • HyperFrames Media Use

    heygen-com/hyperframes

    Finds, generates and edits media for HyperFrames video projects: music, sound effects, images, icons, logos, voiceovers, captions and color grades.

    60k GitHub stars~2.4k tokensUpdated today
    Media & CreativeAuto-check passed
  • Openspec Onboard

    SAP/e-mobility-charging-stations-simulator

    Official

    Guided onboarding for OpenSpec - walk through a complete workflow cycle with narration and real codebase work.

    227 GitHub starsUsed in 25 repos~3.5k tokens
    Media & CreativeAuto-check passed
  • Blog Audio

    AgriciDaniel/claude-blog

    Generate audio narration of blog posts using Google Gemini TTS.

    2.3k GitHub starsUsed in 1 repo~2.2k tokens
    Media & CreativeAuto-check: notes
  • Music

    tadaspetra/loop

    Generate music using ElevenLabs Music API. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 2 repos~827 tokens
    Media & CreativeAuto-check passed
  • Create News Video

    hoquanghai/Auto-Create-Video

    Tạo video tin tức ngắn 9:16 (~60s) từ URL bài báo hoặc file .txt tiếng Việt.

    319 GitHub starsUsed in 1 repo~3.7k tokens
    Media & CreativeAuto-check passed

Questions about Tts Voiceover

What does Tts Voiceover do?

Generate voiceover audio using local VoxCPM2 TTS model with voice cloning. Tts Voiceover is an agent skill from liancheng-zcy/remotion-com-skills. Generate voiceover audio using local VoxCPM2 TTS model with voice cloning.

When should I use Tts Voiceover?

Tts Voiceover fits situations like: the user says anything about: 配音; asks to turn text into speech / narration / voiceover; wants to generate spoken audio from text; even if they dont explicitly mention TTS — if they want narration.

How do I install Tts Voiceover in Claude Code?

Run `npx skills add liancheng-zcy/remotion-com-skills --skill tts-voiceover -a claude-code`. Or copy the skill folder (.claude/skills/tts-voiceover in liancheng-zcy/remotion-com-skills) into .claude/skills/tts-voiceover in your project. Claude Code loads it when a task matches its description.

How do I install Tts Voiceover in Codex?

Run `npx skills add liancheng-zcy/remotion-com-skills --skill tts-voiceover -a codex`. Or copy the skill folder (.claude/skills/tts-voiceover in liancheng-zcy/remotion-com-skills) into .agents/skills/tts-voiceover in your project. Codex loads it when a task matches its description.

Can I use Tts Voiceover in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add liancheng-zcy/remotion-com-skills --skill tts-voiceover -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/tts-voiceover, .gemini/skills/tts-voiceover, .github/skills/tts-voiceover and .opencode/skills/tts-voiceover in your project.

What does Tts Voiceover need to run?

SKILL.md names no scripts, command-line tools or credentials: Tts Voiceover is instructions for the agent only.

Does Tts Voiceover access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Tts Voiceover safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Tts Voiceover use?

Tts Voiceover is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Tts Voiceover use?

About 569 tokens (SKILL.md is roughly 2.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Tts Voiceover?

Skills that share tags, products or a category with Tts Voiceover: MoneyPrinterTurbo Video Generator (harry0703/MoneyPrinterTurbo, 130k stars), HyperFrames Media Use (heygen-com/hyperframes, 60k stars), Openspec Onboard (SAP/e-mobility-charging-stations-simulator, 227 stars) and Blog Audio (AgriciDaniel/claude-blog, 2.3k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Tts Voiceover?

liancheng-zcy (a GitHub user) maintains it in liancheng-zcy/remotion-com-skills, which has 106 GitHub stars. The repository was last updated on May 27, 2026.

Source: liancheng-zcy/remotion-com-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.