Agent skill

Tts Voice Synthesis

by anbeime in anbeime/skill

影片与视频编辑、内容创作者在制作视频配音或有声书时,当需要克隆音色、生成情感化配音或流式实时语音合成请用此技能。支持1.7B高质量与0.6B快速双模型,一键实现多语言方言配音,让语音创作更高效更自然。

No licenceAuto-check passedMedia & Creative

Install Tts Voice Synthesis

skills CLI
$ npx skills add anbeime/skill --skill tts-voice-synthesis -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install anbeime/skill tts-voice-synthesis --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/anbeime/skill.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/tts-voice-synthesis .claude/skills/tts-voice-synthesis && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
tts-voice-synthesis
GitHub stars
7.8k
Token cost
~757 tokens
SKILL.md length
140 words
Files
6 (incl. scripts, references)
Skills in repo
62
Repo updated
First seen
Licence
None found

At a glance

影片与视频编辑、内容创作者在制作视频配音或有声书时,当需要克隆音色、生成情感化配音或流式实时语音合成请用此技能。支持1.7B高质量与0.6B快速双模型,一键实现多语言方言配音,让语音创作更高效更自然。

  • Works in 4 steps: 文本准备 → 选择音色 → 执行合成 → …
  • Tasks that involve Text to speech and voice
  • SKILL.md covers 任务目标, 前置准备, 操作步骤 and 资源索引, plus 2 more sections
  • Runs Python scripts from its folder; calls python

What it does

Tts Voice Synthesis is an agent skill from anbeime/skill. 影片与视频编辑、内容创作者在制作视频配音或有声书时,当需要克隆音色、生成情感化配音或流式实时语音合成请用此技能。支持1.7B高质量与0.6B快速双模型,一键实现多语言方言配音,让语音创作更高效更自然。

Its SKILL.md is about 760 tokens, which your agent loads only when the skill is triggered. The skill folder holds 7 other files, including scripts and reference files (for example `references/emotion_guide.md`, `references/model_config.md` and `references/usage_examples.md`).

It sits in Media & Creative, covering Text to speech and voice. The repository describes itself as: 收录最全、更新最快的技能Skills商店:精选原创技能包(涵盖文档处理、内容创作、编程开发、机器学习、自动化工作流),全部打包好可直接安装使用!同时自动抓取GitHub上万个Skills项目,按分类、更新时间、Star数量整理。The most comprehensive and frequently updated AI Agent skill…

When your agent uses it

  • Tasks that involve Text to speech and voice

Example prompts

  • “/tts-voice-synthesis”

Requirements

  • Python 3

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. 文本准备
  2. 选择音色
  3. 执行合成
  4. 验证输出

What it can do on your machine

Read from SKILL.md and the folder at commit db1e192. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 2 files in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Tts Voice Synthesis loads about 757 tokens when it runs, and up to ~6.5k if it reads all its reference files. Until then it costs about 30 tokens; SKILL.md has 140 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~30
When it runs · the whole SKILL.md, loaded when a task matches
~757
With references · SKILL.md plus every file in references/, read only if the agent opens them
~6.5k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 140 words (~757 tokens).

name
tts-voice-synthesis
dependency.python
torch>=2.0.0 torchaudio>=2.0.0 transformers>=4.35.0 scipy>=1.11.0 numpy>=1.24.0 librosa>=0.10.0 soundfile>=0.12.0 pydub>=0.25.0
dependency.system
# 创建音色模型保存目录 mkdir -p voices mkdir -p output

Read the full SKILL.md on GitHub

Files

SKILL.md and 5 other files (scripts, references) in skills/tts-voice-synthesis of anbeime/skill.

  • SKILL.md
  • references/emotion_guide.md
  • references/model_config.md
  • references/usage_examples.md
  • scripts/tts_generate.py
  • scripts/voice_clone.py

Open the folder on GitHubat commit db1e192

Compare with similar skills

Tts Voice Synthesis next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Tts Voice Synthesis compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Tts Voice Synthesis this skillanbeime/skill7.8k—~757Automated safety check: PassNone
MoneyPrinterTurbo Video Generatorharry0703/MoneyPrinterTurbo130k—~2.1kAutomated safety check: WarnMIT
HyperFrames Media Useheygen-com/hyperframes60k—~2.4kAutomated safety check: PassApache-2.0
Openspec OnboardSAP/e-mobility-charging-stations-simulator22725 repos~3.5kAutomated safety check: PassMIT
Blog AudioAgriciDaniel/claude-blog2.3k1 repos~2.2kAutomated safety check: NotesMIT
Musictadaspetra/loop2962 repos~827Automated safety check: PassMIT

Similar skills

  • MoneyPrinterTurbo Video Generator

    harry0703/MoneyPrinterTurbo

    Installs and runs MoneyPrinterTurbo to turn a topic or script into a finished short video with voice-over, subtitles, stock footage and music.

    130k GitHub stars~2.1k tokensUpdated today
    Media & CreativeAuto-check: warnings
  • HyperFrames Media Use

    heygen-com/hyperframes

    Finds, generates and edits media for HyperFrames video projects: music, sound effects, images, icons, logos, voiceovers, captions and color grades.

    60k GitHub stars~2.4k tokensUpdated today
    Media & CreativeAuto-check passed
  • Openspec Onboard

    SAP/e-mobility-charging-stations-simulator

    Official

    Guided onboarding for OpenSpec - walk through a complete workflow cycle with narration and real codebase work.

    227 GitHub starsUsed in 25 repos~3.5k tokens
    Media & CreativeAuto-check passed
  • Blog Audio

    AgriciDaniel/claude-blog

    Generate audio narration of blog posts using Google Gemini TTS.

    2.3k GitHub starsUsed in 1 repo~2.2k tokens
    Media & CreativeAuto-check: notes
  • Music

    tadaspetra/loop

    Generate music using ElevenLabs Music API. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 2 repos~827 tokens
    Media & CreativeAuto-check passed
  • Create News Video

    hoquanghai/Auto-Create-Video

    Tạo video tin tức ngắn 9:16 (~60s) từ URL bài báo hoặc file .txt tiếng Việt.

    319 GitHub starsUsed in 1 repo~3.7k tokens
    Media & CreativeAuto-check passed

More from anbeime/skill

All 62 skills in this repo
  • Chinese-language skill that produces a 25-second vertical video of a digital shopping-guide avatar for e-commerce, chaining AI image, voice and video generation.

    7.8k GitHub stars~1k tokensUpdated yesterday
    Auto-check passed
  • Produces a full material pack for a three-minute history-of-science explainer video from a theme, era and core conclusion: narration, storyboard, Veo2 prompts and character design.

    7.8k GitHub stars~480 tokensUpdated yesterday
    Auto-check passed
  • Turns a PPT content outline into styled slide images, an interactive HTML viewer and an optional stitched video from a content-planning partner skill.

    7.8k GitHub stars~1.2k tokensUpdated yesterday
    Auto-check: notes
  • Builds or improves a PowerPoint deck through a seven-role workflow: topic analysis, template choice, content planning, writing, image suggestions, editing and PPTX generation.

    7.8k GitHub stars~1.4k tokensUpdated yesterday
    Auto-check passed
  • Turns a document into a narrated roadshow video through ten staged roles, from document analysis and slide planning to audio, subtitles and final composition.

    7.8k GitHub stars~1.4k tokensUpdated yesterday
    Auto-check: notes
  • PPTX Generator

    anbeime/skill

    产品经理与运营专员在准备汇报展示时,将 JSON 数据一键转换为标准 .pptx 文件。自动应用多种布局、图表与表格样式,快速生成专业 PPT,彻底告别繁琐手工排版,让内容产出与展示更高效!

    7.8k GitHub stars~1.4k tokensUpdated yesterday
    Auto-check passed

Questions about Tts Voice Synthesis

What does Tts Voice Synthesis do?

影片与视频编辑、内容创作者在制作视频配音或有声书时,当需要克隆音色、生成情感化配音或流式实时语音合成请用此技能。支持1.7B高质量与0.6B快速双模型,一键实现多语言方言配音,让语音创作更高效更自然。. Tts Voice Synthesis is an agent skill from anbeime/skill.

When should I use Tts Voice Synthesis?

Tts Voice Synthesis fits situations like: tasks that involve Text to speech and voice.

How do I install Tts Voice Synthesis in Claude Code?

Run `npx skills add anbeime/skill --skill tts-voice-synthesis -a claude-code`. Or copy the skill folder (skills/tts-voice-synthesis in anbeime/skill) into .claude/skills/tts-voice-synthesis in your project. Claude Code loads it when a task matches its description.

How do I install Tts Voice Synthesis in Codex?

Run `npx skills add anbeime/skill --skill tts-voice-synthesis -a codex`. Or copy the skill folder (skills/tts-voice-synthesis in anbeime/skill) into .agents/skills/tts-voice-synthesis in your project. Codex loads it when a task matches its description.

Can I use Tts Voice Synthesis in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add anbeime/skill --skill tts-voice-synthesis -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/tts-voice-synthesis, .gemini/skills/tts-voice-synthesis, .github/skills/tts-voice-synthesis and .opencode/skills/tts-voice-synthesis in your project.

What does Tts Voice Synthesis need to run?

Going by SKILL.md and its folder, Tts Voice Synthesis needs Python for the scripts in its folder and the command-line tools its instructions call (python). Our summary lists: Python 3.

Does Tts Voice Synthesis access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Tts Voice Synthesis safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Tts Voice Synthesis use?

No licence was found for Tts Voice Synthesis or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Tts Voice Synthesis use?

About 757 tokens (SKILL.md is roughly 3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 5.8k tokens, read only when the agent opens those files.

What are the alternatives to Tts Voice Synthesis?

Skills that share tags, products or a category with Tts Voice Synthesis: MoneyPrinterTurbo Video Generator (harry0703/MoneyPrinterTurbo, 130k stars), HyperFrames Media Use (heygen-com/hyperframes, 60k stars), Openspec Onboard (SAP/e-mobility-charging-stations-simulator, 227 stars) and Blog Audio (AgriciDaniel/claude-blog, 2.3k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Tts Voice Synthesis?

anbeime (a GitHub user) maintains it in anbeime/skill, which has 7,760 GitHub stars. The repository holds 62 skills in this directory. The repository was last updated on October 9, 2026.

Source: anbeime/skill on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.