Voice Clone Tts
npc-live/clawfirm
声纹克隆和语音合成。上传音频样本克隆声纹,用克隆声纹或预设声纹生成语音。支持多个后端:MiniMax、ElevenLabs、Fish Audio、Azure TTS、OpenAI TTS。支持情绪控制、语速调整、批量生成。触发词:语音合成、TTS、声纹克隆、voice clone、text to speech、配音、旁白。
上传本人语音样本克隆专属音色,再用它合成口播、旁白或带货语音。当用户说“声音克隆、克隆/复刻我的声音、用我的声音配音、定制专属音色”时使用,需要用户自备云端 provider 凭证。使用公共现成音色时改用 tts-voiceover。
$ npx skills add ZJU-REAL/Easel --skill voice-clone -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install ZJU-REAL/Easel voice-clone --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/ZJU-REAL/Easel.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/openclaw/voice-clone .claude/skills/voice-clone && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "voice-clone" agent skill from https://github.com/ZJU-REAL/Easel/tree/main/skills/openclaw/voice-clone into .claude/skills/voice-clone/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-clone", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/ZJU-REAL/Easel/tree/main/skills/openclaw/voice-cloneType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add ZJU-REAL/Easel --skill voice-clone -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install ZJU-REAL/Easel voice-clone --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/ZJU-REAL/Easel.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/openclaw/voice-clone .agents/skills/voice-clone && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "voice-clone" agent skill from https://github.com/ZJU-REAL/Easel/tree/main/skills/openclaw/voice-clone into .agents/skills/voice-clone/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-clone", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add ZJU-REAL/Easel --skill voice-clone -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install ZJU-REAL/Easel voice-clone --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/ZJU-REAL/Easel.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/openclaw/voice-clone .cursor/skills/voice-clone && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "voice-clone" agent skill from https://github.com/ZJU-REAL/Easel/tree/main/skills/openclaw/voice-clone into .cursor/skills/voice-clone/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-clone", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/ZJU-REAL/Easel.git --path skills/openclaw/voice-clone--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add ZJU-REAL/Easel --skill voice-clone -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install ZJU-REAL/Easel voice-clone --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/ZJU-REAL/Easel.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/openclaw/voice-clone .gemini/skills/voice-clone && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "voice-clone" agent skill from https://github.com/ZJU-REAL/Easel/tree/main/skills/openclaw/voice-clone into .gemini/skills/voice-clone/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-clone", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install ZJU-REAL/Easel voice-cloneInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add ZJU-REAL/Easel --skill voice-clone -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/ZJU-REAL/Easel.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/openclaw/voice-clone .github/skills/voice-clone && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "voice-clone" agent skill from https://github.com/ZJU-REAL/Easel/tree/main/skills/openclaw/voice-clone into .github/skills/voice-clone/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-clone", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add ZJU-REAL/Easel --skill voice-clone -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install ZJU-REAL/Easel voice-clone --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/ZJU-REAL/Easel.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/openclaw/voice-clone .opencode/skills/voice-clone && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "voice-clone" agent skill from https://github.com/ZJU-REAL/Easel/tree/main/skills/openclaw/voice-clone into .opencode/skills/voice-clone/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-clone", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
voice-clone上传本人语音样本克隆专属音色,再用它合成口播、旁白或带货语音。当用户说“声音克隆、克隆/复刻我的声音、用我的声音配音、定制专属音色”时使用,需要用户自备云端 provider 凭证。使用公共现成音色时改用 tts-voiceover。
Voice Clone is an agent skill from ZJU-REAL/Easel. 上传本人语音样本克隆专属音色,再用它合成口播、旁白或带货语音。当用户说“声音克隆、克隆/复刻我的声音、用我的声音配音、定制专属音色”时使用,需要用户自备云端 provider 凭证。使用公共现成音色时改用 tts-voiceover。
Its SKILL.md is about 720 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Media & Creative, covering Text to speech and voice. It works with MiniMax and OpenAI. The repository describes itself as: An open-source AI agent for social media — discover trends, create content, publish everywhere, and learn what works across Xiaohongshu, Douyin, Zhihu, Bilibili, and more.🎨一个开源的… The licence is Apache-2.0.
2 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 278f420. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
pythonFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names these keys or tokens, usually read from environment variables:
DASHSCOPE_API_KEYMINIMAX_API_KEYFISH_API_KEYVOICE_API_KEYGEMINI_API_KEYFrom names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Voice Clone loads about 715 tokens when it runs. Until then it costs about 32 tokens; SKILL.md has 175 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check noted patterns worth knowing about, such as sudo or a known installer.
` 到 `AGENTS.md` 末尾给出的 Easel 项目根,确认当前目录有 `.env` 和 `skills/shared/scripts/`,再运行注册表、`check`、`enroll` 或 `clone`。不得改用 workspa选 provider 并在 `.env` 填 key,再 `check` 离线校验:e.py check --provider minimax --env-file .env| provider | 服务 | .env 需配 |y.py configured --group voice --env-file .env`。只有一个可用时显式选择;多个可用且用户没点名时,列出 provider/模型询问本次使用哪个,不按 `VOICE_PROVIDER` 擅自选择。Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from ZJU-REAL/Easel at commit 278f420, republished under its Apache-2.0 licence (© ZJU-REAL). 175 words, ~715 tokens.
.claude/skills/voice-clone/SKILL.md (or your agent's skills folder).用本人语音样本克隆音色,再合成任意文案。走云端 provider(用户自备 key),本地无需 GPU。 全部走
skills/shared/scripts/voice_clone.py。
不想克隆、用现成公共音色见 tts-voiceover(edge-tts,免费无需 key); AI 生成音乐/BGM 见 ai-music;合成后与 BGM 混音见 audio-mix。
配置检查路径铁律:先
cd到AGENTS.md末尾给出的 Easel 项目根,确认当前目录有.env和skills/shared/scripts/,再运行注册表、check、enroll或clone。不得改用 workspace 的./shared/scripts/...,也不得以env/printenv没显示变量为由判断VOICE_BASE_URL/Key 缺失。check支持时显式传--env-file .env。
选 provider 并在 .env 填 key,再 check 离线校验:
python skills/shared/scripts/voice_clone.py check --provider minimax --env-file .env| provider | 服务 | .env 需配 |
|---|---|---|
dashscope | 阿里 CosyVoice 声音复刻 | DASHSCOPE_API_KEY(可选 DASHSCOPE_TTS_MODEL/DASHSCOPE_BASE_URL) |
minimax | MiniMax 语音克隆 | MINIMAX_API_KEY、MINIMAX_GROUP_ID(可选 MINIMAX_MODEL) |
fish-audio | Fish Audio | FISH_API_KEY(可选 FISH_BASE_URL) |
openai-compatible | OpenAI 兼容 /audio/speech | VOICE_API_KEY、VOICE_BASE_URL(预置 voice,非零样本克隆) |
gemini | Google Gemini TTS | GEMINI_API_KEY(可选 GEMINI_TTS_MODEL/GEMINI_VOICE/GEMINI_BASE_URL) |
⚠️ 各 provider 依公开 API 文档实现,端点/模型名可用 env 覆盖以适配实际参数。
执行前先跑 model_registry.py configured --group voice --env-file .env。只有一个可用时显式选择;多个可用且用户没点名时,列出 provider/模型询问本次使用哪个,不按 VOICE_PROVIDER 擅自选择。
| 字段 | 必填 | 说明 |
|---|---|---|
| 语音样本 | 克隆时必填 | 本人清晰无噪的语音(一般 10s-1min,具体看 provider 要求) |
| 文案 | 合成时必填 | 要用克隆音色说出来的文字 |
outputs/主题名/)脚本路径(相对项目根):skills/shared/scripts/voice_clone.py(各子命令支持 -h)。
# minimax:上传样本文件
python skills/shared/scripts/voice_clone.py enroll --provider minimax \
--sample me.mp3 --name my_voice
# dashscope:用公网可访问的样本 URL
python skills/shared/scripts/voice_clone.py enroll --provider dashscope \
--sample-url https://.../me.wav --name myv(fish-audio 用已有 model_id 或内联参考音频,openai-compatible 用预置 voice 名,无需 enroll。)
python skills/shared/scripts/voice_clone.py clone --provider minimax \
--voice-id my_voice --text "大家好,欢迎来到我的频道" --speed 1.0 \
-o outputs/主题名/vo.mp3fish-audio 也可直接给参考音频:--sample ref.mp3 --sample-text "参考音频的文字"。
check 确认 key,再 enroll,再 clone。outputs/主题名/。主流声音克隆云服务:阿里 CosyVoice(声音复刻)、MiniMax(语音克隆 + T2A)、Fish Audio、 OpenAI 兼容 TTS。本地零样本克隆(GPT-SoVITS/CosyVoice 本地)需 GPU,故走云端 provider。
© ZJU-REAL, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in skills/openclaw/voice-clone of ZJU-REAL/Easel.
Open the folder on GitHubat commit 278f420
Voice Clone next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Voice Clone this skillZJU-REAL/Easel | 3.4k | — | ~715 | Automated safety check: Notes | Apache-2.0 | |
| Voice Clone Ttsnpc-live/clawfirm | 156 | — | ~1.1k | Automated safety check: Pass | None | |
| KrillinAI CLI Operatorkrillinai/OpenCreator | 13k | — | ~869 | Automated safety check: Pass | Apache-2.0 | |
| Web Video PresentationConardLi/garden-skills | 13k | — | ~3.5k | Automated safety check: Pass | MIT | |
| Video Translatorshang-zhu/violin | 1.1k | — | ~1k | Automated safety check: Notes | MIT | |
| 9Router Text to Speechdecolua/9router | 31k | — | ~765 | Automated safety check: Pass | MIT |
npc-live/clawfirm
声纹克隆和语音合成。上传音频样本克隆声纹,用克隆声纹或预设声纹生成语音。支持多个后端:MiniMax、ElevenLabs、Fish Audio、Azure TTS、OpenAI TTS。支持情绪控制、语速调整、批量生成。触发词:语音合成、TTS、声纹克隆、voice clone、text to speech、配音、旁白。
krillinai/OpenCreator
Routes agents to the right KrillinAI command for subtitles, dubbing, video rendering, covers and speech, and explains how to read its JSON and manifest output.
ConardLi/garden-skills
Turns an article or spoken script into a click-through, full-screen 16:9 web presentation that looks like a video, with optional synthesized narration.
shang-zhu/violin
Dub a video into another language and generate subtitles using the default Together + Cartesia stack.
decolua/9router
Turns text into spoken audio through a 9Router server, choosing among voices from OpenAI, ElevenLabs, Deepgram, Edge TTS and other providers.
Bomx/super-video-maker-skill
End-to-end AI video production skill for agentic frameworks.
ZJU-REAL/Easel
微信公众号文章排版引擎:把 Markdown / Word(.docx) / PDF / 纯文本转成可直接粘贴进公众号编辑器的 HTML,自动章节编号、关键词标记、引言卡、目录、代码块、图片/GIF、作者签名;主题从 references/theme-index.md…
ZJU-REAL/Easel
微信公众号文章自动创作与发布工具。给定参考文章、文字或文档,自动搜索整理全网相关信息、生成图文并茂的公众号文章,并发布到微信公众号草稿箱。特别强调反 AI 检测写作。
ZJU-REAL/Easel
社媒卡片视觉设计系统:提供配色、中文字体层级、满画幅布局、品类骨架和死空白/密度质检,避免模板化 PPT 与廉价 AI 感。
ZJU-REAL/Easel
生成电商商品视觉方案:主图概念、场景图、详情页视觉方向和 AI 生图 Prompt. An agent skill from ZJU-REAL/Easel.
ZJU-REAL/Easel
将数据或文字内容转化为可视化信息图,支持静态(AntV)和动画 GIF 两种模式。当用户需要制作信息图、数据可视化、流程图、对比图、动画图表、GIF 图表、思维导图、SWOT 分析图时调用。本地渲染信息图/GIF 动画;要单张静态图片 URL 用 chart-visualization,要 CSV/JSON→整页报告用 data-report
ZJU-REAL/Easel
长篇小说/网文连载创作:从世界观、人设和三级大纲写到逐章正文,并用文件化状态维护伏笔、前情和跨章一致性. An agent skill from ZJU-REAL/Easel.
Categories
上传本人语音样本克隆专属音色,再用它合成口播、旁白或带货语音。当用户说“声音克隆、克隆/复刻我的声音、用我的声音配音、定制专属音色”时使用,需要用户自备云端 provider 凭证。使用公共现成音色时改用 tts-voiceover。. Voice Clone is an agent skill from ZJU-REAL/Easel.
Voice Clone fits situations like: tasks that involve Text to speech and voice.
Run `npx skills add ZJU-REAL/Easel --skill voice-clone -a claude-code`. Or copy the skill folder (skills/openclaw/voice-clone in ZJU-REAL/Easel) into .claude/skills/voice-clone in your project. Claude Code loads it when a task matches its description.
Run `npx skills add ZJU-REAL/Easel --skill voice-clone -a codex`. Or copy the skill folder (skills/openclaw/voice-clone in ZJU-REAL/Easel) into .agents/skills/voice-clone in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add ZJU-REAL/Easel --skill voice-clone -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/voice-clone, .gemini/skills/voice-clone, .github/skills/voice-clone and .opencode/skills/voice-clone in your project.
Going by SKILL.md and its folder, Voice Clone needs the command-line tools its instructions call (python) and credentials named DASHSCOPE_API_KEY, MINIMAX_API_KEY, FISH_API_KEY and VOICE_API_KEY. Our summary lists: Python 3; A credential in DASHSCOPE_API_KEY; A credential in MINIMAX_API_KEY.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found notes only (mentions a .env file), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.
Voice Clone is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 715 tokens (SKILL.md is roughly 2.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Voice Clone: Voice Clone Tts (npc-live/clawfirm, 156 stars), KrillinAI CLI Operator (krillinai/OpenCreator, 13k stars), Web Video Presentation (ConardLi/garden-skills, 13k stars) and Video Translator (shang-zhu/violin, 1.1k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
ZJU-REAL (a GitHub organization) maintains it in ZJU-REAL/Easel, which has 3,376 GitHub stars. The repository holds 114 skills in this directory. The repository was last updated on October 9, 2026.
Source: ZJU-REAL/Easel on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.