AI Video Script
zrt-ai-lab/opencode-skills
A skill your agent uses when a request asks for a Chinese-first AI video script with shot plans, image prompts, narration, subtitles, or handoff contracts for scene generation, image generation…
让 Agent 变成能语音对话的机器人:语音文件转文字(支持微信 silk 格式)+ 多音色人格回复(Edge TTS 免费中文音色),全本地零 API 成本。Voice persona chat: transcribe voice messages (incl.
$ npx skills add davepoon/buildwithclaude --skill voice-persona -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install davepoon/buildwithclaude voice-persona --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/davepoon/buildwithclaude.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/all-skills/skills/voice-persona .claude/skills/voice-persona && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "voice-persona" agent skill from https://github.com/davepoon/buildwithclaude/tree/main/plugins/all-skills/skills/voice-persona into .claude/skills/voice-persona/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-persona", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/davepoon/buildwithclaude/tree/main/plugins/all-skills/skills/voice-personaType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add davepoon/buildwithclaude --skill voice-persona -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install davepoon/buildwithclaude voice-persona --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/davepoon/buildwithclaude.git skills-src && mkdir -p .agents/skills && cp -r skills-src/plugins/all-skills/skills/voice-persona .agents/skills/voice-persona && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "voice-persona" agent skill from https://github.com/davepoon/buildwithclaude/tree/main/plugins/all-skills/skills/voice-persona into .agents/skills/voice-persona/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-persona", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add davepoon/buildwithclaude --skill voice-persona -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install davepoon/buildwithclaude voice-persona --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/davepoon/buildwithclaude.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/plugins/all-skills/skills/voice-persona .cursor/skills/voice-persona && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "voice-persona" agent skill from https://github.com/davepoon/buildwithclaude/tree/main/plugins/all-skills/skills/voice-persona into .cursor/skills/voice-persona/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-persona", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/davepoon/buildwithclaude.git --path plugins/all-skills/skills/voice-persona--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add davepoon/buildwithclaude --skill voice-persona -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install davepoon/buildwithclaude voice-persona --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/davepoon/buildwithclaude.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/plugins/all-skills/skills/voice-persona .gemini/skills/voice-persona && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "voice-persona" agent skill from https://github.com/davepoon/buildwithclaude/tree/main/plugins/all-skills/skills/voice-persona into .gemini/skills/voice-persona/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-persona", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install davepoon/buildwithclaude voice-personaInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add davepoon/buildwithclaude --skill voice-persona -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/davepoon/buildwithclaude.git skills-src && mkdir -p .github/skills && cp -r skills-src/plugins/all-skills/skills/voice-persona .github/skills/voice-persona && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "voice-persona" agent skill from https://github.com/davepoon/buildwithclaude/tree/main/plugins/all-skills/skills/voice-persona into .github/skills/voice-persona/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-persona", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add davepoon/buildwithclaude --skill voice-persona -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install davepoon/buildwithclaude voice-persona --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/davepoon/buildwithclaude.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/plugins/all-skills/skills/voice-persona .opencode/skills/voice-persona && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "voice-persona" agent skill from https://github.com/davepoon/buildwithclaude/tree/main/plugins/all-skills/skills/voice-persona into .opencode/skills/voice-persona/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-persona", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
voice-persona让 Agent 变成能语音对话的机器人:语音文件转文字(支持微信 silk 格式)+ 多音色人格回复(Edge TTS 免费中文音色),全本地零 API 成本。Voice persona chat: transcribe voice messages (incl.
Voice Persona is an agent skill from davepoon/buildwithclaude. 让 Agent 变成能语音对话的机器人:语音文件转文字(支持微信 silk 格式)+ 多音色人格回复(Edge TTS 免费中文音色),全本地零 API 成本。Voice persona chat: transcribe voice messages (incl. WeChat silk) and reply in persona voices.
Its SKILL.md is about 950 tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/voice_persona.py`).
It sits in Media & Creative, covering Text to speech and voice, Messaging and chat bots and Transcription. It works with WeChat and Telegram. The repository describes itself as: A single hub to find Claude Skills, Agents, Commands, Hooks, Plugins, and Marketplace collections to extend Claude Code, Claude Desktop, Agent SDK and OpenClaw. The licence is MIT.
Read from SKILL.md and the folder at commit 616deb5. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships 1 file in scripts/ (Python), which the agent can run.
Shell commands in SKILL.md call:
python3pipFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use pip, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names these keys or tokens, usually read from environment variables:
LLM_API_KEYFrom names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Voice Persona loads about 947 tokens when it runs. Until then it costs about 47 tokens; SKILL.md has 221 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
The full file from davepoon/buildwithclaude at commit 616deb5, republished under its MIT licence (© davepoon). 221 words, ~947 tokens.
.claude/skills/voice-persona/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.把 Agent 变成"能听懂语音、会用人格声音回话"的机器人——首个 IM 语音双工闭环技能。
Voice chat for agents: transcribe voice (incl. WeChat silk) → reply with persona voices → back to WeChat voice bubbles. First IM two-way voice skill.
| 方案 | 问题 |
|---|---|
| jozhn/wechat-voice-decode-skill | 只做单向解码转写,无 TTS/人格/语音回复,绑死 openclaw 路径 |
| zhayujie/chatgpt-on-wechat (46k★) | 语音依赖云端 ASR/TTS,个人微信通道封号风险,无 silk 原生闭环 |
| whisperbot / Telegram 转写 bot 族 | 全部单向"语音→文字",无多音色人格语音回复 |
| SillyTavern (33k★) | 角色音色强大但只在自家 UI,不接微信/Telegram 语音消息 |
| 微软云 voice skills(Azure/OpenAI TTS) | 单项能力 + 付费云 API,无 IM 场景 |
别人做"听写"或"朗读",voice-persona 做"在微信/Telegram 里用带人格的嗓音说话"。
从本技能目录运行(本仓库已自带脚本,无需额外 clone):
cd plugins/all-skills/skills/voice-persona # 或直接进入 voice-persona 技能目录
python3 scripts/voice_persona.py demo
# 🎙 语音 → 文字 → [元气少女/沉稳大叔/新闻播报] 三音色回复音频
# ✅ 全链路通过:语音输入 → 转写 → 人格回复 → 语音输出pip install faster-whisper pilk edge-tts # 首次使用安装
# 可选:--llm 增强口吻需要 LLM_API_KEY / OpenAI 兼容 URL(与本集合其他技能一致)所有命令从技能目录内以相对路径调用脚本(无需安装到系统 PATH):
# 1. 语音 → 文字(微信语音直接传 .silk 文件即可)
python3 scripts/voice_persona.py stt wechat_voice.silk
python3 scripts/voice_persona.py stt meeting.m4a --model small
# 2. 文字 → 人格回复 → 语音 mp3
python3 scripts/voice_persona.py speak "明天记得交报告" --persona yunjian --out reply.mp3
# 3. 只取人格化文本(接你自己的 TTS/IM)
python3 scripts/voice_persona.py chat "明天记得交报告" --persona xiaoyi
# 4. 列出人格
python3 scripts/voice_persona.py list
# 5. 音频 → 微信 silk 语音(可回发微信语音气泡)
python3 scripts/voice_persona.py to_silk reply.mp3 --out reply.silk
# 6. 一条命令双向闭环:人格语音 → 微信格式
python3 scripts/voice_persona.py speak "明天早上十点开会" --persona xiaoyi --out r.mp3
python3 scripts/voice_persona.py to_silk r.mp3 # → r.silk(#!SILK_V3)本仓库自带脚本可直接运行;若使用 Hermes Agent,也可从技能源仓库安装以自动接入:
hermes skills install jiawood2006/hermes-skills/skills/voice-persona在 voice_persona.py 顶部 PERSONAS / STYLE_WORDS 增加即可:
| key | 人格 | Edge TTS 音色 | 口吻 |
|---|---|---|---|
| xiaoyi | 元气少女 · 小伊 | zh-CN-XiaoyiNeural | 嘿嘿、啦/呀 |
| xiaoxuan | 温柔知心 · 晓萱 | zh-CN-XiaoxuanNeural | 别担心、慢慢来 |
| yunjian | 沉稳大叔 · 云健 | zh-CN-YunjianNeural | 直说、结论先行 |
| yunyang | 新闻播报 · 云扬 | zh-CN-YunyangNeural | 播报、条理 |
| yunxi | 阳光伙伴 · 云希 | zh-CN-YunxiNeural | 加油、没问题 |
| xiaochen | 随性好友 · 晓辰 | zh-CN-XiaochenNeural | 口语化、像朋友 |
STT 接入:把平台收到的语音消息文件交给 stt 子命令 → 拿到文字进对话流。
Hermes 的 stt.provider: local_command + HERMES_LOCAL_STT_COMMAND 环境变量可直接把微信语音自动接入(指向本技能 scripts/ 下的脚本绝对路径):
HERMES_LOCAL_STT_COMMAND=<python> <...>/skills/voice-persona/scripts/voice_persona.py stt {input_path} --model {model} --output_dir {output_dir} --language {language}TTS 接入:人格化文本 → speak 出 mp3 → 平台发送语音。
pilk.decode() 输出的是无 RIFF 头的裸 PCM,whisper 读不了——要用 pilk.silk_to_wav()(输出标准 wav,已实测)© davepoon, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 1 other file (scripts) in plugins/all-skills/skills/voice-persona of davepoon/buildwithclaude.
Open the folder on GitHubat commit 616deb5
Voice Persona next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Voice Persona this skilldavepoon/buildwithclaude | 3.6k | — | ~947 | Automated safety check: Pass | MIT | |
| AI Video Scriptzrt-ai-lab/opencode-skills | 287 | — | ~1.7k | Automated safety check: Pass | None | |
| FeedgrabiBigQiang/feedgrab | 614 | — | ~2k | Automated safety check: Pass | MIT | |
| Lingzaoatian-create/lingzao-skill | 296 | 1 repos | ~8.8k | Automated safety check: Pass | MIT | |
| Ffmpeg Best Practicelinyqh/speclip-skills | 110 | — | ~2.3k | Automated safety check: Pass | None | |
| Voice Memoletta-ai/lettabot | 327 | — | ~487 | Automated safety check: Pass | Apache-2.0 |
zrt-ai-lab/opencode-skills
A skill your agent uses when a request asks for a Chinese-first AI video script with shot plans, image prompts, narration, subtitles, or handoff contracts for scene generation, image generation…
iBigQiang/feedgrab
Universal content grabber — fetch any URL and return structured Markdown.
atian-create/lingzao-skill
Use Lingzao creator-content tools for Xiaohongshu/XHS, Douyin, and WeChat official-account public content.
linyqh/speclip-skills
Lean FFmpeg playbook for reliable video compression, WeChat-compatible MP4 export, clip stitching, audio mixing, subtitle handling, and quick fallback decisions.
letta-ai/lettabot
Reply with voice memos using text-to-speech. An agent skill from letta-ai/lettabot.
kangarooking/kangarooking-skills
Download or open videos and recover platform captions, audio transcripts, keyframes, screen text, visual facts, and editing observations as a plain multimodaltranscript.md.
davepoon/buildwithclaude
Build, update, and apply iOS design specifications using Apple Human Interface Guidelines (HIG) source data.
davepoon/buildwithclaude
Download YouTube videos with customizable quality and format options.
davepoon/buildwithclaude
A skill your agent uses when the user asks to "analyze video", "watch this video", "what happens in this video", "describe this clip", "review this footage", "classify these videos", "compare…
davepoon/buildwithclaude
Discover Atlas Cloud image and video models, inspect their live schemas, and submit one confirmed media generation request with bounded GET polling.
davepoon/buildwithclaude
面向没有编程经验的用户,把想法做成可试用的浏览器插件,并完成检查、商店材料、审核提交和上线验证;也用于继续已有插件、排错和发布新版。用户说“帮我做个插件”“把插件上架”“继续我的插件”时使用。普通网站开发、仅查询插件知识不触发。
davepoon/buildwithclaude
Toolkit for creating animated GIFs optimized for Slack, with validators for size constraints and composable animation primitives.
让 Agent 变成能语音对话的机器人:语音文件转文字(支持微信 silk 格式)+ 多音色人格回复(Edge TTS 免费中文音色),全本地零 API 成本。Voice persona chat: transcribe voice messages (incl. Voice Persona is an agent skill from davepoon/buildwithclaude. 让 Agent 变成能语音对话的机器人:语音文件转文字(支持微信 silk 格式)+ 多音色人格回复(Edge TTS 免费中文音色),全本地零 API 成本。Voice persona chat: transcribe voice messages (incl.
Voice Persona fits situations like: tasks that involve Text to speech and voice; tasks that involve Messaging and chat bots; tasks that involve Transcription.
Run `npx skills add davepoon/buildwithclaude --skill voice-persona -a claude-code`. Or copy the skill folder (plugins/all-skills/skills/voice-persona in davepoon/buildwithclaude) into .claude/skills/voice-persona in your project. Claude Code loads it when a task matches its description.
Run `npx skills add davepoon/buildwithclaude --skill voice-persona -a codex`. Or copy the skill folder (plugins/all-skills/skills/voice-persona in davepoon/buildwithclaude) into .agents/skills/voice-persona in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add davepoon/buildwithclaude --skill voice-persona -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/voice-persona, .gemini/skills/voice-persona, .github/skills/voice-persona and .opencode/skills/voice-persona in your project.
Going by SKILL.md and its folder, Voice Persona needs Python for the scripts in its folder, the command-line tools its instructions call (python3 and pip) and credentials named LLM_API_KEY. Our summary lists: Python 3; A credential in LLM_API_KEY.
SKILL.md contains no URLs. Its commands use pip, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
Voice Persona is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.
About 947 tokens (SKILL.md is roughly 3.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Voice Persona: AI Video Script (zrt-ai-lab/opencode-skills, 287 stars), Feedgrab (iBigQiang/feedgrab, 614 stars), Lingzao (atian-create/lingzao-skill, 296 stars) and Ffmpeg Best Practice (linyqh/speclip-skills, 110 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
davepoon (a GitHub user) maintains it in davepoon/buildwithclaude, which has 3,610 GitHub stars. The repository holds 247 skills in this directory. The repository was last updated on October 9, 2026.
Source: davepoon/buildwithclaude on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.