Agent skill

Aliyun Videoretalk

by cinience in cinience/alicloud-skills

A skill your agent uses when replacing lip sync in existing videos with Alibaba Cloud Model Studio VideoRetalk (videoretalk).

MITAuto-check passedMedia & Creative

Install Aliyun Videoretalk

skills CLI
$ npx skills add cinience/alicloud-skills --skill aliyun-videoretalk -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install cinience/alicloud-skills aliyun-videoretalk --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/cinience/alicloud-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/ai/video/aliyun-videoretalk .claude/skills/aliyun-videoretalk && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
aliyun-videoretalk
GitHub stars
397
Token cost
~735 tokens
SKILL.md length
258 words
Files
4 (incl. scripts, references)
Skills in repo
96
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses when replacing lip sync in existing videos with Alibaba Cloud Model Studio VideoRetalk (videoretalk).

  • Replacing lip sync in existing videos with Alibaba Cloud Model Studio VideoRetalk (videoretalk)
  • SKILL.md covers Validation, Output And Evidence, Critical model names and Prerequisites, plus 6 more sections
  • Runs Python scripts from its folder; calls python; reaches dashscope.aliyuncs.com; needs DASHSCOPE_API_KEY
  • Creating dubbed videos

What it does

Aliyun Videoretalk is an agent skill from cinience/alicloud-skills. Use when replacing lip sync in existing videos with Alibaba Cloud Model Studio VideoRetalk (videoretalk). Use when creating dubbed videos, replacing narration, or synchronizing a talking-head video to a new speech track.

Its SKILL.md is about 740 tokens, which your agent loads only when the skill is triggered. The skill folder holds 6 other files, including scripts and reference files (for example `agents/openai.yaml`, `references/sources.md` and `scripts/prepare_retalk_request.py`).

It sits in Media & Creative, covering Text to speech and voice. It works with Alibaba Cloud. The repository describes itself as: alibaba cloud skills,qwen ,wan and all skills. The licence is MIT.

When your agent uses it

  • Replacing lip sync in existing videos with Alibaba Cloud Model Studio VideoRetalk (videoretalk)
  • Creating dubbed videos
  • Replacing narration
  • Synchronizing a talking-head video to a new speech track

Example prompts

  • “/aliyun-videoretalk”

Requirements

  • Python 3
  • A credential in DASHSCOPE_API_KEY

What it can do on your machine

Read from SKILL.md and the folder at commit 1818263. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • dashscope.aliyuncs.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • DASHSCOPE_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Aliyun Videoretalk loads about 735 tokens when it runs, and up to ~806 if it reads all its reference files. Until then it costs about 60 tokens; SKILL.md has 258 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~60
When it runs · the whole SKILL.md, loaded when a task matches
~735
With references · SKILL.md plus every file in references/, read only if the agent opens them
~806

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from cinience/alicloud-skills at commit 1818263, republished under its MIT licence (© cinience). 258 words, ~735 tokens.

Download SKILL.mdSave it as .claude/skills/aliyun-videoretalk/SKILL.md (or your agent's skills folder). This skill also uses 3 other files; get the full folder from GitHub.
name
aliyun-videoretalk
description
Use when replacing lip sync in existing videos with Alibaba Cloud Model Studio VideoRetalk (`videoretalk`). Use when creating dubbed videos, replacing narration, or synchronizing a talking-head video to a new speech track.
version
1.0.0

Category: provider

Model Studio VideoRetalk

Validation

bash
mkdir -p output/aliyun-videoretalk
python -m py_compile skills/ai/video/aliyun-videoretalk/scripts/prepare_retalk_request.py && echo "py_compile_ok" > output/aliyun-videoretalk/validate.txt

Pass criteria: command exits 0 and output/aliyun-videoretalk/validate.txt is generated.

Output And Evidence

  • Save normalized request payloads, target face selection settings, and task polling snapshots under output/aliyun-videoretalk/.
  • Record the exact video/audio input URLs and whether video_extension was enabled.

Use VideoRetalk when the input is already a person video and the job is to replace lip sync with a new speech track.

Critical model names

Use this exact model string:

  • videoretalk

Prerequisites

  • This model currently only supports China mainland (Beijing).
  • API is HTTP async only; there is no online console experience.
  • Set DASHSCOPE_API_KEY in your environment, or add dashscope_api_key to ~/.alibabacloud/credentials.

Normalized interface (video.retalk)

Request
  • model (string, optional): default videoretalk
  • video_url (string, required)
  • audio_url (string, required)
  • ref_image_url (string, optional): target face when input video contains multiple faces
  • video_extension (bool, optional): extend video to match longer audio
  • query_face_threshold (int, optional): 120 to 200
Response
  • task_id (string)
  • task_status (string)
  • video_url (string, when finished)
  • usage (object, optional)

Endpoint and execution model

  • Submit task: POST https://dashscope.aliyuncs.com/api/v1/services/aigc/image2video/video-synthesis/
  • Poll task: GET https://dashscope.aliyuncs.com/api/v1/tasks/{task_id}
  • HTTP calls are async only and must set header X-DashScope-Async: enable.

Quick start

bash
python skills/ai/video/aliyun-videoretalk/scripts/prepare_retalk_request.py \
  --video-url "https://example.com/talking-head.mp4" \
  --audio-url "https://example.com/new-voice.wav" \
  --video-extension

Operational guidance

  • Keep input videos front-facing and close enough for stable face tracking.
  • If the video contains multiple faces, provide ref_image_url to anchor the intended target.
  • If the new audio is longer than the input video, decide explicitly whether to extend the picture track or truncate the audio.
  • URLs must be public HTTP/HTTPS links; local file paths are not accepted by the API.

Output location

  • Default output: output/aliyun-videoretalk/request.json
  • Override base dir with OUTPUT_DIR.

References

  • references/sources.md

© cinience, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 3 other files (scripts, references) in skills/ai/video/aliyun-videoretalk of cinience/alicloud-skills.

  • SKILL.md
  • agents/openai.yaml
  • references/sources.md
  • scripts/prepare_retalk_request.py

Open the folder on GitHubat commit 1818263

Compare with similar skills

Aliyun Videoretalk next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Aliyun Videoretalk compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Aliyun Videoretalk this skillcinience/alicloud-skills397—~735Automated safety check: PassMIT
Bailian Media Generationmodelstudioai/cli542—~2kAutomated safety check: PassApache-2.0
Dashscopecalesthio/OpenMontage66k—~1.5kAutomated safety check: NotesAGPL-3.0
Musictadaspetra/loop2962 repos~827Automated safety check: PassMIT
Sound Effectstadaspetra/loop2962 repos~1.1kAutomated safety check: PassMIT
Edu Math Videowy51ai/edulab1.4k—~2.5kAutomated safety check: NotesApache-2.0

Similar skills

  • Bailian Media Generation

    modelstudioai/cli

    Chinese-language entry point into Alibaba Cloud Bailian's image, video and speech generation and understanding, routed through separate image, video, speech and vision commands.

    542 GitHub stars~2k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Dashscope

    calesthio/OpenMontage

    DashScope (Alibaba Cloud Bailian / 阿里云百炼) integration — image generation (qwen-image-2.0-pro), text-to-speech (qwen3-tts-flash), and ASR with word-level timestamps (qwen3-asr-flash-filetrans).

    66k GitHub stars~1.5k tokensUpdated 6 days ago
    Media & CreativeAuto-check: notes
  • Music

    tadaspetra/loop

    Generate music using ElevenLabs Music API. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 2 repos~827 tokens
    Media & CreativeAuto-check passed
  • Sound Effects

    tadaspetra/loop

    Generate sound effects from text descriptions using ElevenLabs.

    296 GitHub starsUsed in 2 repos~1.1k tokens
    Media & CreativeAuto-check passed
  • Edu Math Video

    wy51ai/edulab

    A skill your agent uses when asked to make an explainer / walkthrough video (讲解视频、解题视频、例题精讲、微课) for a math problem (数学题, geometry, algebra, functions, motion/行程 problems), from a problem screenshot…

    1.4k GitHub stars~2.5k tokensUpdated today
    Media & CreativeAuto-check: notes
  • Book Video Factory

    bytec-ai/book-video-factory

    通用的多账号图书短视频生产工作流。用于用户希望建立图书号项目目录、配置账号级片头/声音/BGM/视觉规范,或只提供一本书后依次完成资料研究、口播稿、分镜、图片、配音、字幕、预览与成片导出。适用于新建工作区、批量管理多个账号、继续已有单书任务和检查生产状态;不绑定特定研究、图片、TTS、转录或视频渲染供应商。

    321 GitHub stars~1.4k tokensUpdated 2 mo ago
    Media & CreativeAuto-check: notes

More from cinience/alicloud-skills

All 96 skills in this repo
  • Aliyun Skill Creator

    cinience/alicloud-skills

    A skill your agent uses when creating, migrating, or optimizing skills for this alicloud-skills repository.

    397 GitHub stars~2.8k tokensUpdated 1 mo ago
    Auto-check passed
  • Alicloud Acs Agent Sandbox

    cinience/alicloud-skills

    Bootstrap, create, connect to, operate, secure, scale, upgrade, troubleshoot, inspect, and tear down Alibaba Cloud Container Compute Service (ACS) Agent Sandbox environments.

    397 GitHub stars~2.7k tokensUpdated 1 mo ago
    Auto-check passed
  • Alicloud Acs Cluster

    cinience/alicloud-skills

    Create, inspect, connect to, inventory, and delete Alibaba Cloud Container Compute Service (ACS) clusters through the official CS OpenAPI.

    397 GitHub stars~1.7k tokensUpdated 1 mo ago
    Auto-check passed
  • Aliyun Adb Mysql

    cinience/alicloud-skills

    A skill your agent uses when managing Alibaba Cloud AnalyticDB for MySQL (ADB) via OpenAPI/SDK, including the user needs AnalyticDB resource lifecycle and configuration operations, status checks, or…

    397 GitHub stars~708 tokensUpdated 1 mo ago
    Auto-check passed
  • Aliyun Aicontent Generate

    cinience/alicloud-skills

    A skill your agent uses when managing Alibaba Cloud AIContent (AiContent) via OpenAPI/SDK, including the user needs AI content generation or content workflow operations in Alibaba Cloud, including…

    397 GitHub stars~734 tokensUpdated 1 mo ago
    Auto-check passed
  • Aliyun Aimiaobi Generate

    cinience/alicloud-skills

    A skill your agent uses when managing Alibaba Cloud Quan Miao (AiMiaoBi) via OpenAPI/SDK, including the user asks for Alibaba Cloud MiaoBi content operations, including listing resources…

    397 GitHub stars~724 tokensUpdated 1 mo ago
    Auto-check passed

Works with

Questions about Aliyun Videoretalk

What does Aliyun Videoretalk do?

A skill your agent uses when replacing lip sync in existing videos with Alibaba Cloud Model Studio VideoRetalk (videoretalk). Aliyun Videoretalk is an agent skill from cinience/alicloud-skills. Use when replacing lip sync in existing videos with Alibaba Cloud Model Studio VideoRetalk (videoretalk).

When should I use Aliyun Videoretalk?

Aliyun Videoretalk fits situations like: replacing lip sync in existing videos with Alibaba Cloud Model Studio VideoRetalk (videoretalk); creating dubbed videos; replacing narration; synchronizing a talking-head video to a new speech track.

How do I install Aliyun Videoretalk in Claude Code?

Run `npx skills add cinience/alicloud-skills --skill aliyun-videoretalk -a claude-code`. Or copy the skill folder (skills/ai/video/aliyun-videoretalk in cinience/alicloud-skills) into .claude/skills/aliyun-videoretalk in your project. Claude Code loads it when a task matches its description.

How do I install Aliyun Videoretalk in Codex?

Run `npx skills add cinience/alicloud-skills --skill aliyun-videoretalk -a codex`. Or copy the skill folder (skills/ai/video/aliyun-videoretalk in cinience/alicloud-skills) into .agents/skills/aliyun-videoretalk in your project. Codex loads it when a task matches its description.

Can I use Aliyun Videoretalk in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add cinience/alicloud-skills --skill aliyun-videoretalk -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/aliyun-videoretalk, .gemini/skills/aliyun-videoretalk, .github/skills/aliyun-videoretalk and .opencode/skills/aliyun-videoretalk in your project.

What does Aliyun Videoretalk need to run?

Going by SKILL.md and its folder, Aliyun Videoretalk needs Python for the scripts in its folder, the command-line tools its instructions call (python) and credentials named DASHSCOPE_API_KEY. Our summary lists: Python 3; A credential in DASHSCOPE_API_KEY.

Does Aliyun Videoretalk access the network?

SKILL.md names 1 domain. In commands or code: dashscope.aliyuncs.com; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.

Is Aliyun Videoretalk safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Aliyun Videoretalk use?

Aliyun Videoretalk is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Aliyun Videoretalk use?

About 735 tokens (SKILL.md is roughly 2.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 71 tokens, read only when the agent opens those files.

What are the alternatives to Aliyun Videoretalk?

Skills that share tags, products or a category with Aliyun Videoretalk: Bailian Media Generation (modelstudioai/cli, 542 stars), Dashscope (calesthio/OpenMontage, 66k stars), Music (tadaspetra/loop, 296 stars) and Sound Effects (tadaspetra/loop, 296 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Aliyun Videoretalk?

cinience (a GitHub user) maintains it in cinience/alicloud-skills, which has 397 GitHub stars. The repository holds 96 skills in this directory. The repository was last updated on August 11, 2026.

Source: cinience/alicloud-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.