Agent skill

Aliyun Wan Digital Human

by cinience in cinience/alicloud-skills

A skill your agent uses when generating talking, singing, or presentation videos from a single character image and audio with Alibaba Cloud Model Studio digital-human model wan2.2-s2v.

MITAuto-check passedMedia & Creative

Install Aliyun Wan Digital Human

skills CLI
$ npx skills add cinience/alicloud-skills --skill aliyun-wan-digital-human -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install cinience/alicloud-skills aliyun-wan-digital-human --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/cinience/alicloud-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/ai/video/aliyun-wan-digital-human .claude/skills/aliyun-wan-digital-human && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
aliyun-wan-digital-human
GitHub stars
397
Token cost
~701 tokens
SKILL.md length
228 words
Files
4 (incl. scripts, references)
Skills in repo
96
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses when generating talking, singing, or presentation videos from a single character image and audio with Alibaba Cloud Model Studio digital-human model wan2.2-s2v.

  • Generating talking
  • SKILL.md covers Validation, Output And Evidence, Critical model names and Prerequisites, plus 5 more sections
  • Runs Python scripts from its folder; calls python; needs DASHSCOPE_API_KEY
  • Presentation videos from a single character image and audio with Alibaba Cloud Model Studio digital-human model wan2.2-s2v

What it does

Aliyun Wan Digital Human is an agent skill from cinience/alicloud-skills. Use when generating talking, singing, or presentation videos from a single character image and audio with Alibaba Cloud Model Studio digital-human model wan2.2-s2v. Use when creating narrated avatar videos, singing portraits, or broadcast-style talking-head clips.

Its SKILL.md is about 700 tokens, which your agent loads only when the skill is triggered. The skill folder holds 6 other files, including scripts and reference files (for example `agents/openai.yaml`, `references/sources.md` and `scripts/prepare_digital_human_request.py`).

It sits in Media & Creative. It works with Alibaba Cloud. The repository describes itself as: alibaba cloud skills,qwen ,wan and all skills. The licence is MIT.

When your agent uses it

  • Generating talking
  • Presentation videos from a single character image and audio with Alibaba Cloud Model Studio digital-human model wan2.2-s2v
  • Creating narrated avatar videos
  • Singing portraits

Example prompts

  • “/aliyun-wan-digital-human”

Requirements

  • Python 3
  • A credential in DASHSCOPE_API_KEY

What it can do on your machine

Read from SKILL.md and the folder at commit 1818263. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • DASHSCOPE_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Aliyun Wan Digital Human loads about 701 tokens when it runs, and up to ~752 if it reads all its reference files. Until then it costs about 73 tokens; SKILL.md has 228 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~73
When it runs · the whole SKILL.md, loaded when a task matches
~701
With references · SKILL.md plus every file in references/, read only if the agent opens them
~752

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from cinience/alicloud-skills at commit 1818263, republished under its MIT licence (© cinience). 228 words, ~701 tokens.

Download SKILL.mdSave it as .claude/skills/aliyun-wan-digital-human/SKILL.md (or your agent's skills folder). This skill also uses 3 other files; get the full folder from GitHub.
name
aliyun-wan-digital-human
description
Use when generating talking, singing, or presentation videos from a single character image and audio with Alibaba Cloud Model Studio digital-human model `wan2.2-s2v`. Use when creating narrated avatar videos, singing portraits, or broadcast-style talking-head clips.
version
1.0.0

Category: provider

Model Studio Digital Human

Validation

bash
mkdir -p output/aliyun-wan-digital-human
python -m py_compile skills/ai/video/aliyun-wan-digital-human/scripts/prepare_digital_human_request.py && echo "py_compile_ok" > output/aliyun-wan-digital-human/validate.txt

Pass criteria: command exits 0 and output/aliyun-wan-digital-human/validate.txt is generated.

Output And Evidence

  • Save normalized request payloads, chosen resolution, and task polling snapshots under output/aliyun-wan-digital-human/.
  • Record image/audio URLs and whether the input image passed detection.

Use this skill for image + audio driven speaking, singing, or presenting characters.

Critical model names

Use these exact model strings:

  • wan2.2-s2v-detect
  • wan2.2-s2v

Selection guidance:

  • Run wan2.2-s2v-detect first to validate the image.
  • Use wan2.2-s2v for the actual video generation job.

Prerequisites

  • China mainland (Beijing) only.
  • Set DASHSCOPE_API_KEY in your environment, or add dashscope_api_key to ~/.alibabacloud/credentials.
  • Input audio should contain clear speech or singing, and input image should depict a clear subject.

Normalized interface (video.digital_human)

Detect Request
  • model (string, optional): default wan2.2-s2v-detect
  • image_url (string, required)
Generate Request
  • model (string, optional): default wan2.2-s2v
  • image_url (string, required)
  • audio_url (string, required)
  • resolution (string, optional): 480P or 720P
  • scenario (string, optional): talk, sing, or perform
Response
  • task_id (string)
  • task_status (string)
  • video_url (string, when finished)

Quick start

bash
python skills/ai/video/aliyun-wan-digital-human/scripts/prepare_digital_human_request.py \
  --image-url "https://example.com/anchor.png" \
  --audio-url "https://example.com/voice.mp3" \
  --resolution 720P \
  --scenario talk

Operational guidance

  • Use a portrait, half-body, or full-body image with a clear face and stable framing.
  • Match audio length to the desired output duration; the output follows the audio length up to the model limit.
  • Keep image and audio as public HTTP/HTTPS URLs.
  • If the image fails detection, do not proceed directly to video generation.

Output location

  • Default output: output/aliyun-wan-digital-human/request.json
  • Override base dir with OUTPUT_DIR.

References

  • references/sources.md

© cinience, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 3 other files (scripts, references) in skills/ai/video/aliyun-wan-digital-human of cinience/alicloud-skills.

  • SKILL.md
  • agents/openai.yaml
  • references/sources.md
  • scripts/prepare_digital_human_request.py

Open the folder on GitHubat commit 1818263

Compare with similar skills

Aliyun Wan Digital Human next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Aliyun Wan Digital Human compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Aliyun Wan Digital Human this skillcinience/alicloud-skills397—~701Automated safety check: PassMIT
Bailian Media Generationmodelstudioai/cli542—~2kAutomated safety check: PassApache-2.0
Dashscopecalesthio/OpenMontage66k—~1.5kAutomated safety check: NotesAGPL-3.0
Image GenerationQwenLM/qwen-code-examples143—~568Automated safety check: PassMIT
Cap Cinematic Demo GeneratorCapSoftware/Cap23k—~2.4kAutomated safety check: PassCustom licence
Musictadaspetra/loop2962 repos~827Automated safety check: PassMIT

Similar skills

  • Bailian Media Generation

    modelstudioai/cli

    Chinese-language entry point into Alibaba Cloud Bailian's image, video and speech generation and understanding, routed through separate image, video, speech and vision commands.

    542 GitHub stars~2k tokensUpdated 2 days ago
    Media & CreativeAuto-check passed
  • Dashscope

    calesthio/OpenMontage

    DashScope (Alibaba Cloud Bailian / 阿里云百炼) integration — image generation (qwen-image-2.0-pro), text-to-speech (qwen3-tts-flash), and ASR with word-level timestamps (qwen3-asr-flash-filetrans).

    66k GitHub stars~1.5k tokensUpdated 7 days ago
    Media & CreativeAuto-check: notes
  • Image Generation

    QwenLM/qwen-code-examples

    Image generation skill based on Alibaba Cloud DashScope, supporting the creation of high-quality hand-drawn or standard images from user descriptions.

    143 GitHub stars~568 tokensUpdated 4 mo ago
    Media & CreativeAuto-check passed
  • Turns any URL into a short cinematic product-demo video on macOS, scouting the page, recording it with virtual input, then treating the clip with Cap's 3D camera and music.

    23k GitHub stars~2.4k tokensUpdated today
    Media & CreativeAuto-check passed
  • Music

    tadaspetra/loop

    Generate music using ElevenLabs Music API. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 2 repos~827 tokens
    Media & CreativeAuto-check passed
  • Add Icon

    wei/socialify

    Add or update supported language, framework, library, tool, or platform icons in Socialify.

    2.2k GitHub stars~826 tokensUpdated 3 days ago
    Media & CreativeAuto-check passed

More from cinience/alicloud-skills

All 96 skills in this repo
  • Aliyun Skill Creator

    cinience/alicloud-skills

    A skill your agent uses when creating, migrating, or optimizing skills for this alicloud-skills repository.

    397 GitHub stars~2.8k tokensUpdated 2 mo ago
    Auto-check passed
  • Alicloud Acs Agent Sandbox

    cinience/alicloud-skills

    Bootstrap, create, connect to, operate, secure, scale, upgrade, troubleshoot, inspect, and tear down Alibaba Cloud Container Compute Service (ACS) Agent Sandbox environments.

    397 GitHub stars~2.7k tokensUpdated 2 mo ago
    Auto-check passed
  • Alicloud Acs Cluster

    cinience/alicloud-skills

    Create, inspect, connect to, inventory, and delete Alibaba Cloud Container Compute Service (ACS) clusters through the official CS OpenAPI.

    397 GitHub stars~1.7k tokensUpdated 2 mo ago
    Auto-check passed
  • Aliyun Adb Mysql

    cinience/alicloud-skills

    A skill your agent uses when managing Alibaba Cloud AnalyticDB for MySQL (ADB) via OpenAPI/SDK, including the user needs AnalyticDB resource lifecycle and configuration operations, status checks, or…

    397 GitHub stars~708 tokensUpdated 2 mo ago
    Auto-check passed
  • Aliyun Aicontent Generate

    cinience/alicloud-skills

    A skill your agent uses when managing Alibaba Cloud AIContent (AiContent) via OpenAPI/SDK, including the user needs AI content generation or content workflow operations in Alibaba Cloud, including…

    397 GitHub stars~734 tokensUpdated 2 mo ago
    Auto-check passed
  • Aliyun Aimiaobi Generate

    cinience/alicloud-skills

    A skill your agent uses when managing Alibaba Cloud Quan Miao (AiMiaoBi) via OpenAPI/SDK, including the user asks for Alibaba Cloud MiaoBi content operations, including listing resources…

    397 GitHub stars~724 tokensUpdated 2 mo ago
    Auto-check passed

Works with

Questions about Aliyun Wan Digital Human

What does Aliyun Wan Digital Human do?

A skill your agent uses when generating talking, singing, or presentation videos from a single character image and audio with Alibaba Cloud Model Studio digital-human model wan2.2-s2v. Aliyun Wan Digital Human is an agent skill from cinience/alicloud-skills.2-s2v.

When should I use Aliyun Wan Digital Human?

Aliyun Wan Digital Human fits situations like: generating talking; presentation videos from a single character image and audio with Alibaba Cloud Model Studio digital-human model wan2.2-s2v; creating narrated avatar videos; singing portraits.

How do I install Aliyun Wan Digital Human in Claude Code?

Run `npx skills add cinience/alicloud-skills --skill aliyun-wan-digital-human -a claude-code`. Or copy the skill folder (skills/ai/video/aliyun-wan-digital-human in cinience/alicloud-skills) into .claude/skills/aliyun-wan-digital-human in your project. Claude Code loads it when a task matches its description.

How do I install Aliyun Wan Digital Human in Codex?

Run `npx skills add cinience/alicloud-skills --skill aliyun-wan-digital-human -a codex`. Or copy the skill folder (skills/ai/video/aliyun-wan-digital-human in cinience/alicloud-skills) into .agents/skills/aliyun-wan-digital-human in your project. Codex loads it when a task matches its description.

Can I use Aliyun Wan Digital Human in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add cinience/alicloud-skills --skill aliyun-wan-digital-human -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/aliyun-wan-digital-human, .gemini/skills/aliyun-wan-digital-human, .github/skills/aliyun-wan-digital-human and .opencode/skills/aliyun-wan-digital-human in your project.

What does Aliyun Wan Digital Human need to run?

Going by SKILL.md and its folder, Aliyun Wan Digital Human needs Python for the scripts in its folder, the command-line tools its instructions call (python) and credentials named DASHSCOPE_API_KEY. Our summary lists: Python 3; A credential in DASHSCOPE_API_KEY.

Does Aliyun Wan Digital Human access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Aliyun Wan Digital Human safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Aliyun Wan Digital Human use?

Aliyun Wan Digital Human is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Aliyun Wan Digital Human use?

About 701 tokens (SKILL.md is roughly 2.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 51 tokens, read only when the agent opens those files.

What are the alternatives to Aliyun Wan Digital Human?

Skills that share tags, products or a category with Aliyun Wan Digital Human: Bailian Media Generation (modelstudioai/cli, 542 stars), Dashscope (calesthio/OpenMontage, 66k stars), Image Generation (QwenLM/qwen-code-examples, 143 stars) and Cap Cinematic Demo Generator (CapSoftware/Cap, 23k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Aliyun Wan Digital Human?

cinience (a GitHub user) maintains it in cinience/alicloud-skills, which has 397 GitHub stars. The repository holds 96 skills in this directory. The repository was last updated on August 11, 2026.

Source: cinience/alicloud-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.