Agent skill

Aliyun Qwen Tts Voice Clone

by cinience in cinience/alicloud-skills

A skill your agent uses when cloning voices with Alibaba Cloud Model Studio Qwen TTS VC models.

MITAuto-check passedMedia & Creative

Install Aliyun Qwen Tts Voice Clone

skills CLI
$ npx skills add cinience/alicloud-skills --skill aliyun-qwen-tts-voice-clone -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install cinience/alicloud-skills aliyun-qwen-tts-voice-clone --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/cinience/alicloud-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/ai/audio/aliyun-qwen-tts-voice-clone .claude/skills/aliyun-qwen-tts-voice-clone && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
aliyun-qwen-tts-voice-clone
GitHub stars
397
Token cost
~661 tokens
SKILL.md length
208 words
Files
4 (incl. scripts, references)
Skills in repo
96
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses when cloning voices with Alibaba Cloud Model Studio Qwen TTS VC models.

  • Works in 4 steps: Confirm user intent, region,… → Run one minimal read-only query first to… → Execute the target operation with… → …
  • Cloning voices with Alibaba Cloud Model Studio Qwen TTS VC models
  • SKILL.md covers Critical model names, Prerequisites, Normalized interface… and Operational guidance, plus 6 more sections
  • Runs Python scripts from its folder; calls python3 and python; needs DASHSCOPE_API_KEY

What it does

Aliyun Qwen Tts Voice Clone is an agent skill from cinience/alicloud-skills. Use when cloning voices with Alibaba Cloud Model Studio Qwen TTS VC models. Use when creating cloned voices from sample audio and synthesizing text with cloned timbre.

Its SKILL.md is about 660 tokens, which your agent loads only when the skill is triggered. The skill folder holds 6 other files, including scripts and reference files (for example `agents/openai.yaml`, `references/sources.md` and `scripts/prepare_voice_clone_request.py`).

It sits in Media & Creative, covering Text to speech and voice. It works with Qwen and Alibaba Cloud. The repository describes itself as: alibaba cloud skills,qwen ,wan and all skills. The licence is MIT.

When your agent uses it

  • Cloning voices with Alibaba Cloud Model Studio Qwen TTS VC models
  • Creating cloned voices from sample audio and synthesizing text with cloned timbre

Example prompts

  • “/aliyun-qwen-tts-voice-clone”

Requirements

  • Python 3
  • A credential in DASHSCOPE_API_KEY

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Confirm user intent, region, identifiers, and whether the operation is read-only or mutating.
  2. Run one minimal read-only query first to verify connectivity and permissions.
  3. Execute the target operation with explicit parameters and bounded scope.
  4. Verify results and save output/evidence files.

What it can do on your machine

Read from SKILL.md and the folder at commit 1818263. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python3
    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • DASHSCOPE_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Aliyun Qwen Tts Voice Clone loads about 661 tokens when it runs, and up to ~694 if it reads all its reference files. Until then it costs about 49 tokens; SKILL.md has 208 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~49
When it runs · the whole SKILL.md, loaded when a task matches
~661
With references · SKILL.md plus every file in references/, read only if the agent opens them
~694

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from cinience/alicloud-skills at commit 1818263, republished under its MIT licence (© cinience). 208 words, ~661 tokens.

Download SKILL.mdSave it as .claude/skills/aliyun-qwen-tts-voice-clone/SKILL.md (or your agent's skills folder). This skill also uses 3 other files; get the full folder from GitHub.
name
aliyun-qwen-tts-voice-clone
description
Use when cloning voices with Alibaba Cloud Model Studio Qwen TTS VC models. Use when creating cloned voices from sample audio and synthesizing text with cloned timbre.
version
1.0.0

Category: provider

Model Studio Qwen TTS Voice Clone

Use voice cloning models to replicate timbre from enrollment audio samples.

Critical model names

Use one of these exact model strings:

  • qwen3-tts-vc-2026-01-22
  • qwen3-tts-vc-realtime-2026-01-15

Prerequisites

  • Install SDK in a virtual environment:
bash
python3 -m venv .venv
. .venv/bin/activate
python -m pip install dashscope
  • Set DASHSCOPE_API_KEY in your environment, or add dashscope_api_key to ~/.alibabacloud/credentials.

Normalized interface (tts.voice_clone)

Request
  • text (string, required)
  • voice_sample (string | bytes, required) enrollment sample
  • voice_name (string, optional)
  • stream (bool, optional)
Response
  • audio_url (string) or streaming PCM chunks
  • voice_id (string)
  • request_id (string)

Operational guidance

  • Use clean speech samples with low background noise.
  • Respect consent and policy requirements for cloned voices.
  • Persist generated voice_id and reuse for future synthesis requests.

Local helper script

Prepare a normalized request JSON and validate response schema:

bash
.venv/bin/python skills/ai/audio/aliyun-qwen-tts-voice-clone/scripts/prepare_voice_clone_request.py \
  --text "Welcome to this voice-clone demo" \
  --voice-sample "https://example.com/voice-sample.wav"

Output location

  • Default output: output/ai-audio-tts-voice-clone/audio/
  • Override base dir with OUTPUT_DIR.

Validation

bash
mkdir -p output/aliyun-qwen-tts-voice-clone
for f in skills/ai/audio/aliyun-qwen-tts-voice-clone/scripts/*.py; do
  python3 -m py_compile "$f"
done
echo "py_compile_ok" > output/aliyun-qwen-tts-voice-clone/validate.txt

Pass criteria: command exits 0 and output/aliyun-qwen-tts-voice-clone/validate.txt is generated.

Output And Evidence

  • Save artifacts, command outputs, and API response summaries under output/aliyun-qwen-tts-voice-clone/.
  • Include key parameters (region/resource id/time range) in evidence files for reproducibility.

Workflow

  1. Confirm user intent, region, identifiers, and whether the operation is read-only or mutating.
  2. Run one minimal read-only query first to verify connectivity and permissions.
  3. Execute the target operation with explicit parameters and bounded scope.
  4. Verify results and save output/evidence files.

References

  • references/sources.md

© cinience, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 3 other files (scripts, references) in skills/ai/audio/aliyun-qwen-tts-voice-clone of cinience/alicloud-skills.

  • SKILL.md
  • agents/openai.yaml
  • references/sources.md
  • scripts/prepare_voice_clone_request.py

Open the folder on GitHubat commit 1818263

Compare with similar skills

Aliyun Qwen Tts Voice Clone next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Aliyun Qwen Tts Voice Clone compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Aliyun Qwen Tts Voice Clone this skillcinience/alicloud-skills397—~661Automated safety check: PassMIT
Bailian Media Generationmodelstudioai/cli542—~2kAutomated safety check: PassApache-2.0
Dashscopecalesthio/OpenMontage66k—~1.5kAutomated safety check: NotesAGPL-3.0
Audio Ttssecond-state/qwen3_tts_rs233—~1.6kAutomated safety check: PassNone
Qianwen Audio TtsQianWen-AI/qianwen-ai105—~4.2kAutomated safety check: NotesApache-2.0
Z Qwen Audio Studiotjxj/z-skills548—~716Automated safety check: PassNone

Similar skills

  • Bailian Media Generation

    modelstudioai/cli

    Chinese-language entry point into Alibaba Cloud Bailian's image, video and speech generation and understanding, routed through separate image, video, speech and vision commands.

    542 GitHub stars~2k tokensUpdated 2 days ago
    Media & CreativeAuto-check passed
  • Dashscope

    calesthio/OpenMontage

    DashScope (Alibaba Cloud Bailian / 阿里云百炼) integration — image generation (qwen-image-2.0-pro), text-to-speech (qwen3-tts-flash), and ASR with word-level timestamps (qwen3-asr-flash-filetrans).

    66k GitHub stars~1.5k tokensUpdated 8 days ago
    Media & CreativeAuto-check: notes
  • Audio Tts

    second-state/qwen3_tts_rs

    Generate speech audio from text using Qwen3 TTS, or clone a voice from reference audio.

    233 GitHub stars~1.6k tokensUpdated 4 mo ago
    Media & CreativeAuto-check passed
  • Qianwen Audio Tts

    QianWen-AI/qianwen-ai

    Synthesize speech from text with Qwen TTS models. An agent skill from QianWen-AI/qianwen-ai.

    105 GitHub stars~4.2k tokensUpdated yesterday
    Media & CreativeAuto-check: notes
  • Z Qwen Audio Studio

    tjxj/z-skills

    A skill your agent uses when creating complete generated audio with qwen-audio-3.1-tts-next, including podcasts, radio drama, advertisements, multiple speakers, reference voices, ambience, sound…

    548 GitHub stars~716 tokensUpdated 19 days ago
    Media & CreativeAuto-check passed
  • Voiceover

    GTKottman/mortiflix-oss

    Narration, sound effects and music beds in the voice the studio set up (ElevenLabs, Qwen3-TTS on this machine's GPU, or the owner's own voice from the recording booth): lines from the approved…

    499 GitHub stars~1.8k tokensUpdated 2 days ago
    Media & CreativeAuto-check passed

More from cinience/alicloud-skills

All 96 skills in this repo
  • Aliyun Skill Creator

    cinience/alicloud-skills

    A skill your agent uses when creating, migrating, or optimizing skills for this alicloud-skills repository.

    397 GitHub stars~2.8k tokensUpdated 2 mo ago
    Auto-check passed
  • Alicloud Acs Agent Sandbox

    cinience/alicloud-skills

    Bootstrap, create, connect to, operate, secure, scale, upgrade, troubleshoot, inspect, and tear down Alibaba Cloud Container Compute Service (ACS) Agent Sandbox environments.

    397 GitHub stars~2.7k tokensUpdated 2 mo ago
    Auto-check passed
  • Alicloud Acs Cluster

    cinience/alicloud-skills

    Create, inspect, connect to, inventory, and delete Alibaba Cloud Container Compute Service (ACS) clusters through the official CS OpenAPI.

    397 GitHub stars~1.7k tokensUpdated 2 mo ago
    Auto-check passed
  • Aliyun Adb Mysql

    cinience/alicloud-skills

    A skill your agent uses when managing Alibaba Cloud AnalyticDB for MySQL (ADB) via OpenAPI/SDK, including the user needs AnalyticDB resource lifecycle and configuration operations, status checks, or…

    397 GitHub stars~708 tokensUpdated 2 mo ago
    Auto-check passed
  • Aliyun Aicontent Generate

    cinience/alicloud-skills

    A skill your agent uses when managing Alibaba Cloud AIContent (AiContent) via OpenAPI/SDK, including the user needs AI content generation or content workflow operations in Alibaba Cloud, including…

    397 GitHub stars~734 tokensUpdated 2 mo ago
    Auto-check passed
  • Aliyun Aimiaobi Generate

    cinience/alicloud-skills

    A skill your agent uses when managing Alibaba Cloud Quan Miao (AiMiaoBi) via OpenAPI/SDK, including the user asks for Alibaba Cloud MiaoBi content operations, including listing resources…

    397 GitHub stars~724 tokensUpdated 2 mo ago
    Auto-check passed

Questions about Aliyun Qwen Tts Voice Clone

What does Aliyun Qwen Tts Voice Clone do?

A skill your agent uses when cloning voices with Alibaba Cloud Model Studio Qwen TTS VC models. Aliyun Qwen Tts Voice Clone is an agent skill from cinience/alicloud-skills. Use when cloning voices with Alibaba Cloud Model Studio Qwen TTS VC models.

When should I use Aliyun Qwen Tts Voice Clone?

Aliyun Qwen Tts Voice Clone fits situations like: cloning voices with Alibaba Cloud Model Studio Qwen TTS VC models; creating cloned voices from sample audio and synthesizing text with cloned timbre.

How do I install Aliyun Qwen Tts Voice Clone in Claude Code?

Run `npx skills add cinience/alicloud-skills --skill aliyun-qwen-tts-voice-clone -a claude-code`. Or copy the skill folder (skills/ai/audio/aliyun-qwen-tts-voice-clone in cinience/alicloud-skills) into .claude/skills/aliyun-qwen-tts-voice-clone in your project. Claude Code loads it when a task matches its description.

How do I install Aliyun Qwen Tts Voice Clone in Codex?

Run `npx skills add cinience/alicloud-skills --skill aliyun-qwen-tts-voice-clone -a codex`. Or copy the skill folder (skills/ai/audio/aliyun-qwen-tts-voice-clone in cinience/alicloud-skills) into .agents/skills/aliyun-qwen-tts-voice-clone in your project. Codex loads it when a task matches its description.

Can I use Aliyun Qwen Tts Voice Clone in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add cinience/alicloud-skills --skill aliyun-qwen-tts-voice-clone -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/aliyun-qwen-tts-voice-clone, .gemini/skills/aliyun-qwen-tts-voice-clone, .github/skills/aliyun-qwen-tts-voice-clone and .opencode/skills/aliyun-qwen-tts-voice-clone in your project.

What does Aliyun Qwen Tts Voice Clone need to run?

Going by SKILL.md and its folder, Aliyun Qwen Tts Voice Clone needs Python for the scripts in its folder, the command-line tools its instructions call (python3 and python) and credentials named DASHSCOPE_API_KEY. Our summary lists: Python 3; A credential in DASHSCOPE_API_KEY.

Does Aliyun Qwen Tts Voice Clone access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Aliyun Qwen Tts Voice Clone safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Aliyun Qwen Tts Voice Clone use?

Aliyun Qwen Tts Voice Clone is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Aliyun Qwen Tts Voice Clone use?

About 661 tokens (SKILL.md is roughly 2.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 33 tokens, read only when the agent opens those files.

What are the alternatives to Aliyun Qwen Tts Voice Clone?

Skills that share tags, products or a category with Aliyun Qwen Tts Voice Clone: Bailian Media Generation (modelstudioai/cli, 542 stars), Dashscope (calesthio/OpenMontage, 66k stars), Audio Tts (second-state/qwen3_tts_rs, 233 stars) and Qianwen Audio Tts (QianWen-AI/qianwen-ai, 105 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Aliyun Qwen Tts Voice Clone?

cinience (a GitHub user) maintains it in cinience/alicloud-skills, which has 397 GitHub stars. The repository holds 96 skills in this directory. The repository was last updated on August 11, 2026.

Source: cinience/alicloud-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.