Agent skill

Aliyun Qwen Tts

by cinience in cinience/alicloud-skills

A skill your agent uses when generating human-like speech audio with Model Studio DashScope Qwen TTS models (qwen3-tts-flash, qwen3-tts-instruct-flash).

MITAuto-check passedMedia & Creative

Install Aliyun Qwen Tts

skills CLI
$ npx skills add cinience/alicloud-skills --skill aliyun-qwen-tts -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install cinience/alicloud-skills aliyun-qwen-tts --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/cinience/alicloud-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/ai/audio/aliyun-qwen-tts .claude/skills/aliyun-qwen-tts && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
aliyun-qwen-tts
GitHub stars
397
Token cost
~951 tokens
SKILL.md length
279 words
Files
5 (incl. scripts, references)
Skills in repo
96
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses when generating human-like speech audio with Model Studio DashScope Qwen TTS models (qwen3-tts-flash, qwen3-tts-instruct-flash).

  • Works in 4 steps: Confirm user intent, region,… → Run one minimal read-only query first to… → Execute the target operation with… → …
  • Generating human-like speech audio with Model Studio DashScope Qwen TTS models (qwen3-tts-flash
  • SKILL.md covers Validation, Output And Evidence, Critical model names and Prerequisites, plus 7 more sections
  • Runs Python scripts from its folder; calls python and python3; reaches dashscope-intl.aliyuncs.com and dashscope.aliyuncs.com; needs DASHSCOPE_API_KEY

What it does

Aliyun Qwen Tts is an agent skill from cinience/alicloud-skills. Use when generating human-like speech audio with Model Studio DashScope Qwen TTS models (qwen3-tts-flash, qwen3-tts-instruct-flash). Use when converting text to speech, producing voice lines for short drama/news videos, or documenting TTS request/response fields for DashScope.

Its SKILL.md is about 950 tokens, which your agent loads only when the skill is triggered. The skill folder holds 7 other files, including scripts and reference files (for example `agents/openai.yaml`, `references/api_reference.md` and `references/sources.md`).

It sits in Media & Creative, covering Text to speech and voice. It works with Qwen and Alibaba Cloud. The repository describes itself as: alibaba cloud skills,qwen ,wan and all skills. The licence is MIT.

When your agent uses it

  • Generating human-like speech audio with Model Studio DashScope Qwen TTS models (qwen3-tts-flash
  • Qwen3-tts-instruct-flash)
  • Converting text to speech
  • Producing voice lines for short drama/news videos

Example prompts

  • “/aliyun-qwen-tts”

Requirements

  • Python 3
  • A credential in DASHSCOPE_API_KEY

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Confirm user intent, region, identifiers, and whether the operation is read-only or mutating.
  2. Run one minimal read-only query first to verify connectivity and permissions.
  3. Execute the target operation with explicit parameters and bounded scope.
  4. Verify results and save output/evidence files.

What it can do on your machine

Read from SKILL.md and the folder at commit 1818263. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python
    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • dashscope-intl.aliyuncs.com
    • dashscope.aliyuncs.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • DASHSCOPE_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Aliyun Qwen Tts loads about 951 tokens when it runs, and up to ~1.4k if it reads all its reference files. Until then it costs about 73 tokens; SKILL.md has 279 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~73
When it runs · the whole SKILL.md, loaded when a task matches
~951
With references · SKILL.md plus every file in references/, read only if the agent opens them
~1.4k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from cinience/alicloud-skills at commit 1818263, republished under its MIT licence (© cinience). 279 words, ~951 tokens.

Download SKILL.mdSave it as .claude/skills/aliyun-qwen-tts/SKILL.md (or your agent's skills folder). This skill also uses 4 other files; get the full folder from GitHub.
name
aliyun-qwen-tts
description
Use when generating human-like speech audio with Model Studio DashScope Qwen TTS models (qwen3-tts-flash, qwen3-tts-instruct-flash). Use when converting text to speech, producing voice lines for short drama/news videos, or documenting TTS request/response fields for DashScope.
version
1.0.0

Category: provider

Model Studio Qwen TTS

Validation

bash
mkdir -p output/aliyun-qwen-tts
python -m py_compile skills/ai/audio/aliyun-qwen-tts/scripts/generate_tts.py && echo "py_compile_ok" > output/aliyun-qwen-tts/validate.txt

Pass criteria: command exits 0 and output/aliyun-qwen-tts/validate.txt is generated.

Output And Evidence

  • Save generated audio links, sample audio files, and request payloads to output/aliyun-qwen-tts/.
  • Keep one validation log per execution.

Critical model names

Use one of the recommended models:

  • qwen3-tts-flash
  • qwen3-tts-instruct-flash
  • qwen3-tts-instruct-flash-2026-01-26

Prerequisites

  • Install SDK (recommended in a venv to avoid PEP 668 limits):
bash
python3 -m venv .venv
. .venv/bin/activate
python -m pip install dashscope
  • Set DASHSCOPE_API_KEY in your environment, or add dashscope_api_key to ~/.alibabacloud/credentials (env takes precedence).

Normalized interface (tts.generate)

Request
  • text (string, required)
  • voice (string, required)
  • language_type (string, optional; default Auto)
  • instruction (string, optional; recommended for instruct models)
  • stream (bool, optional; default false)
Response
  • audio_url (string, when stream=false)
  • audio_base64_pcm (string, when stream=true)
  • sample_rate (int, 24000)
  • format (string, wav or pcm depending on mode)

Quick start (Python + DashScope SDK)

python
import os
import dashscope

# Prefer env var for auth: export DASHSCOPE_API_KEY=...
# Or use ~/.alibabacloud/credentials with dashscope_api_key under [default].
# Beijing region; for Singapore use: https://dashscope-intl.aliyuncs.com/api/v1
dashscope.base_http_api_url = "https://dashscope.aliyuncs.com/api/v1"

text = "Hello, this is a short voice line."
response = dashscope.MultiModalConversation.call(
    model="qwen3-tts-instruct-flash",
    api_key=os.getenv("DASHSCOPE_API_KEY"),
    text=text,
    voice="Cherry",
    language_type="English",
    instruction="Warm and calm tone, slightly slower pace.",
    stream=False,
)

audio_url = response.output.audio.url
print(audio_url)

Streaming notes

  • stream=True returns Base64-encoded PCM chunks at 24kHz.
  • Decode chunks and play or concatenate to a pcm buffer.
  • The response contains finish_reason == "stop" when the stream ends.

Operational guidance

  • Keep requests concise; split long text into multiple calls if you hit size or timeout errors.
  • Use language_type consistent with the text to improve pronunciation.
  • Use instruction only when you need explicit style/tone control.
  • Cache by (text, voice, language_type) to avoid repeat costs.

Output location

  • Default output: output/aliyun-qwen-tts/audio/
  • Override base dir with OUTPUT_DIR.

Workflow

  1. Confirm user intent, region, identifiers, and whether the operation is read-only or mutating.
  2. Run one minimal read-only query first to verify connectivity and permissions.
  3. Execute the target operation with explicit parameters and bounded scope.
  4. Verify results and save output/evidence files.

References

  • references/api_reference.md for parameter mapping and streaming example.

  • Realtime mode is provided by skills/ai/audio/aliyun-qwen-tts-realtime/.

  • Voice cloning/design are provided by skills/ai/audio/aliyun-qwen-tts-voice-clone/ and skills/ai/audio/aliyun-qwen-tts-voice-design/.

  • Source list: references/sources.md

© cinience, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 4 other files (scripts, references) in skills/ai/audio/aliyun-qwen-tts of cinience/alicloud-skills.

  • SKILL.md
  • agents/openai.yaml
  • references/api_reference.md
  • references/sources.md
  • scripts/generate_tts.py

Open the folder on GitHubat commit 1818263

Compare with similar skills

Aliyun Qwen Tts next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Aliyun Qwen Tts compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Aliyun Qwen Tts this skillcinience/alicloud-skills397—~951Automated safety check: PassMIT
Bailian Media Generationmodelstudioai/cli542—~2kAutomated safety check: PassApache-2.0
Dashscopecalesthio/OpenMontage66k—~1.5kAutomated safety check: NotesAGPL-3.0
Audio Ttssecond-state/qwen3_tts_rs233—~1.6kAutomated safety check: PassNone
Qianwen Audio TtsQianWen-AI/qianwen-ai105—~4.2kAutomated safety check: NotesApache-2.0
Z Qwen Audio Studiotjxj/z-skills548—~716Automated safety check: PassNone

Similar skills

  • Bailian Media Generation

    modelstudioai/cli

    Chinese-language entry point into Alibaba Cloud Bailian's image, video and speech generation and understanding, routed through separate image, video, speech and vision commands.

    542 GitHub stars~2k tokensUpdated 2 days ago
    Media & CreativeAuto-check passed
  • Dashscope

    calesthio/OpenMontage

    DashScope (Alibaba Cloud Bailian / 阿里云百炼) integration — image generation (qwen-image-2.0-pro), text-to-speech (qwen3-tts-flash), and ASR with word-level timestamps (qwen3-asr-flash-filetrans).

    66k GitHub stars~1.5k tokensUpdated 8 days ago
    Media & CreativeAuto-check: notes
  • Audio Tts

    second-state/qwen3_tts_rs

    Generate speech audio from text using Qwen3 TTS, or clone a voice from reference audio.

    233 GitHub stars~1.6k tokensUpdated 4 mo ago
    Media & CreativeAuto-check passed
  • Qianwen Audio Tts

    QianWen-AI/qianwen-ai

    Synthesize speech from text with Qwen TTS models. An agent skill from QianWen-AI/qianwen-ai.

    105 GitHub stars~4.2k tokensUpdated yesterday
    Media & CreativeAuto-check: notes
  • Z Qwen Audio Studio

    tjxj/z-skills

    A skill your agent uses when creating complete generated audio with qwen-audio-3.1-tts-next, including podcasts, radio drama, advertisements, multiple speakers, reference voices, ambience, sound…

    548 GitHub stars~716 tokensUpdated 19 days ago
    Media & CreativeAuto-check passed
  • Voiceover

    GTKottman/mortiflix-oss

    Narration, sound effects and music beds in the voice the studio set up (ElevenLabs, Qwen3-TTS on this machine's GPU, or the owner's own voice from the recording booth): lines from the approved…

    499 GitHub stars~1.8k tokensUpdated 2 days ago
    Media & CreativeAuto-check passed

More from cinience/alicloud-skills

All 96 skills in this repo
  • Aliyun Skill Creator

    cinience/alicloud-skills

    A skill your agent uses when creating, migrating, or optimizing skills for this alicloud-skills repository.

    397 GitHub stars~2.8k tokensUpdated 2 mo ago
    Auto-check passed
  • Alicloud Acs Agent Sandbox

    cinience/alicloud-skills

    Bootstrap, create, connect to, operate, secure, scale, upgrade, troubleshoot, inspect, and tear down Alibaba Cloud Container Compute Service (ACS) Agent Sandbox environments.

    397 GitHub stars~2.7k tokensUpdated 2 mo ago
    Auto-check passed
  • Alicloud Acs Cluster

    cinience/alicloud-skills

    Create, inspect, connect to, inventory, and delete Alibaba Cloud Container Compute Service (ACS) clusters through the official CS OpenAPI.

    397 GitHub stars~1.7k tokensUpdated 2 mo ago
    Auto-check passed
  • Aliyun Adb Mysql

    cinience/alicloud-skills

    A skill your agent uses when managing Alibaba Cloud AnalyticDB for MySQL (ADB) via OpenAPI/SDK, including the user needs AnalyticDB resource lifecycle and configuration operations, status checks, or…

    397 GitHub stars~708 tokensUpdated 2 mo ago
    Auto-check passed
  • Aliyun Aicontent Generate

    cinience/alicloud-skills

    A skill your agent uses when managing Alibaba Cloud AIContent (AiContent) via OpenAPI/SDK, including the user needs AI content generation or content workflow operations in Alibaba Cloud, including…

    397 GitHub stars~734 tokensUpdated 2 mo ago
    Auto-check passed
  • Aliyun Aimiaobi Generate

    cinience/alicloud-skills

    A skill your agent uses when managing Alibaba Cloud Quan Miao (AiMiaoBi) via OpenAPI/SDK, including the user asks for Alibaba Cloud MiaoBi content operations, including listing resources…

    397 GitHub stars~724 tokensUpdated 2 mo ago
    Auto-check passed

Questions about Aliyun Qwen Tts

What does Aliyun Qwen Tts do?

A skill your agent uses when generating human-like speech audio with Model Studio DashScope Qwen TTS models (qwen3-tts-flash, qwen3-tts-instruct-flash). Aliyun Qwen Tts is an agent skill from cinience/alicloud-skills. Use when generating human-like speech audio with Model Studio DashScope Qwen TTS models (qwen3-tts-flash, qwen3-tts-instruct-flash).

When should I use Aliyun Qwen Tts?

Aliyun Qwen Tts fits situations like: generating human-like speech audio with Model Studio DashScope Qwen TTS models (qwen3-tts-flash; qwen3-tts-instruct-flash); converting text to speech; producing voice lines for short drama/news videos.

How do I install Aliyun Qwen Tts in Claude Code?

Run `npx skills add cinience/alicloud-skills --skill aliyun-qwen-tts -a claude-code`. Or copy the skill folder (skills/ai/audio/aliyun-qwen-tts in cinience/alicloud-skills) into .claude/skills/aliyun-qwen-tts in your project. Claude Code loads it when a task matches its description.

How do I install Aliyun Qwen Tts in Codex?

Run `npx skills add cinience/alicloud-skills --skill aliyun-qwen-tts -a codex`. Or copy the skill folder (skills/ai/audio/aliyun-qwen-tts in cinience/alicloud-skills) into .agents/skills/aliyun-qwen-tts in your project. Codex loads it when a task matches its description.

Can I use Aliyun Qwen Tts in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add cinience/alicloud-skills --skill aliyun-qwen-tts -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/aliyun-qwen-tts, .gemini/skills/aliyun-qwen-tts, .github/skills/aliyun-qwen-tts and .opencode/skills/aliyun-qwen-tts in your project.

What does Aliyun Qwen Tts need to run?

Going by SKILL.md and its folder, Aliyun Qwen Tts needs Python for the scripts in its folder, the command-line tools its instructions call (python and python3) and credentials named DASHSCOPE_API_KEY. Our summary lists: Python 3; A credential in DASHSCOPE_API_KEY.

Does Aliyun Qwen Tts access the network?

SKILL.md names 2 domains. In commands or code: dashscope-intl.aliyuncs.com and dashscope.aliyuncs.com; the agent is likely to contact these when it follows the instructions. This is read from the text; nothing was executed.

Is Aliyun Qwen Tts safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Aliyun Qwen Tts use?

Aliyun Qwen Tts is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Aliyun Qwen Tts use?

About 951 tokens (SKILL.md is roughly 3.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 451 tokens, read only when the agent opens those files.

What are the alternatives to Aliyun Qwen Tts?

Skills that share tags, products or a category with Aliyun Qwen Tts: Bailian Media Generation (modelstudioai/cli, 542 stars), Dashscope (calesthio/OpenMontage, 66k stars), Audio Tts (second-state/qwen3_tts_rs, 233 stars) and Qianwen Audio Tts (QianWen-AI/qianwen-ai, 105 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Aliyun Qwen Tts?

cinience (a GitHub user) maintains it in cinience/alicloud-skills, which has 397 GitHub stars. The repository holds 96 skills in this directory. The repository was last updated on August 11, 2026.

Source: cinience/alicloud-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.