Agent skill

Z Qwen Audio Studio

by tjxj in tjxj/z-skills

A skill your agent uses when creating complete generated audio with qwen-audio-3.1-tts-next, including podcasts, radio drama, advertisements, multiple speakers, reference voices, ambience, sound…

No licenceAuto-check passedMedia & Creative

Install Z Qwen Audio Studio

skills CLI
$ npx skills add tjxj/z-skills --skill z-qwen-audio-studio -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install tjxj/z-skills z-qwen-audio-studio --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/tjxj/z-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/z-qwen-audio-studio .claude/skills/z-qwen-audio-studio && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
z-qwen-audio-studio
GitHub stars
548
Token cost
~716 tokens
SKILL.md length
276 words
Files
8 (incl. scripts, references)
Skills in repo
15
Repo updated
First seen
Licence
None found

At a glance

A skill your agent uses when creating complete generated audio with qwen-audio-3.1-tts-next, including podcasts, radio drama, advertisements, multiple speakers, reference voices, ambience, sound…

  • Works in 5 steps: Identify the mode: narration,… → Preserve the user's exact dialogue and… → For reference recordings, bind files in… → …
  • Creating complete generated audio with qwen-audio-3.1-tts-next
  • SKILL.md covers First use, Workflow, Failure handling and Security
  • Runs Python scripts from its folder; calls python3; needs DASHSCOPE_API_KEY

What it does

Z Qwen Audio Studio is an agent skill from tjxj/z-skills. Use when creating complete generated audio with qwen-audio-3.1-tts-next, including podcasts, radio drama, advertisements, multiple speakers, reference voices, ambience, sound effects, or background music. Trigger on 全景音频、双人播客、广播剧、环境声、动作音效、参考音频 and TTS Next requests. Exclude qwen-audio-3.1-tts-flash system-voice TTS.

Its SKILL.md is about 720 tokens, which your agent loads only when the skill is triggered. The skill folder holds 12 other files, including scripts and reference files (for example `agents/openai.yaml`, `evals/evals.json` and `references/api-contract.md`).

It sits in Media & Creative, covering Text to speech and voice, Podcasting and Music and audio generation. It works with Qwen. The repository describes itself as: A collection of reusable skills.

When your agent uses it

  • Creating complete generated audio with qwen-audio-3.1-tts-next
  • Including podcasts
  • Multiple speakers
  • Reference voices

Example prompts

  • “/z-qwen-audio-studio”

Requirements

  • Python 3
  • A credential in DASHSCOPE_API_KEY

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. Identify the mode: narration, advertisement, podcast, drama, or auto.
  2. Preserve the user's exact dialogue and requested sound events. Read references/prompt-patterns.md when the request needs prompt…
  3. For reference recordings, bind files in order to @voice1, @voice2, and @voice3. Each file must be WAV, MP3, or OGG Opus, at most 30…
  4. Compile and inspect the prompt before a paid generation
  5. Generate into an explicit output directory

What it can do on your machine

Read from SKILL.md and the folder at commit a29467e. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • DASHSCOPE_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Z Qwen Audio Studio loads about 716 tokens when it runs, and up to ~2k if it reads all its reference files. Until then it costs about 84 tokens; SKILL.md has 276 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~84
When it runs · the whole SKILL.md, loaded when a task matches
~716
With references · SKILL.md plus every file in references/, read only if the agent opens them
~2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 276 words (~716 tokens).

“Use Alibaba Cloud qwen-audio-3.1-tts-next to generate a complete audio scene from a prompt and up to three reference recordings.”

— opening of SKILL.md by tjxj
name
z-qwen-audio-studio

Read the full SKILL.md on GitHub

Files

SKILL.md and 7 other files (scripts, references) in z-qwen-audio-studio of tjxj/z-skills.

  • SKILL.md
  • agents/openai.yaml
  • evals/evals.json
  • references/api-contract.md
  • references/prompt-patterns.md
  • references/setup.md
  • scripts/qwen_audio_studio.py
  • tests/test_qwen_audio_studio.py

Open the folder on GitHubat commit a29467e

Compare with similar skills

Z Qwen Audio Studio next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Z Qwen Audio Studio compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Z Qwen Audio Studio this skilltjxj/z-skills548—~716Automated safety check: PassNone
AI Podcast CreationNeverSight/learn-skills.dev2171 repos~2kAutomated safety check: PassNone
Audio Processingvivy-yi/xiaohongshu-skills481—~5.4kAutomated safety check: PassNone
VoiceoverGTKottman/mortiflix-oss499—~1.8kAutomated safety check: PassAGPL-3.0
Musictadaspetra/loop2962 repos~827Automated safety check: PassMIT
Sound Effectstadaspetra/loop2962 repos~1.1kAutomated safety check: PassMIT

Similar skills

  • AI Podcast Creation

    NeverSight/learn-skills.dev

    Create AI-powered podcasts with text-to-speech, music, and audio editing.

    217 GitHub starsUsed in 1 repo~2k tokens
    Media & CreativeAuto-check passed
  • Audio Processing

    vivy-yi/xiaohongshu-skills

    A skill your agent uses when processing audio for Xiaohongshu content, editing voiceovers, improving sound quality, creating podcasts, or producing audio-based posts

    481 GitHub stars~5.4k tokensUpdated 8 mo ago
    Media & CreativeAuto-check passed
  • Voiceover

    GTKottman/mortiflix-oss

    Narration, sound effects and music beds in the voice the studio set up (ElevenLabs, Qwen3-TTS on this machine's GPU, or the owner's own voice from the recording booth): lines from the approved…

    499 GitHub stars~1.8k tokensUpdated 2 days ago
    Media & CreativeAuto-check passed
  • Music

    tadaspetra/loop

    Generate music using ElevenLabs Music API. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 2 repos~827 tokens
    Media & CreativeAuto-check passed
  • Sound Effects

    tadaspetra/loop

    Generate sound effects from text descriptions using ElevenLabs.

    296 GitHub starsUsed in 2 repos~1.1k tokens
    Media & CreativeAuto-check passed
  • Podcast

    zarazhangrui/personalized-podcast

    Generate a podcast episode from content you provide. An agent skill from zarazhangrui/personalized-podcast.

    438 GitHub stars~2.3k tokensUpdated 6 mo ago
    Media & CreativeAuto-check: notes

More from tjxj/z-skills

All 15 skills in this repo
  • A skill your agent uses whenever the user wants answers, summaries, comparisons, decisions, quotations, persona simulations, articles, scripts, or claim checks that must stay grounded in…

    548 GitHub stars~1.7k tokensUpdated 19 days ago
    Auto-check passed
  • Z Mail Reader

    tjxj/z-skills

    A skill your agent uses when user wants to read emails, check inbox, fetch emails via IMAP, download email attachments, extract inline images, or summarize email content.

    548 GitHub stars~912 tokensUpdated 19 days ago
    Auto-check passed
  • Z Md Excel

    tjxj/z-skills

    Extract all Markdown tables from a .md file and export to Excel (.xlsx).

    548 GitHub stars~728 tokensUpdated 19 days ago
    Auto-check passed
  • Z Md To Word

    tjxj/z-skills

    Convert local Markdown files into Word documents. An agent skill from tjxj/z-skills.

    548 GitHub stars~815 tokensUpdated 19 days ago
    Auto-check passed
  • 将文章、讲稿或技术主题制作成王虹学术报告气质的 16:9 Notability 手写风 HTML 幻灯片,逐页导出 PNG,并默认同时生成可逐条播放动画的时间轴版 HTML。触发词:王虹PPT风格、王虹手写PPT、Notability学术手写幻灯片、手写网页PPT、手写PPT、数学家手写报告风、手写PPT时间轴版、手写演示动画。

    548 GitHub stars~952 tokensUpdated 19 days ago
    Auto-check passed
  • A skill your agent uses whenever the user asks to talk with, roleplay, interview, quote, fact-check, or apply the reasoning of 梁文锋/Liang Wenfeng from the bundled May 20 investor-meeting transcript.

    548 GitHub stars~1.4k tokensUpdated 19 days ago
    Auto-check passed

Works with

Questions about Z Qwen Audio Studio

What does Z Qwen Audio Studio do?

A skill your agent uses when creating complete generated audio with qwen-audio-3.1-tts-next, including podcasts, radio drama, advertisements, multiple speakers, reference voices, ambience, sound…. Z Qwen Audio Studio is an agent skill from tjxj/z-skills.1-tts-next, including podcasts, radio drama, advertisements, multiple speakers, reference voices, ambience, sound effects, or background music.

When should I use Z Qwen Audio Studio?

Z Qwen Audio Studio fits situations like: creating complete generated audio with qwen-audio-3.1-tts-next; including podcasts; multiple speakers; reference voices.

How do I install Z Qwen Audio Studio in Claude Code?

Run `npx skills add tjxj/z-skills --skill z-qwen-audio-studio -a claude-code`. Or copy the skill folder (z-qwen-audio-studio in tjxj/z-skills) into .claude/skills/z-qwen-audio-studio in your project. Claude Code loads it when a task matches its description.

How do I install Z Qwen Audio Studio in Codex?

Run `npx skills add tjxj/z-skills --skill z-qwen-audio-studio -a codex`. Or copy the skill folder (z-qwen-audio-studio in tjxj/z-skills) into .agents/skills/z-qwen-audio-studio in your project. Codex loads it when a task matches its description.

Can I use Z Qwen Audio Studio in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add tjxj/z-skills --skill z-qwen-audio-studio -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/z-qwen-audio-studio, .gemini/skills/z-qwen-audio-studio, .github/skills/z-qwen-audio-studio and .opencode/skills/z-qwen-audio-studio in your project.

What does Z Qwen Audio Studio need to run?

Going by SKILL.md and its folder, Z Qwen Audio Studio needs Python for the scripts in its folder, the command-line tools its instructions call (python3) and credentials named DASHSCOPE_API_KEY. Our summary lists: Python 3; A credential in DASHSCOPE_API_KEY.

Does Z Qwen Audio Studio access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Z Qwen Audio Studio safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Z Qwen Audio Studio use?

No licence was found for Z Qwen Audio Studio or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Z Qwen Audio Studio use?

About 716 tokens (SKILL.md is roughly 2.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 1.3k tokens, read only when the agent opens those files.

What are the alternatives to Z Qwen Audio Studio?

Skills that share tags, products or a category with Z Qwen Audio Studio: AI Podcast Creation (NeverSight/learn-skills.dev, 217 stars), Audio Processing (vivy-yi/xiaohongshu-skills, 481 stars), Voiceover (GTKottman/mortiflix-oss, 499 stars) and Music (tadaspetra/loop, 296 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Z Qwen Audio Studio?

tjxj (a GitHub user) maintains it in tjxj/z-skills, which has 548 GitHub stars. The repository holds 15 skills in this directory. The repository was last updated on September 22, 2026.

Source: tjxj/z-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.