Agent skill

Audio Studio

by WrongStack in WrongStack/WrongStack

Compose music, lyrics, soundscapes and voiceovers, or implement browser audio feedback.

MITAuto-check passedMedia & Creative

Install Audio Studio

skills CLI
$ npx skills add WrongStack/WrongStack --skill audio-studio -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install WrongStack/WrongStack audio-studio --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/WrongStack/WrongStack.git skills-src && mkdir -p .claude/skills && cp -r skills-src/packages/core/skills/audio-studio .claude/skills/audio-studio && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
audio-studio
GitHub stars
371
Token cost
~910 tokens
SKILL.md length
399 words
Files
1
Skills in repo
100
Repo updated
First seen
Licence
MIT

At a glance

Compose music, lyrics, soundscapes and voiceovers, or implement browser audio feedback.

  • Works in 6 steps: Define duration, language,… → Separate style directions from… → For narration, mark pronunciation,… → …
  • Other music prompts
  • SKILL.md covers Selection card, Overview, Rules and Workflow, plus 3 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Audio Studio is an agent skill from WrongStack/WrongStack. Compose music, lyrics, soundscapes and voiceovers, or implement browser audio feedback. Use when creating Suno or other music prompts, TTS narration, a soundtrack or interactive sound; use media-production for the final video timeline and mux.

Its SKILL.md is about 910 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering Text to speech and voice. It works with Suno. The repository describes itself as: An AI coding agent that reads your code, edits files, runs commands, and reasons through bugs — across a terminal REPL, a full-screen TUI, and a browser UI, while you keep your… The licence is MIT.

When your agent uses it

  • Other music prompts
  • Interactive sound
  • Use media-production for the final video timeline and mux

Example prompts

  • “/audio-studio”

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. Define duration, language, instrumental/vocal intent, pacing, instrumentation,
  2. Separate style directions from sung/spoken text. Section tags may guide a music
  3. For narration, mark pronunciation, pauses and emotional changes only in syntax
  4. Use authorized voices, lyrics and licensed assets. Check commercial-use terms
  5. Distinguish a prompt from a generated audio file. Deliver the actual file only
  6. For browser audio, create/resume AudioContext after a user gesture, provide mute

What it can do on your machine

Read from SKILL.md and the folder at commit a744bdc. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • help.suno.com
    • elevenlabs.io

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Audio Studio loads about 910 tokens when it runs. Until then it costs about 64 tokens; SKILL.md has 399 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~64
When it runs · the whole SKILL.md, loaded when a task matches
~910

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from WrongStack/WrongStack at commit a744bdc, republished under its MIT licence (© WrongStack). 399 words, ~910 tokens.

Download SKILL.mdSave it as .claude/skills/audio-studio/SKILL.md (or your agent's skills folder).
name
audio-studio
description
Compose music, lyrics, soundscapes and voiceovers, or implement browser audio feedback. Use when creating Suno or other music prompts, TTS narration, a soundtrack or interactive sound; use media-production for the final video timeline and mux.
version
1.2.1
required-capabilities
filesystem.read, filesystem.write
optional-capabilities
verification.run, web.research
trigger
creating Suno or other music prompts, TTS narration, a soundtrack or interactive sound; use media-production for the final video timeline and mux.
metadata.routing-group
media

Audio Studio

Selection card

  • Task: Prepare music, narration and audio delivery. / TR: Müzik, seslendirme ve ses teslimi hazırla.
  • Start: Identify the existing engine, scene, timeline and delivery format.
  • Finish: apply the acceptance checks below; report observed results and unresolved constraints.

Overview

Suno v6 is the current flagship model (v6-mini is the accessible fast variant), verified against its official model guide on 2026-10-09. ElevenLabs offers Eleven v3 alongside models optimized for latency. Choose the latest available model appropriate to the task; record its exact id, access tier and output limits. Do not describe an SDK version as a model version.

Rules

  1. Define duration, language, instrumental/vocal intent, pacing, instrumentation, reference mood and delivery format before generation. Use reasonable defaults when the brief is sufficient.
  2. Separate style directions from sung/spoken text. Section tags may guide a music model; they do not guarantee BPM, exact timestamps or duration.
  3. For narration, mark pronunciation, pauses and emotional changes only in syntax the chosen provider supports. Do not force a gender, accent or imitation.
  4. Use authorized voices, lyrics and licensed assets. Check commercial-use terms for the user's access tier. Keep API keys outside source and output logs.
  5. Distinguish a prompt from a generated audio file. Deliver the actual file only after listening to it and validating its duration, clipping and channels.
  6. For browser audio, create/resume AudioContext after a user gesture, provide mute and volume control, and release oscillators and audio nodes when finished.
Show full SKILL.md (157 more words)Show less

Workflow

  1. Inspect existing audio code and provider/tool availability. Verify current model documentation before a paid generation or SDK upgrade.
  2. Draft the musical structure or narration; synchronize it to the storyboard when the task includes video.
  3. Generate a short sample first when voice, pronunciation or cost is uncertain. Bound retries and preserve accepted takes.
  4. Listen for missing words, unintended sung directions, harsh cuts and distortion. Measure loudness/true peak against the destination specification, rather than applying a universal loudness number.
  5. Retain a lossless master when available; export a playback-compatible preview. Record voice/model, sample rate, channels, duration and edits.

Sources

Suno models, ElevenLabs models. SDK snapshot: ElevenLabs client 1.27.0 and fal client 1.10.1, checked 2026-10-09; refresh npm before installation.

Acceptance checks

  • Listen to the delivered audio; verify duration, clipping, channels and the intended voice/music balance.

Skills in scope

  • media-production — audio placement, subtitles and delivery.
  • motion-canvas-video — synchronize scene cues to narration.
  • manim-video — narration pacing for explanations.

© WrongStack, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in packages/core/skills/audio-studio of WrongStack/WrongStack.

Open the folder on GitHubat commit a744bdc

Compare with similar skills

Audio Studio next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Audio Studio compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Audio Studio this skillWrongStack/WrongStack371—~910Automated safety check: PassMIT
AI MusicZJU-REAL/Easel3.4k—~896Automated safety check: NotesApache-2.0
AI Music And Soundsocial-media-skills/skills134—~1.9kAutomated safety check: PassMIT
Audio Jinglesanqiufong/slides-from-anything1321 repos~1.1kAutomated safety check: PassApache-2.0
Sunosocial-media-skills/skills134—~2.1kAutomated safety check: PassMIT
Musictadaspetra/loop2962 repos~827Automated safety check: PassMIT

Similar skills

  • AI Music

    ZJU-REAL/Easel

    AI 音乐 / BGM 生成:给短视频、社媒内容生成原创背景音乐 / 配乐 / 纯音乐。通过可插拔 provider(阿里 DashScope / Suno 类第三方 API)文生音乐,异步提交→轮询→下载,产物可再裁剪/归一化或加到视频。当用户说“AI 音乐”“AI 配乐”“生成 BGM”“背景音乐”“原创音乐”“AI 作曲”“纯音乐”“给视频配乐”“做首曲子”时使用。与…

    3.4k GitHub stars~896 tokensUpdated yesterday
    Media & CreativeAuto-check: notes
  • AI Music And Sound

    social-media-skills/skills

    The AI music + sound-design skill for social -- original/licensed audio beds and sound design for Reels/TikToks/Shorts/videos.

    134 GitHub stars~1.9k tokensUpdated 8 days ago
    Media & CreativeAuto-check passed
  • Audio Jingle

    sanqiufong/slides-from-anything

    Audio generation skill — jingles, beds, voiceover, and sound effects.

    132 GitHub starsUsed in 1 repo~1.1k tokens
    Media & CreativeAuto-check passed
  • Suno

    social-media-skills/skills

    The Suno craft skill — generate full songs, brand music, and audio (vocals, lyrics, stems) with the right tier, honest rights, and structure control.

    134 GitHub stars~2.1k tokensUpdated 8 days ago
    Media & CreativeAuto-check passed
  • Music

    tadaspetra/loop

    Generate music using ElevenLabs Music API. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 2 repos~827 tokens
    Media & CreativeAuto-check passed
  • Sound Effects

    tadaspetra/loop

    Generate sound effects from text descriptions using ElevenLabs.

    296 GitHub starsUsed in 2 repos~1.1k tokens
    Media & CreativeAuto-check passed

More from WrongStack/WrongStack

All 100 skills in this repo
  • Tech Stack

    WrongStack/WrongStack

    Validate and upgrade dependencies against live registries and official migration guides in any ecosystem.

    371 GitHub stars~1.5k tokensUpdated today
    Auto-check passed
  • Skill Creator

    WrongStack/WrongStack

    Create, improve and validate WrongStack SKILL.md bundles with precise discovery, progressive resources and current runtime contracts.

    371 GitHub stars~1.3k tokensUpdated today
    Auto-check passed
  • Bug Hunter

    WrongStack/WrongStack

    A skill your agent uses when scanning source code for bugs, anti-patterns, code smells, or quality issues in a codebase, or when running a proof-driven bug hunt that must find, prove, fix, and…

    371 GitHub stars~2.4k tokensUpdated today
    Auto-check passed
  • Design Craft

    WrongStack/WrongStack

    Design or substantially improve user-facing interfaces with a product-specific visual direction, content hierarchy, and rendered critique.

    371 GitHub stars~2.3k tokensUpdated today
    Auto-check passed
  • Design Critique

    WrongStack/WrongStack

    A skill your agent uses to audit an interface that already exists and say precisely why it looks generated, templated, or unfinished — a scored rubric across composition, typography, color, states…

    371 GitHub stars~2.5k tokensUpdated today
    Auto-check passed
  • Mailbox Bridge

    WrongStack/WrongStack

    A skill your agent uses when external coding agents (Claude Code, Aider, custom scripts) need to participate in the project's shared WrongStack mailbox, or when a user asks to "expose the mailbox"…

    371 GitHub stars~2.6k tokensUpdated today
    Auto-check passed

Works with

Questions about Audio Studio

What does Audio Studio do?

Compose music, lyrics, soundscapes and voiceovers, or implement browser audio feedback. Audio Studio is an agent skill from WrongStack/WrongStack. Compose music, lyrics, soundscapes and voiceovers, or implement browser audio feedback.

When should I use Audio Studio?

Audio Studio fits situations like: other music prompts; interactive sound; use media-production for the final video timeline and mux.

How do I install Audio Studio in Claude Code?

Run `npx skills add WrongStack/WrongStack --skill audio-studio -a claude-code`. Or copy the skill folder (packages/core/skills/audio-studio in WrongStack/WrongStack) into .claude/skills/audio-studio in your project. Claude Code loads it when a task matches its description.

How do I install Audio Studio in Codex?

Run `npx skills add WrongStack/WrongStack --skill audio-studio -a codex`. Or copy the skill folder (packages/core/skills/audio-studio in WrongStack/WrongStack) into .agents/skills/audio-studio in your project. Codex loads it when a task matches its description.

Can I use Audio Studio in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add WrongStack/WrongStack --skill audio-studio -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/audio-studio, .gemini/skills/audio-studio, .github/skills/audio-studio and .opencode/skills/audio-studio in your project.

What does Audio Studio need to run?

SKILL.md names no scripts, command-line tools or credentials: Audio Studio is instructions for the agent only.

Does Audio Studio access the network?

SKILL.md names 2 domains. As links in the text: help.suno.com and elevenlabs.io. This is read from the text; nothing was executed.

Is Audio Studio safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Audio Studio use?

Audio Studio is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Audio Studio use?

About 910 tokens (SKILL.md is roughly 3.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Audio Studio?

Skills that share tags, products or a category with Audio Studio: AI Music (ZJU-REAL/Easel, 3.4k stars), AI Music And Sound (social-media-skills/skills, 134 stars), Audio Jingle (sanqiufong/slides-from-anything, 132 stars) and Suno (social-media-skills/skills, 134 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Audio Studio?

WrongStack (a GitHub organization) maintains it in WrongStack/WrongStack, which has 371 GitHub stars. The repository holds 100 skills in this directory. The repository was last updated on October 10, 2026.

Source: WrongStack/WrongStack on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.