Agent skill

Audio Track

by ucsandman in ucsandman/marketing-studio

A skill your agent uses when the user wants music, voiceover, narration, or a soundtrack added to a video asset, OR wants standalone generated audio for any purpose (e.g.

MITAuto-check: notesMedia & Creative

Install Audio Track

skills CLI
$ npx skills add ucsandman/marketing-studio --skill audio-track -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install ucsandman/marketing-studio audio-track --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/ucsandman/marketing-studio.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/audio-track .claude/skills/audio-track && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
audio-track
GitHub stars
251
Token cost
~1.1k tokens
SKILL.md length
567 words
Files
1
Skills in repo
13
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses when the user wants music, voiceover, narration, or a soundtrack added to a video asset, OR wants standalone generated audio for any purpose (e.g.

  • Works in 5 steps: Shared-repo guard + toolchain per… → Copy source of truth:… → Run it (VO lines and the exact-length… → …
  • The user wants music
  • SKILL.md covers Recipe A0: bespoke film, Recipe A: video soundtrack and Recipe B: standalone audio
  • Calls node; needs ELEVENLABS_API_KEY

What it does

Audio Track is an agent skill from ucsandman/marketing-studio. Use when the user wants music, voiceover, narration, or a soundtrack added to a video asset, OR wants standalone generated audio for any purpose (e.g. "/audio-track", "add music to the launch video", "narrate the demo", "make a 30 second music sting", "generate a voiceover mp3", "jingle for the intro").

Its SKILL.md is about 1.1k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering Text to speech and voice and Music and audio generation. The repository describes itself as: Create product launch assets with Remotion, then review and distribute them through guarded, dry-run-first workflows. The licence is MIT.

When your agent uses it

  • The user wants music
  • A soundtrack added to a video asset
  • Wants standalone generated audio for any purpose (e.g

Example prompts

  • “/audio-track”
  • “add music to the launch video”
  • “narrate the demo”
  • “/audio-track”

Requirements

  • A credential in ELEVENLABS_API_KEY

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. Shared-repo guard + toolchain per marketing-studio.
  2. Copy source of truth: scripts/build--film-audio.mjs (copy postflop's).
  3. Run it (VO lines and the exact-length bed hit the API once; re-runs reuse files,
  4. node scripts/score-film.mjs --project — refuses
  5. Listen-proof: node scripts/verify-cue.mjs on one

What it can do on your machine

Read from SKILL.md and the folder at commit 66cd1c3. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • node

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • ELEVENLABS_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Audio Track loads about 1.1k tokens when it runs. Until then it costs about 79 tokens; SKILL.md has 567 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~79
When it runs · the whole SKILL.md, loaded when a task matches
~1.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NoteMentions a .env fileSKILL.md:24
    h need `ELEVENLABS_API_KEY` in the repo `.env` (missing key = feeder exits 2;

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from ucsandman/marketing-studio at commit 66cd1c3, republished under its MIT licence (© ucsandman). 567 words, ~1,125 tokens.

Download SKILL.mdSave it as .claude/skills/audio-track/SKILL.md (or your agent's skills folder).
name
audio-track
description
Use when the user wants music, voiceover, narration, or a soundtrack added to a video asset, OR wants standalone generated audio for any purpose (e.g. "/audio-track", "add music to the launch video", "narrate the demo", "make a 30 second music sting", "generate a voiceover mp3", "jingle for the intro").

Audio Track

REQUIRED BACKGROUND: marketing-studio skill. Work in ${CLAUDE_SKILL_DIR}/../... Read ${CLAUDE_SKILL_DIR}/../../docs/playbook/audio.md first (endpoints, ducking, manifest contract).

Every film gets this pass; it is not opt-in (CLAUDE.md: a film is not done without audio, and audio means narration). node scripts/check-audio.mjs <brand> --project <product> is the exit gate for all three recipes that touch video.

Three modes; pick by what the target is:

  • Bespoke film (studio/src/films/<brand>/, e.g. PostflopFilm) — post-render scoring with scripts/build-<brand>-film-audio.mjs + scripts/score-film.mjs. Recipe A0.
  • Video soundtrack — a directed LaunchVideo composition scored with music and voiceover inside the product workspace. Recipe A.
  • Standalone audio — a music track and/or narration mp3s delivered as files (a sting, a jingle, a narration for something outside this studio). Recipe B.

Both need ELEVENLABS_API_KEY in the repo .env (missing key = feeder exits 2; videos stay silent). Generation costs real money: one pass, trim copy rather than regenerate blindly. Free tier returns 402; Starter or above required.

Recipe A0: bespoke film

  1. Shared-repo guard + toolchain per marketing-studio.
  2. Copy source of truth: scripts/build-<brand>-film-audio.mjs (copy postflop's). Narration lines carry the film SECOND they start on, computed from the film's timeline.ts so a re-timed shot moves its line with it; SFX cues carry the frame the shot lands something on (Stamp/Rule/click constants + the shot's from), max two per beat. Budget ~2.6 words/second minus a 150ms breath per line.
  3. Run it (VO lines and the exact-length bed hit the API once; re-runs reuse files, --force <id,...|music> regenerates). Read the printed start/end table: every line inside its shot.
  4. node scripts/score-film.mjs <brand> <product-film> --project <product> — refuses overlapping lines (trim copy, re-run the builder) and zero-VO manifests; masters by measured gain and verifies the DELIVERED file (I within 0.5 of TARGET_I, TP <= -1).
  5. Listen-proof: node scripts/verify-cue.mjs <scored.mp4> <voStart> <voDur> on one narration window and one bed-only window (voice ~6 dB above the bed), then node scripts/check-audio.mjs <brand> --project <product> (exit 0), then SEND the scored file.
Show full SKILL.md (252 more words)Show less

Recipe A: video soundtrack

  1. Shared-repo guard + toolchain per marketing-studio.
  2. Copy source of truth: scripts/build-<brand>-audio.mjs (copy the noban one for a new brand). VO lines are keyed by act, written FOR THE EAR ("dot gg", not ".gg"), one line per act, terse. Music prompt describes the brand's sonic character.
  3. Run it with --project <product> (music takes 1-3 min). Check every line's duration fits its act; trim TEXT if not, re-run with --force.
  4. Merge with --project <product>, then render into the product workspace using its public directory. Picture and sound consume the same shot plan; explicit audioRef maps voice independently of visual source IDs, while null means no narration.
  5. Listen-proof: verify the mp4 has an audio stream (ffprobe), then SEND the video — audio is approved by ear, by the user. Ducking feel (base 0.35 / duck 0.12) is tunable in studio/src/lib/audioMix.ts if redlined.

Recipe B: standalone audio

  1. Shared-repo guard per marketing-studio. The feeder code lives in the engine, but all generated audio lives in the product workspace.
  2. Music: node feeders/audio/client.mjs music --project <product> --brand <brand> --prompt "<sonic character>" --length-ms <n> --out <product-output>.mp3 (exact-length, commercially licensed on paid plans; 3s-120s). Voiceover: write {lines: [{id, text}]} to a temp JSON (text written for the ear), then run the voice feeder with --project <product> --brand <brand> and a product-owned output. Durations print as ... OK: <name> <ms>ms; probe --file <mp3> re-measures.
  3. Listen-proof: SEND the product-owned mp3(s) to the user. Machine loudness, beat, and cue reports remain incomplete until a human listens.

© ucsandman, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/audio-track of ucsandman/marketing-studio.

Open the folder on GitHubat commit 66cd1c3

Compare with similar skills

Audio Track next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Audio Track compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Audio Track this skillucsandman/marketing-studio251—~1.1kAutomated safety check: NotesMIT
HyperFrames Media Useheygen-com/hyperframes59k—~2.4kAutomated safety check: PassApache-2.0
Musictadaspetra/loop2962 repos~827Automated safety check: PassMIT
Sound Effectstadaspetra/loop2962 repos~1.1kAutomated safety check: PassMIT
Characteristic VoiceNoizAI/skills526—~1.8kAutomated safety check: PassNone
Sound FxNoizAI/skills526—~1.4kAutomated safety check: PassNone

Similar skills

  • HyperFrames Media Use

    heygen-com/hyperframes

    Finds, generates and edits media for HyperFrames video projects: music, sound effects, images, icons, logos, voiceovers, captions and color grades.

    59k GitHub stars~2.4k tokensUpdated today
    Media & CreativeAuto-check passed
  • Music

    tadaspetra/loop

    Generate music using ElevenLabs Music API. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 2 repos~827 tokens
    Media & CreativeAuto-check passed
  • Sound Effects

    tadaspetra/loop

    Generate sound effects from text descriptions using ElevenLabs.

    296 GitHub starsUsed in 2 repos~1.1k tokens
    Media & CreativeAuto-check passed
  • A skill your agent uses whenever the user wants speech to sound more human, companion-like, or emotionally expressive.

    526 GitHub stars~1.8k tokensUpdated 11 days ago
    Media & CreativeAuto-check passed
  • Sound Fx

    NoizAI/skills

    A skill your agent uses whenever the user wants to generate sound effects, ambient audio, or short audio clips from a text description.

    526 GitHub stars~1.4k tokensUpdated 11 days ago
    Media & CreativeAuto-check passed
  • Z Qwen Audio Studio

    tjxj/z-skills

    A skill your agent uses when creating complete generated audio with qwen-audio-3.1-tts-next, including podcasts, radio drama, advertisements, multiple speakers, reference voices, ambience, sound…

    547 GitHub stars~716 tokensUpdated 17 days ago
    Media & CreativeAuto-check passed

More from ucsandman/marketing-studio

All 13 skills in this repo
  • Frontend Verify

    ucsandman/marketing-studio

    Verify frontend changes end to end after editing a web app, instead of manually clicking through pages.

    251 GitHub stars~2.4k tokensUpdated 1 mo ago
    Auto-check passed
  • Launch

    ucsandman/marketing-studio

    Take a developed project end-to-end through launch — domain, hosting, payments, email infra, algorithm-researched copy, and multi-platform distribution — with zero copy-paste.

    251 GitHub stars~2.2k tokensUpdated 1 mo ago
    Auto-check: notes
  • Announce

    ucsandman/marketing-studio

    A skill your agent uses when a product or feature is ready to go public — after shipping, when the user says "announce it", "go live", "tell people about it", or wants a release marketed end to end.

    251 GitHub stars~1.5k tokensUpdated 1 mo ago
    Auto-check passed
  • Launch Video

    ucsandman/marketing-studio

    A skill your agent uses when the user wants a full launch video, hero video, or 20–60s product film combining product proof, brand, narration, music, and authored motion.

    251 GitHub stars~980 tokensUpdated 1 mo ago
    Auto-check passed
  • Marketing

    ucsandman/marketing-studio

    A skill your agent uses when the user wants the complete marketing asset suite for a product in one run — "/marketing", "build all the marketing assets", "generate everything for the launch", "all…

    251 GitHub stars~5.5k tokensUpdated 1 mo ago
    Auto-check passed
  • Marketing Studio

    ucsandman/marketing-studio

    A skill your agent uses when generating any brand video, animation, image, or audio asset (logo reveal, social clip, product demo, launch video, OG image, README GIF, music, voiceover) for any…

    251 GitHub stars~1k tokensUpdated 1 mo ago
    Auto-check passed

Questions about Audio Track

What does Audio Track do?

A skill your agent uses when the user wants music, voiceover, narration, or a soundtrack added to a video asset, OR wants standalone generated audio for any purpose (e.g. Audio Track is an agent skill from ucsandman/marketing-studio.g.

When should I use Audio Track?

Audio Track fits situations like: the user wants music; A soundtrack added to a video asset; wants standalone generated audio for any purpose (e.g.

How do I install Audio Track in Claude Code?

Run `npx skills add ucsandman/marketing-studio --skill audio-track -a claude-code`. Or copy the skill folder (skills/audio-track in ucsandman/marketing-studio) into .claude/skills/audio-track in your project. Claude Code loads it when a task matches its description.

How do I install Audio Track in Codex?

Run `npx skills add ucsandman/marketing-studio --skill audio-track -a codex`. Or copy the skill folder (skills/audio-track in ucsandman/marketing-studio) into .agents/skills/audio-track in your project. Codex loads it when a task matches its description.

Can I use Audio Track in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add ucsandman/marketing-studio --skill audio-track -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/audio-track, .gemini/skills/audio-track, .github/skills/audio-track and .opencode/skills/audio-track in your project.

What does Audio Track need to run?

Going by SKILL.md and its folder, Audio Track needs the command-line tools its instructions call (node) and credentials named ELEVENLABS_API_KEY. Our summary lists: A credential in ELEVENLABS_API_KEY.

Does Audio Track access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Audio Track safe to install?

Our automated static check of SKILL.md found notes only (mentions a .env file), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.

What licence does Audio Track use?

Audio Track is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Audio Track use?

About 1.1k tokens (SKILL.md is roughly 4.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Audio Track?

Skills that share tags, products or a category with Audio Track: HyperFrames Media Use (heygen-com/hyperframes, 59k stars), Music (tadaspetra/loop, 296 stars), Sound Effects (tadaspetra/loop, 296 stars) and Characteristic Voice (NoizAI/skills, 526 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Audio Track?

ucsandman (a GitHub user) maintains it in ucsandman/marketing-studio, which has 251 GitHub stars. The repository holds 13 skills in this directory. The repository was last updated on September 6, 2026.

Source: ucsandman/marketing-studio on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.