Agent skill

Scenario Elevenlabs

by scenario-labs in scenario-labs/skills

A skill your agent uses when generating or transforming audio with ElevenLabs models on Scenario via MCP: text-to-speech with inline emotion tags, music from a prompt or sectioned songs, sound…

MITAuto-check passedMedia & Creative

Install Scenario Elevenlabs

skills CLI
$ npx skills add scenario-labs/skills --skill scenario-elevenlabs -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install scenario-labs/skills scenario-elevenlabs --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/scenario-labs/skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/scenario-elevenlabs .claude/skills/scenario-elevenlabs && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
scenario-elevenlabs
GitHub stars
946
Token cost
~2k tokens
SKILL.md length
998 words
Files
1
Skills in repo
146
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses when generating or transforming audio with ElevenLabs models on Scenario via MCP: text-to-speech with inline emotion tags, music from a prompt or sectioned songs, sound…

  • Works in 6 steps: search with target="models",… → model_schema_get with that id: fields… → upload_asset the trailer video (see the… → …
  • Transforming audio with ElevenLabs models on Scenario via MCP: text-to-speech with inline emotion tags
  • SKILL.md covers Overview, Quick reference, Speech: tags on v3, dials… and Music: one prompt or thirty…, plus 3 more sections
  • Calls npx

What it does

Scenario Elevenlabs is an agent skill from scenario-labs/skills. Use when generating or transforming audio with ElevenLabs models on Scenario via MCP: text-to-speech with inline emotion tags, music from a prompt or sectioned songs, sound effects and loops, dubbing audio or video into another language, re-voicing a recording with a preset or cloned voice, or isolating speech from noise. Keywords: ElevenLabs, Eleven v3, Multilingual v2, Turbo 2.5, Music v2, Sound Effects 2, Dubbing, Voice Changer, Speech to Speech, Voice Isolator, TTS, SFX, localization.

Its SKILL.md is about 2k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering Text to speech and voice. It works with ElevenLabs and Model Context Protocol. The repository describes itself as: Get production-ready images, video, audio, and 3D from any AI agent: skills that pick the right model, price before spending, and keep characters and brands consistent through… The licence is MIT.

When your agent uses it

  • Transforming audio with ElevenLabs models on Scenario via MCP: text-to-speech with inline emotion tags
  • Music from a prompt
  • Sectioned songs
  • Sound effects and loops

Example prompts

  • “/scenario-elevenlabs”

Requirements

  • Node.js

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. search with target="models", query="elevenlabs dubbing", public=true. Prefer the newest non-deprecated hit, e.g…
  2. model_schema_get with that id: fields and defaults before anything else.
  3. upload_asset the trailer video (see the scenario skill) to get an asset id.
  4. model_run with that model_id, dry_run=true, and the exact parameters={"file": "asset_x", "targetLang": "es", "keyterms": ["Aetherfall"…
  5. Repeat model_run with wait=false, then jobs_wait with the returned job id, re-called with pending_job_ids on timeout, never a second…
  6. asset_display the result (a video, because the input was) and confirm the keyterms survived; asset_download with no format to save it.

What it can do on your machine

Read from SKILL.md and the folder at commit f6f8ab7. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • npx

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use npx, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Scenario Elevenlabs loads about 2k tokens when it runs. Until then it costs about 128 tokens; SKILL.md has 998 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~128
When it runs · the whole SKILL.md, loaded when a task matches
~2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from scenario-labs/skills at commit f6f8ab7, republished under its MIT licence (© scenario-labs). 998 words, ~2,012 tokens.

Download SKILL.mdSave it as .claude/skills/scenario-elevenlabs/SKILL.md (or your agent's skills folder).
name
scenario-elevenlabs
description
Use when generating or transforming audio with ElevenLabs models on Scenario via MCP: text-to-speech with inline emotion tags, music from a prompt or sectioned songs, sound effects and loops, dubbing audio or video into another language, re-voicing a recording with a preset or cloned voice, or isolating speech from noise. Keywords: ElevenLabs, Eleven v3, Multilingual v2, Turbo 2.5, Music v2, Sound Effects 2, Dubbing, Voice Changer, Speech to Speech, Voice Isolator, TTS, SFX, localization.
license
MIT

Scenario ElevenLabs Audio

Overview

ElevenLabs covers five audio lanes on Scenario: speech, music, sound effects, dubbing, and voice transforms. Eleven members ship side by side at authoring time, one job each: the deciding move is picking the right member, not coaxing one model into a mode. Discover them with search and treat model_schema_get as the contract: the members agree on almost nothing.

Connection and the core loop: see the scenario skill in this repo; model-agnostic audio work: the scenario-audio skill. If a sibling skill named here is missing from your available skills, ask the user to install it (npx skills add scenario-labs/skills --skill <name>); unattended, proceed from tool schemas and flag the gap.

Quick reference

JobMemberDeciding inputs
Expressive speech, dialogEleven v3text with inline tags
Narration, pinned languageMultilingual 2languageCode
Fast or bulk speechTurbo 2.5same shape, lower cost
Music from one promptMusic v2prompt, durationSeconds, forceInstrumental
Structured songMusic Advanced v2required sections array
SFX and loopsSound Effects 2text, promptInfluence, loop
Localize, protect namesDubbing v2file, targetLang, keyterms
Localize, speaker controlDubbingnumSpeakers, dropBackgroundAudio
Re-voice a recordingSpeech to Speech, Voice Changeraudio (up to 5 min) plus a voice
Clean a noisy voiceVoice Isolatoraudio

The speaking members share one voice contract: voiceId takes a cloned ElevenLabs voice and always wins; publicVoice picks a preset (21 at authoring time) and is ignored beside it; neither set means the Adam preset. Voice Changer is Speech to Speech plus delivery dials (stability, similarityBoost, styleExaggeration, useSpeakerBoost). WAV output exists on the TTS members alone; outputFormat elsewhere is MP3 or Opus, absent on Dubbing and Voice Isolator. seed gives repeatability, except on Sound Effects 2, both Dubbing members, and Voice Isolator. Cost rides the content: text on TTS (40,000 characters at authoring time), durationSeconds on Music v2, sections on Advanced, file or audio on the dubbing and voice members, so dry_run long jobs and member comparisons.

Speech: tags on v3, dials elsewhere

Eleven v3 reads inline audio tags in the text, [whispers], [excited], [sighs], to steer delivery moment to moment across 70+ languages, and carries multi-speaker dialogue. Multilingual 2 and Turbo 2.5 do not read tags: direction there lives in stability, styleExaggeration, and speed (0.7 to 1.2), and languageCode (ISO 639-1) pins the language, a field v3 lacks. Turbo trades expressiveness for cost, roughly half the other two per run at authoring time.

Music: one prompt or thirty sections

Music v2 takes one prompt (mood, genre, instruments, tempo), durationSeconds (3 to 600 at authoring time, cost impact), and forceInstrumental to suppress vocals: asking in prose is unreliable. Music Advanced v2 requires sections, up to 30 ordered segments at authoring time, each with text, its own durationSeconds (3 to 120), positiveStyles and negativeStyles (up to 10 each), and contextAdherence (high binds a segment to its neighbors, low frees it). Section grammar: square brackets label ([Verse], [Chorus]), curly braces direct ({soft piano intro}), and plain text is sung as lyrics. Advanced has no instrumental flag, so any plain text will be sung. Neither music member takes numOutputs: one run is one composition, and sung lyrics land differently on every draw, so a batch of takes is several parallel runs with distinct seed values, each priced alone on its duration, which costs what a batch field would have and returns separate files with real boundaries. Two [Verse] blocks in one run to get two takes returns one file with no split point.

Show full SKILL.md (421 more words)Show less

Dubbing replaces the track, not the lips

Both Dubbing members take an audio or video file with a required targetLang, and the output follows the input kind: video in, dubbed video out. Neither re-animates lips. Source auto-detection is spelled differently: sourceLang: "auto" on v2, an empty string on the older member. Pick v2 for keyterms (names, brands, and jargon preserved verbatim through translation); pick the older Dubbing for numSpeakers (0 auto-detects, up to 10), dropBackgroundAudio, disableVoiceCloning, and highestResolution video. Their prices differ several-fold, so dry_run both when either fits. Speech to Speech, Voice Changer and Voice Isolator, by contrast, take audio only, whatever the catalog blurb says: for a clip, pull the track with model_scenario-audio-extract, run the member, then lay the result over the clip in model_scenario-compose-video, the clip as a video layer with mute: true (fixed ids: each is Scenario's single deterministic tool for its operation, so discovery would only re-derive it). The same compositor lays a Music v2 bed under a video; layer contract in scenario-video-assembly.

Worked example: dub a trailer into Spanish

  1. search with target="models", query="elevenlabs dubbing", public=true. Prefer the newest non-deprecated hit, e.g. model_elevenlabs-dubbing-v2 (a live hit at authoring time: re-discover each session).
  2. model_schema_get with that id: fields and defaults before anything else.
  3. upload_asset the trailer video (see the scenario skill) to get an asset id.
  4. model_run with that model_id, dry_run=true, and the exact parameters={"file": "asset_x", "targetLang": "es", "keyterms": ["Aetherfall", "Kestrel Squad"]}. The file drives the price: re-estimate per input.
  5. Repeat model_run with wait=false, then jobs_wait with the returned job id, re-called with pending_job_ids on timeout, never a second model_run.
  6. asset_display the result (a video, because the input was) and confirm the keyterms survived; asset_download with no format to save it.

Common mistakes

  • Setting publicVoice next to voiceId and expecting it to apply: a set voiceId always wins.
  • Inline tags on Multilingual 2 or Turbo 2.5: that grammar is Eleven v3's; elsewhere a tag can be read aloud.
  • Writing "instrumental" in a Music v2 prompt instead of setting forceInstrumental; on Music Advanced there is no flag and plain section text is always sung.
  • Expecting lip-sync from Dubbing: the track changes, the picture does not.
  • Reaching for numOutputs on the music members, or repeating one long run hoping for a different take without changing seed: same seed and settings reproduce the same music.
  • Carrying one member's caps to another: 40,000 characters, 600 seconds, 30 sections, and 5 minutes of input audio are each true of one lane and false of the next.

© scenario-labs, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/scenario-elevenlabs of scenario-labs/skills.

Open the folder on GitHubat commit f6f8ab7

Compare with similar skills

Scenario Elevenlabs next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Scenario Elevenlabs compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Scenario Elevenlabs this skillscenario-labs/skills946—~2kAutomated safety check: PassMIT
Remotion ProductionDojoCodingLabs/remotion-superpowers132—~1.1kAutomated safety check: PassMIT
Text To Speechcalesthio/OpenMontage66k—~2.4kAutomated safety check: PassAGPL-3.0
Story Narratorhassancs91/claude-image-generation102—~2.5kAutomated safety check: PassMIT
Elevenlabs Agentsjezweb/claude-skills1.1k—~3.3kAutomated safety check: PassMIT
Making Demo Videosnukeop/nuclear19k—~892Automated safety check: PassAGPL-3.0

Similar skills

  • Remotion Production

    DojoCodingLabs/remotion-superpowers

    Full video production workflow for Remotion projects. An agent skill from DojoCodingLabs/remotion-superpowers.

    132 GitHub stars~1.1k tokensUpdated 7 days ago
    Media & CreativeAuto-check passed
  • Text To Speech

    calesthio/OpenMontage

    Generate speech audio from text using HeyGen's Starfish TTS model.

    66k GitHub stars~2.4k tokensUpdated 8 days ago
    Media & CreativeAuto-check passed
  • Story Narrator

    hassancs91/claude-image-generation

    Generates one expressive narration MP3 per scene for an English storybook, using the ElevenLabs MCP (texttospeech).

    102 GitHub stars~2.5k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Elevenlabs Agents

    jezweb/claude-skills

    Build conversational AI voice agents on the ElevenLabs platform.

    1.1k GitHub stars~3.3k tokensUpdated 2 days ago
    AI & LLM EngineeringAuto-check passed
  • Making Demo Videos

    nukeop/nuclear

    A skill your agent uses when making a demo, tutorial, or feature video of Nuclear.

    19k GitHub stars~892 tokensUpdated 5 days ago
    Media & CreativeAuto-check passed
  • Music

    tadaspetra/loop

    Generate music using ElevenLabs Music API. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 2 repos~827 tokens
    Media & CreativeAuto-check passed

More from scenario-labs/skills

All 146 skills in this repo
  • Scenario Blender Grease Pencil

    scenario-labs/skills

    A skill your agent uses when drawing or animating with Grease Pencil in Blender 5.x from Python: 2D or 2.5D illustration, frame-by-frame animation, a cutout or part-based 2D character, strokes with…

    946 GitHub stars~4.5k tokensUpdated yesterday
    Auto-check passed
  • Scenario Blender Hair

    scenario-labs/skills

    A skill your agent uses when grooming hair or fur in Blender with hair curves, such as a character hairstyle, animal fur, procedural fur in geometry nodes, or hair cards and mesh hair for games.

    946 GitHub stars~4.7k tokensUpdated yesterday
    Auto-check passed
  • A skill your agent uses when lighting, rendering or compositing in Blender: light a character, product or hero shot, interior at dusk or night, three-point or motivated lighting, sun and sky, HDRI…

    946 GitHub stars~5k tokensUpdated yesterday
    Auto-check passed
  • Scenario Chatgpt Pet Create

    scenario-labs/skills

    A skill your agent uses when creating a ChatGPT pet or Codex pet with Scenario: hatching an animated companion from a text idea, a character, mascot or brand cue, or reference photos and art; making…

    946 GitHub stars~3.6k tokensUpdated yesterday
    Auto-check passed
  • Scenario Godot Animation

    scenario-labs/skills

    A skill your agent uses when animating characters or scenes in Godot 4.7: AnimationPlayer clips and RESET, AnimationTree state machines and blend spaces built in code, Mixamo or glTF import, loop…

    946 GitHub stars~4.7k tokensUpdated yesterday
    Auto-check passed
  • Scenario Godot Audio

    scenario-labs/skills

    A skill your agent uses when adding or fixing sound in Godot 4.7: audio buses and effects, volume sliders, 'too many sounds', combat audio with hundreds of enemies, sounds clipping or distorting, 3D…

    946 GitHub stars~4.7k tokensUpdated yesterday
    Auto-check passed

Questions about Scenario Elevenlabs

What does Scenario Elevenlabs do?

A skill your agent uses when generating or transforming audio with ElevenLabs models on Scenario via MCP: text-to-speech with inline emotion tags, music from a prompt or sectioned songs, sound…. Scenario Elevenlabs is an agent skill from scenario-labs/skills. Use when generating or transforming audio with ElevenLabs models on Scenario via MCP: text-to-speech with inline emotion tags, music from a prompt or sectioned songs, sound effects and loops, dubbing audio or video into another language, re-voicing a recording with a preset or cloned voice, or isolating speech from noise.

When should I use Scenario Elevenlabs?

Scenario Elevenlabs fits situations like: transforming audio with ElevenLabs models on Scenario via MCP: text-to-speech with inline emotion tags; music from a prompt; sectioned songs; sound effects and loops.

How do I install Scenario Elevenlabs in Claude Code?

Run `npx skills add scenario-labs/skills --skill scenario-elevenlabs -a claude-code`. Or copy the skill folder (skills/scenario-elevenlabs in scenario-labs/skills) into .claude/skills/scenario-elevenlabs in your project. Claude Code loads it when a task matches its description.

How do I install Scenario Elevenlabs in Codex?

Run `npx skills add scenario-labs/skills --skill scenario-elevenlabs -a codex`. Or copy the skill folder (skills/scenario-elevenlabs in scenario-labs/skills) into .agents/skills/scenario-elevenlabs in your project. Codex loads it when a task matches its description.

Can I use Scenario Elevenlabs in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add scenario-labs/skills --skill scenario-elevenlabs -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/scenario-elevenlabs, .gemini/skills/scenario-elevenlabs, .github/skills/scenario-elevenlabs and .opencode/skills/scenario-elevenlabs in your project.

What does Scenario Elevenlabs need to run?

Going by SKILL.md and its folder, Scenario Elevenlabs needs the command-line tools its instructions call (npx). Our summary lists: Node.js.

Does Scenario Elevenlabs access the network?

SKILL.md contains no URLs. Its commands use npx, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Scenario Elevenlabs safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Scenario Elevenlabs use?

Scenario Elevenlabs is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Scenario Elevenlabs use?

About 2k tokens (SKILL.md is roughly 8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Scenario Elevenlabs?

Skills that share tags, products or a category with Scenario Elevenlabs: Remotion Production (DojoCodingLabs/remotion-superpowers, 132 stars), Text To Speech (calesthio/OpenMontage, 66k stars), Story Narrator (hassancs91/claude-image-generation, 102 stars) and Elevenlabs Agents (jezweb/claude-skills, 1.1k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Scenario Elevenlabs?

scenario-labs (a GitHub organization) maintains it in scenario-labs/skills, which has 946 GitHub stars. The repository holds 146 skills in this directory. The repository was last updated on October 10, 2026.

Source: scenario-labs/skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.