Agent skill

Elevenlabs Voice

by vellum-ai in vellum-ai/vellum-assistant

Select and tune an ElevenLabs TTS voice - curated voice list, custom/cloned voices via API key, and tuning parameters

MITAuto-check passedMedia & Creative

Install Elevenlabs Voice

skills CLI
$ npx skills add vellum-ai/vellum-assistant --skill elevenlabs-voice -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install vellum-ai/vellum-assistant elevenlabs-voice --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/vellum-ai/vellum-assistant.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/elevenlabs-voice .claude/skills/elevenlabs-voice && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
elevenlabs-voice
GitHub stars
1.4k
Token cost
~2.3k tokens
SKILL.md length
943 words
Files
2 (incl. assets)
Skills in repo
108
Repo updated
First seen
Licence
MIT

At a glance

Select and tune an ElevenLabs TTS voice - curated voice list, custom/cloned voices via API key, and tuning parameters

  • Works in 2 steps: Switch to managed vellum… → Switch to BYO elevenlabs…
  • Tasks that involve Text to speech and voice
  • SKILL.md covers Overview, Getting to an ElevenLabs voice…, Choose a Voice and ElevenLabs API Key Setup, plus 3 more sections
  • Calls curl and python3; reaches api.elevenlabs.io

What it does

Elevenlabs Voice is an agent skill from vellum-ai/vellum-assistant. Select and tune an ElevenLabs TTS voice - curated voice list, custom/cloned voices via API key, and tuning parameters

Its SKILL.md is about 2.3k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including assets. Compatibility notes: Designed for Vellum personal assistants

It sits in Media & Creative, covering Text to speech and voice. It works with ElevenLabs and Deepgram. The repository describes itself as: An AI Assistant that’s easy to setup, does your work 24/7, knows your preferences and gets better over time. The licence is MIT.

When your agent uses it

  • Tasks that involve Text to speech and voice

Example prompts

  • “/elevenlabs-voice”

Requirements

  • Python 3
  • Compatibility (from SKILL.md): Designed for Vellum personal assistants

Workflow steps

2 steps, taken from the first numbered list in SKILL.md.

  1. Switch to managed vellum (voice_config_update setting="tts_provider" value="vellum") — no ElevenLabs key needed; requires a platform…
  2. Switch to BYO elevenlabs (voice_config_update setting="tts_provider" value="elevenlabs") — requires an ElevenLabs API key (setup below)…

What it can do on your machine

Read from SKILL.md and the folder at commit 33cc983. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • curl
    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • api.elevenlabs.io

    Also links to:

    • elevenlabs.io

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

  • Compatibility

    Designed for Vellum personal assistants

    From compatibility in the SKILL.md frontmatter.

Context cost

Elevenlabs Voice loads about 2.3k tokens when it runs. Until then it costs about 34 tokens; SKILL.md has 943 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~34
When it runs · the whole SKILL.md, loaded when a task matches
~2.3k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from vellum-ai/vellum-assistant at commit 33cc983, republished under its MIT licence (© vellum-ai). 943 words, ~2,332 tokens.

Download SKILL.mdSave it as .claude/skills/elevenlabs-voice/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
elevenlabs-voice
description
Select and tune an ElevenLabs TTS voice - curated voice list, custom/cloned voices via API key, and tuning parameters
compatibility
Designed for Vellum personal assistants
metadata.icon
assets/icon.svg
metadata.emoji
🗣️

Overview

ElevenLabs provides text-to-speech voices for both in-app TTS and phone calls. Change the voice with the voice_config_update tool — it writes the voice to whichever TTS provider is currently active and pushes to the macOS app via SSE in one call:

voice_config_update setting="tts_voice_id" value="<voice-id>"

The voice lives under the active provider, not always ElevenLabs. The config key depends on services.tts.provider: elevenlabs → services.tts.providers.elevenlabs.voiceId, vellum (managed) → services.tts.providers.vellum.model, deepgram → services.tts.providers.deepgram.model. The voice_config_update tool (and the assistant tts voice <id> CLI command) handle this routing for you. Do NOT assistant config set services.tts.providers.elevenlabs.voiceId ... blindly — on a managed (vellum) assistant that field is ignored, so the write "succeeds" but the voice never changes. See Setting the voice for the CLI fallback.

The tables below apply when the active provider is elevenlabs (BYO key) or managed vellum. On managed assistants they are the only ElevenLabs voices — the platform bills per rate-carded model and rejects anything else at synthesis (the write succeeds but the voice fails on the next turn), so don't offer library or cloned voices unless the elevenlabs provider is active with its own API key. With a BYO key (setup below) and elevenlabs active, any voice id works. Other BYO TTS providers (deepgram, xai, fish-audio, …) use their own voice/model identifiers — never write an ElevenLabs voice id to them; see Getting to an ElevenLabs voice from another provider.

Getting to an ElevenLabs voice from another provider

Check the active provider first: assistant config get services.tts.provider.

  • Already on managed vellum? No provider change needed. The managed platform supports both ElevenLabs and Deepgram voices — services.tts.providers.vellum.model accepts either an ElevenLabs voice id or an Aura model id, so switching between them is just another voice_config_update call.
  • On a BYO provider (e.g. deepgram) and the user wants an ElevenLabs voice? Two options — ask which they prefer. Switch with the voice_config_update tool, not raw assistant config set — the tool validates the switch (e.g. rejects vellum when no platform connection exists, which a raw config write would leave silently broken):
    1. Switch to managed vellum (voice_config_update setting="tts_provider" value="vellum") — no ElevenLabs key needed; requires a platform connection and bills managed credits. Bonus: they keep access to both the ElevenLabs and Deepgram catalogs.
    2. Switch to BYO elevenlabs (voice_config_update setting="tts_provider" value="elevenlabs") — requires an ElevenLabs API key (setup below); usage bills their ElevenLabs account directly.

After either switch, set the voice with voice_config_update as usual.

Choose a Voice

Pick a voice that matches your identity and the user's preferences. Offer to show the full list if they want to choose themselves.

Female voices
VoiceStyleVoice ID
SarahSoft, young, approachableEXAVITQu4vr4xnSDxMaL
AliceConfident, BritishXb7hH8MSUJpSbSDYk0k2
MatildaWarm, friendly, youngXrExE9yKIg1WjnnlVkGX
LilyWarm, BritishpFZP5JQG7iQjIQuC4Bku
Male voices
VoiceStyleVoice ID
AdamDeep, middle-aged, professionalpNInz6obpgDQGcFmaJgB
BillTrustworthy, AmericanpqHfZKP75CvOlQylNhV4
GeorgeWarm, British, distinguishedJBFqnCBsd6RMkjVDRZzb
DanielAuthoritative, BritishonwK4e9ZLuTAKqWW03F9
CharlieCasual, AustralianIKne3meq5aSn9XLyUdCD
LiamYoung, articulateTX3LPaxmHKxFdv7VOQHJ

These are ElevenLabs' current premade voices. Do not use retired legacy ids (Antoni, Josh, Arnold, Rachel, Charlotte, Amelia, …): ElevenLabs silently remaps them to different voices — synthesis succeeds but speaks as the wrong voice.

Show full SKILL.md (440 more words)Show less
Setting the voice

Preferred — the tool. It writes to the active provider's voice field and pushes to the macOS app via SSE (ttsVoiceId) in one call:

voice_config_update setting="tts_voice_id" value="<selected-voice-id>"

CLI fallback (only if the voice_config_update tool is unavailable). Use assistant tts voice, which routes to the active provider's config key for you — do not hand-write assistant config set services.tts.providers.elevenlabs.voiceId ...:

bash
assistant tts voice "<selected-voice-id>"

Setting services.tts.providers.elevenlabs.voiceId directly while the active provider is vellum (or any non-elevenlabs provider) is the #1 cause of "I changed the voice but it didn't change" — that field is ignored by the active provider, so the write reports success but nothing changes. If you must use config set, first check assistant config get services.tts.provider and write the matching key (vellum → services.tts.providers.vellum.model, deepgram → services.tts.providers.deepgram.model).

Verify it worked by reading back the key for the active provider, e.g. for a managed assistant:

bash
assistant config get services.tts.providers.vellum.model

The change hot-applies to the next voice turn (live voice and phone read the config fresh each turn).

Tell the user what voice you chose and why, but also offer to show all available voices so they can choose for themselves.

ElevenLabs API Key Setup

For advanced voice selection (browsing the full library, custom/cloned voices), the user needs an ElevenLabs API key. A free tier is available at https://elevenlabs.io.

To collect the API key securely:

bash
assistant credentials prompt --service elevenlabs --field api_key --label "ElevenLabs API Key"

Advanced Voice Selection (with API key)

Users with an ElevenLabs API key can go beyond the curated list above — only when the active provider is elevenlabs (BYO key). On managed vellum, stay with the curated voices: the platform only accepts its rate-carded subset, and an unoffered id is persisted successfully but fails on the next spoken turn.

Check for an existing key
bash
assistant credentials inspect --service elevenlabs --field api_key --json
Browse the voice library
bash
curl -s "https://api.elevenlabs.io/v2/voices?category=premade&page_size=50" \
  -H "xi-api-key: $(assistant credentials reveal --service elevenlabs --field api_key)" | python3 -m json.tool
Search for a specific style
bash
curl -s "https://api.elevenlabs.io/v2/voices?search=warm+female&page_size=10" \
  -H "xi-api-key: $(assistant credentials reveal --service elevenlabs --field api_key)" | python3 -m json.tool
Custom and cloned voices

If the user has created a custom voice or voice clone in their ElevenLabs account, they can use its voice ID directly. These voices work in both in-app TTS and phone calls.

Preview voices

Each voice in the API response includes a preview_url with an audio sample the user can listen to before deciding.

Set the chosen voice

After the user picks a voice from the library:

voice_config_update setting="tts_voice_id" value="<selected-voice-id>"

Voice Tuning

Fine-tune how the selected voice sounds. These parameters apply to all ElevenLabs modes (in-app TTS and phone calls) when the active provider is elevenlabs — managed (vellum) synthesis does not read them:

bash
# Playback speed (0.7 = slower, 1.0 = normal, 1.2 = faster)
assistant config set services.tts.providers.elevenlabs.speed 1.0

# Stability (0.0 = more expressive/variable, 1.0 = more consistent/monotone)
assistant config set services.tts.providers.elevenlabs.stability 0.5

# Similarity boost (0.0 = more creative, 1.0 = closer to original voice)
assistant config set services.tts.providers.elevenlabs.similarityBoost 0.75

Lower stability makes the voice more expressive but less predictable - good for conversational calls. Higher stability is better for scripted or formal contexts.

Voice Model Tuning

By default, synthesis uses ElevenLabs' eleven_multilingual_v2 model. To use a different model (e.g. a lower-latency one), set a model ID:

bash
assistant config set services.tts.providers.elevenlabs.voiceModelId "eleven_flash_v2_5"

To clear and revert to the default model:

bash
assistant config set services.tts.providers.elevenlabs.voiceModelId ""

© vellum-ai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (assets) in skills/elevenlabs-voice of vellum-ai/vellum-assistant.

  • SKILL.md
  • assets/icon.svg

Open the folder on GitHubat commit 33cc983

Compare with similar skills

Elevenlabs Voice next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Elevenlabs Voice compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Elevenlabs Voice this skillvellum-ai/vellum-assistant1.4k—~2.3kAutomated safety check: PassMIT
9Router Text to Speechdecolua/9router31k—~765Automated safety check: PassMIT
Keirouter Ttsmydisha/keirouter147—~599Automated safety check: PassMIT
Voice AI Developmentdavila7/claude-code-templates33k5 repos~2.1kAutomated safety check: PassMIT
Musictadaspetra/loop2962 repos~827Automated safety check: PassMIT
ElevenLabs Voiceover Generatordigitalsamba/claude-code-video-toolkit2.2k1 repos~2.7kAutomated safety check: NotesMIT

Similar skills

  • 9Router Text to Speech

    decolua/9router

    Turns text into spoken audio through a 9Router server, choosing among voices from OpenAI, ElevenLabs, Deepgram, Edge TTS and other providers.

    31k GitHub stars~765 tokensUpdated 2 days ago
    Media & CreativeAuto-check passed
  • Keirouter Tts

    mydisha/keirouter

    Text-to-speech via KeiRouter /v1/audio/speech using OpenAI / ElevenLabs / Deepgram / Edge TTS / Google TTS / Inworld voices.

    147 GitHub stars~599 tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Voice AI Development

    davila7/claude-code-templates

    Expert in building voice AI applications - from real-time voice agents to voice-enabled apps.

    33k GitHub starsUsed in 5 repos~2.1k tokens
    Media & CreativeAuto-check passed
  • Music

    tadaspetra/loop

    Generate music using ElevenLabs Music API. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 2 repos~827 tokens
    Media & CreativeAuto-check passed
  • ElevenLabs Voiceover Generator

    digitalsamba/claude-code-video-toolkit

    Generates narration, sound effects and cloned voices through the ElevenLabs API, with model and setting choices tuned to the content's style.

    2.2k GitHub starsUsed in 1 repo~2.7k tokens
    Media & CreativeAuto-check: notes
  • Sound Effects

    tadaspetra/loop

    Generate sound effects from text descriptions using ElevenLabs.

    296 GitHub starsUsed in 2 repos~1.1k tokens
    Media & CreativeAuto-check passed

More from vellum-ai/vellum-assistant

All 108 skills in this repo
  • Vellum GitHub App Setup

    vellum-ai/vellum-assistant

    Create and configure a GitHub App so the assistant can push commits, open PRs, and comment under its own bot identity.

    1.4k GitHub stars~3.1k tokensUpdated yesterday
    Auto-check passed
  • Discord App Setup

    vellum-ai/vellum-assistant

    Connect a Discord bot to the assistant via the Discord Gateway with guided application creation and intent configuration

    1.4k GitHub stars~4.2k tokensUpdated yesterday
    Auto-check passed
  • Sentry App Setup

    vellum-ai/vellum-assistant

    Create and configure a Sentry internal integration so the assistant can manage issues, alerts, and releases under its own identity

    1.4k GitHub stars~1.3k tokensUpdated yesterday
    Auto-check passed
  • Memory Corpus Ingest

    vellum-ai/vellum-assistant

    Ingest a large dataset into memory as a skimmed map. An agent skill from vellum-ai/vellum-assistant.

    1.4k GitHub stars~3k tokensUpdated yesterday
    Auto-check: notes
  • Plugin Builder

    vellum-ai/vellum-assistant

    A skill your agent uses when the user wants to build, scaffold, ship, or edit a Vellum plugin that bundles multiple surfaces (hooks, tools, skills, and more) into one installable package.

    1.4k GitHub stars~3.1k tokensUpdated yesterday
    Auto-check passed
  • Slack App Setup

    vellum-ai/vellum-assistant

    Connect a Slack app to the Vellum Assistant via Socket Mode.

    1.4k GitHub stars~2.5k tokensUpdated yesterday
    Auto-check: warnings

Questions about Elevenlabs Voice

What does Elevenlabs Voice do?

Select and tune an ElevenLabs TTS voice - curated voice list, custom/cloned voices via API key, and tuning parameters. Elevenlabs Voice is an agent skill from vellum-ai/vellum-assistant.

When should I use Elevenlabs Voice?

Elevenlabs Voice fits situations like: tasks that involve Text to speech and voice.

How do I install Elevenlabs Voice in Claude Code?

Run `npx skills add vellum-ai/vellum-assistant --skill elevenlabs-voice -a claude-code`. Or copy the skill folder (skills/elevenlabs-voice in vellum-ai/vellum-assistant) into .claude/skills/elevenlabs-voice in your project. Claude Code loads it when a task matches its description.

How do I install Elevenlabs Voice in Codex?

Run `npx skills add vellum-ai/vellum-assistant --skill elevenlabs-voice -a codex`. Or copy the skill folder (skills/elevenlabs-voice in vellum-ai/vellum-assistant) into .agents/skills/elevenlabs-voice in your project. Codex loads it when a task matches its description.

Can I use Elevenlabs Voice in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add vellum-ai/vellum-assistant --skill elevenlabs-voice -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/elevenlabs-voice, .gemini/skills/elevenlabs-voice, .github/skills/elevenlabs-voice and .opencode/skills/elevenlabs-voice in your project.

What does Elevenlabs Voice need to run?

Going by SKILL.md and its folder, Elevenlabs Voice needs the command-line tools its instructions call (curl and python3). Our summary lists: Python 3. Compatibility (from SKILL.md): Designed for Vellum personal assistants.

Does Elevenlabs Voice access the network?

SKILL.md names 2 domains. In commands or code: api.elevenlabs.io; the agent is likely to contact it when it follows the instructions. As links in the text: elevenlabs.io. This is read from the text; nothing was executed.

Is Elevenlabs Voice safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Elevenlabs Voice use?

Elevenlabs Voice is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Elevenlabs Voice use?

About 2.3k tokens (SKILL.md is roughly 9.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Elevenlabs Voice?

Skills that share tags, products or a category with Elevenlabs Voice: 9Router Text to Speech (decolua/9router, 31k stars), Keirouter Tts (mydisha/keirouter, 147 stars), Voice AI Development (davila7/claude-code-templates, 33k stars) and Music (tadaspetra/loop, 296 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Elevenlabs Voice?

vellum-ai (a GitHub organization) maintains it in vellum-ai/vellum-assistant, which has 1,408 GitHub stars. The repository holds 108 skills in this directory. The repository was last updated on October 9, 2026.

Source: vellum-ai/vellum-assistant on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.