9Router Text to Speech
decolua/9router
Turns text into spoken audio through a 9Router server, choosing among voices from OpenAI, ElevenLabs, Deepgram, Edge TTS and other providers.
Complete voice configuration in chat - desktop Talk and PTT shortcuts, microphone permissions, ElevenLabs/Deepgram TTS, and troubleshooting
$ npx skills add vellum-ai/vellum-assistant --skill voice-setup -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install vellum-ai/vellum-assistant voice-setup --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/vellum-ai/vellum-assistant.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/voice-setup .claude/skills/voice-setup && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "voice-setup" agent skill from https://github.com/vellum-ai/vellum-assistant/tree/main/skills/voice-setup into .claude/skills/voice-setup/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-setup", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/vellum-ai/vellum-assistant/tree/main/skills/voice-setupType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add vellum-ai/vellum-assistant --skill voice-setup -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install vellum-ai/vellum-assistant voice-setup --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/vellum-ai/vellum-assistant.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/voice-setup .agents/skills/voice-setup && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "voice-setup" agent skill from https://github.com/vellum-ai/vellum-assistant/tree/main/skills/voice-setup into .agents/skills/voice-setup/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-setup", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add vellum-ai/vellum-assistant --skill voice-setup -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install vellum-ai/vellum-assistant voice-setup --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/vellum-ai/vellum-assistant.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/voice-setup .cursor/skills/voice-setup && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "voice-setup" agent skill from https://github.com/vellum-ai/vellum-assistant/tree/main/skills/voice-setup into .cursor/skills/voice-setup/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-setup", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/vellum-ai/vellum-assistant.git --path skills/voice-setup--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add vellum-ai/vellum-assistant --skill voice-setup -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install vellum-ai/vellum-assistant voice-setup --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/vellum-ai/vellum-assistant.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/voice-setup .gemini/skills/voice-setup && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "voice-setup" agent skill from https://github.com/vellum-ai/vellum-assistant/tree/main/skills/voice-setup into .gemini/skills/voice-setup/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-setup", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install vellum-ai/vellum-assistant voice-setupInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add vellum-ai/vellum-assistant --skill voice-setup -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/vellum-ai/vellum-assistant.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/voice-setup .github/skills/voice-setup && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "voice-setup" agent skill from https://github.com/vellum-ai/vellum-assistant/tree/main/skills/voice-setup into .github/skills/voice-setup/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-setup", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add vellum-ai/vellum-assistant --skill voice-setup -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install vellum-ai/vellum-assistant voice-setup --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/vellum-ai/vellum-assistant.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/voice-setup .opencode/skills/voice-setup && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "voice-setup" agent skill from https://github.com/vellum-ai/vellum-assistant/tree/main/skills/voice-setup into .opencode/skills/voice-setup/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "voice-setup", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
voice-setupComplete voice configuration in chat - desktop Talk and PTT shortcuts, microphone permissions, ElevenLabs/Deepgram TTS, and troubleshooting
Voice Setup is an agent skill from vellum-ai/vellum-assistant. Complete voice configuration in chat - desktop Talk and PTT shortcuts, microphone permissions, ElevenLabs/Deepgram TTS, and troubleshooting
Its SKILL.md is about 2.8k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including assets. Compatibility notes: Designed for Vellum personal assistants
It sits in Media & Creative, covering Text to speech and voice. It works with Deepgram, ElevenLabs and macOS. The repository describes itself as: An AI Assistant that’s easy to setup, does your work 24/7, knows your preferences and gets better over time. The licence is MIT.
4 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 844117a. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
No scripts in the folder and no shell commands in SKILL.md (its code samples are bash and powershell).
From the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Designed for Vellum personal assistants
From compatibility in the SKILL.md frontmatter.
Voice Setup loads about 2.8k tokens when it runs. Until then it costs about 38 tokens; SKILL.md has 1,451 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from vellum-ai/vellum-assistant at commit 844117a, republished under its MIT licence (© vellum-ai). 1,451 words, ~2,809 tokens.
.claude/skills/voice-setup/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.You are helping the user set up and troubleshoot voice features entirely within this conversation. Use the client_os: line in <turn_context> to choose the macOS or Windows instructions below. Do not give macOS commands or key names to a Windows user, or Windows guidance to a macOS user.
Before using a desktop tool, check client_os:
macos or windows, follow that platform's branch.web, ios, android, or absent, do not call open_system_settings or give desktop shortcut instructions. Explain that permissions and the Talk shortcut must be configured from the Mac or Windows desktop app. You can still complete provider, voice, language, and timeout configuration in the current conversation.voice_config_update changes shared voice settings such as the legacy macOS PTT activation key, conversation timeout, speech providers, and TTS voice ID.open_system_settings opens the correct macOS System Settings or Windows Settings privacy page. Call it only when client_os is macos or windows, and pass that value as platform.navigate_settings_tab opens Vellum settings. Use it for review, or when the desktop-owned Talk shortcut must be recorded in the app.assistant credentials prompt collects API keys securely for ElevenLabs or Deepgram.The desktop Talk shortcut is client-owned. On current desktop clients, configure it in the Voice settings shortcut control instead of treating voice_config_update setting="activation_key" as a global shortcut editor. Use voice_config_update for the shared settings it owns. Use activation_key only for the legacy macOS PTT activation setting.
Walk through each relevant section in order. Skip sections the user does not need, and ask before moving to the next section.
Microphone permission is not reported in <channel_capabilities>. Ask whether Vellum shows a microphone permission warning or whether the user already granted access.
If access is denied or the user is unsure:
open_system_settings with pane: "microphone" and the current platform.If the user confirms access is granted, continue without opening system settings.
On macOS, first determine whether the user means the current desktop Talk shortcut or the legacy PTT activation setting. Windows supports the desktop Talk shortcut only.
The Talk shortcut starts or ends a voice conversation.
Ask which behavior they want, then use navigate_settings_tab with tab: "Voice" so they can record the desktop-owned shortcut. Do not claim that voice_config_update changed this shortcut.
This setting is macOS-only. If a Mac user explicitly wants the legacy hold-to-talk setting, offer only values accepted by voice_config_update:
fnfn_shiftctrlnoneAfter the user chooses, call voice_config_update with setting: "activation_key" and the matching canonical value.
On Windows, do not offer or call the legacy activation setting. It has no Windows client consumer. Use the desktop Talk shortcut flow instead.
Ask whether the user wants high-quality text-to-speech voices through ElevenLabs or Deepgram. Standard TTS works without this optional setup.
The included ElevenLabs Voice and Deepgram Voice skills provide the provider-specific setup flow, including API key collection, voice selection, and tuning.
Check the active provider first with assistant config get services.tts.provider. voice_config_update writes the voice to the active provider, and each bring-your-own provider accepts its own voice IDs. If the preferred provider does not match the active provider, collect any required API key before switching:
voice_config_update setting="tts_provider" value="deepgram"The managed vellum provider accepts both supported ElevenLabs and Deepgram voice IDs, so it does not require a provider switch. Then follow the matching included voice skill.
The active provider's voice setting controls both in-app TTS and phone calls.
After setup:
navigate_settings_tab and tab: "Voice".Desktop Talk starts a live voice session. Its audio is transcribed through the assistant's configured speech-to-text provider over the live voice connection. The Windows native helper provides partials only for one-shot dictation from the microphone button. Ask which surface the user tested before troubleshooting missing text.
open_system_settings with pane: "speech_recognition" and the current platform.assistant config get services.stt.provider.voice_config_update setting="stt_provider" and collect any required credential securely before retrying.This path applies to the microphone button's one-shot dictation, not Desktop Talk.
open_system_settings.voice_config_update changes should apply immediately. Verify the persisted value with the relevant config command.For persistent issues, use the matching log path.
macOS:
log stream --predicate 'subsystem == "com.vellum.assistant"' --level debugLook for voice and speech categories.
Windows PowerShell:
$log = Get-ChildItem "$env:APPDATA\Vellum*\logs\vellum.log" |
Sort-Object LastWriteTime -Descending |
Select-Object -First 1
Get-Content $log.FullName -WaitFor Desktop Talk, look for live voice, WebSocket, and speech-to-text provider errors. For one-shot dictation, look for [win-helper], dictation, permission, and recognizer messages. Do not ask the user to share transcript contents from logs.
navigate_settings_tab for review and for the desktop-owned Talk shortcut, which must be recorded in the client.© vellum-ai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 1 other file (assets) in skills/voice-setup of vellum-ai/vellum-assistant.
Open the folder on GitHubat commit 844117a
Voice Setup next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Voice Setup this skillvellum-ai/vellum-assistant | 1.4k | — | ~2.8k | Automated safety check: Pass | MIT | |
| 9Router Text to Speechdecolua/9router | 30k | — | ~765 | Automated safety check: Pass | MIT | |
| Keirouter Ttsmydisha/keirouter | 147 | — | ~599 | Automated safety check: Pass | MIT | |
| Voice AI Developmentdavila7/claude-code-templates | 32k | 5 repos | ~2.1k | Automated safety check: Pass | MIT | |
| C Voicedaxaur/openpaw | 174 | — | ~401 | Automated safety check: Pass | MIT | |
| Musictadaspetra/loop | 296 | 2 repos | ~827 | Automated safety check: Pass | MIT |
decolua/9router
Turns text into spoken audio through a 9Router server, choosing among voices from OpenAI, ElevenLabs, Deepgram, Edge TTS and other providers.
mydisha/keirouter
Text-to-speech via KeiRouter /v1/audio/speech using OpenAI / ElevenLabs / Deepgram / Edge TTS / Google TTS / Inworld voices.
davila7/claude-code-templates
Expert in building voice AI applications - from real-time voice agents to voice-enabled apps.
daxaur/openpaw
Convert speech to text using sag (ElevenLabs STT) and synthesize speech using say (macOS built-in TTS).
tadaspetra/loop
Generate music using ElevenLabs Music API. An agent skill from tadaspetra/loop.
digitalsamba/claude-code-video-toolkit
Generates narration, sound effects and cloned voices through the ElevenLabs API, with model and setting choices tuned to the content's style.
vellum-ai/vellum-assistant
Create and configure a GitHub App so the assistant can push commits, open PRs, and comment under its own bot identity.
vellum-ai/vellum-assistant
Connect a Discord bot to the assistant via the Discord Gateway with guided application creation and intent configuration
vellum-ai/vellum-assistant
Create and configure a Sentry internal integration so the assistant can manage issues, alerts, and releases under its own identity
vellum-ai/vellum-assistant
Ingest a large dataset into memory as a skimmed map. An agent skill from vellum-ai/vellum-assistant.
vellum-ai/vellum-assistant
A skill your agent uses when the user wants to build, scaffold, ship, or edit a Vellum plugin that bundles multiple surfaces (hooks, tools, skills, and more) into one installable package.
vellum-ai/vellum-assistant
Connect a Slack app to the Vellum Assistant via Socket Mode.
Works with
Categories
Complete voice configuration in chat - desktop Talk and PTT shortcuts, microphone permissions, ElevenLabs/Deepgram TTS, and troubleshooting. Voice Setup is an agent skill from vellum-ai/vellum-assistant.
Voice Setup fits situations like: tasks that involve Text to speech and voice.
Run `npx skills add vellum-ai/vellum-assistant --skill voice-setup -a claude-code`. Or copy the skill folder (skills/voice-setup in vellum-ai/vellum-assistant) into .claude/skills/voice-setup in your project. Claude Code loads it when a task matches its description.
Run `npx skills add vellum-ai/vellum-assistant --skill voice-setup -a codex`. Or copy the skill folder (skills/voice-setup in vellum-ai/vellum-assistant) into .agents/skills/voice-setup in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add vellum-ai/vellum-assistant --skill voice-setup -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/voice-setup, .gemini/skills/voice-setup, .github/skills/voice-setup and .opencode/skills/voice-setup in your project.
SKILL.md names no scripts, command-line tools or credentials: Voice Setup is instructions for the agent only. Compatibility (from SKILL.md): Designed for Vellum personal assistants.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Voice Setup is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 2.8k tokens (SKILL.md is roughly 11k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Voice Setup: 9Router Text to Speech (decolua/9router, 30k stars), Keirouter Tts (mydisha/keirouter, 147 stars), Voice AI Development (davila7/claude-code-templates, 32k stars) and C Voice (daxaur/openpaw, 174 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
vellum-ai (a GitHub organization) maintains it in vellum-ai/vellum-assistant, which has 1,400 GitHub stars. The repository holds 108 skills in this directory. The repository was last updated on October 9, 2026.
Source: vellum-ai/vellum-assistant on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.