Agent skill

9Router Text to Speech

by decolua in decolua/9router

Turns text into spoken audio through a 9Router server, choosing among voices from OpenAI, ElevenLabs, Deepgram, Edge TTS and other providers.

MITAuto-check passedMedia & Creative

Install 9Router Text to Speech

skills CLI
$ npx skills add decolua/9router --skill 9router-tts -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install decolua/9router 9router-tts --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/decolua/9router.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/9router-tts .claude/skills/9router-tts && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
9router-tts
GitHub stars
31k
Token cost
~765 tokens
SKILL.md length
156 words
Files
1
Skills in repo
9
Repo updated
First seen
Licence
MIT

At a glance

Turns text into spoken audio through a 9Router server, choosing among voices from OpenAI, ElevenLabs, Deepgram, Edge TTS and other providers.

  • Converting a script or article into an MP3 voiceover
  • SKILL.md covers Discover, Endpoint, Examples and Response shape, plus 1 more section
  • Calls curl and jq; needs NINEROUTER_KEY
  • Reading a block of text aloud with a chosen provider voice

What it does

The agent calls a running 9Router instance to produce speech. It first lists the available voices from the `/v1/models/tts` endpoint, then posts the text to `/v1/audio/speech` with the chosen voice ID in the `model` field. By default the reply is raw MP3 bytes, and adding `?response_format=json` returns base64 audio together with its format.

Voice naming differs by provider: OpenAI takes a model and voice pair, ElevenLabs a model ID and voice ID, and Deepgram uses token auth. Edge TTS and Google TTS need no credentials, and a local-device option uses the operating system voice and requires `ffmpeg`. The skill gives curl and JavaScript examples that save the audio to a file.

When your agent uses it

  • Converting a script or article into an MP3 voiceover
  • Reading a block of text aloud with a chosen provider voice
  • Listing which text-to-speech voices a 9Router server offers

Example prompts

  • “Turn the intro in ./script/intro.txt into an MP3 voiceover and save it as intro.mp3.”
  • “List the text-to-speech voices my 9Router server offers for Vietnamese.”
  • “Read this paragraph aloud with an OpenAI voice and give me the audio file.”

Requirements

  • A running 9Router server, with its address in `NINEROUTER_URL`
  • `NINEROUTER_KEY` when the server has authentication enabled
  • `ffmpeg` for the local-device voice option

What it can do on your machine

Read from SKILL.md and the folder at commit ce4460e. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • curl
    • jq

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use curl, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • NINEROUTER_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

9Router Text to Speech loads about 765 tokens when it runs. Until then it costs about 64 tokens; SKILL.md has 156 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~64
When it runs · the whole SKILL.md, loaded when a task matches
~765

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from decolua/9router at commit ce4460e, republished under its MIT licence (© decolua). 156 words, ~765 tokens.

Download SKILL.mdSave it as .claude/skills/9router-tts/SKILL.md (or your agent's skills folder).
name
9router-tts
description
Text-to-speech via 9Router /v1/audio/speech using OpenAI / ElevenLabs / Deepgram / Edge TTS / Google TTS / Hyperbolic / Inworld voices. Use when the user wants to convert text to speech, generate audio, voiceover, narrate, or read text aloud.

9Router — Text-to-Speech

Requires NINEROUTER_URL (and NINEROUTER_KEY if auth enabled). See https://raw.githubusercontent.com/decolua/9router/refs/heads/master/skills/9router/SKILL.md for setup.

Discover

bash
# 1) List models
curl $NINEROUTER_URL/v1/models/tts | jq '.data[].id'
# 2) Per-model metadata (params, voicesUrl if voice-by-id)
curl "$NINEROUTER_URL/v1/models/info?id=el/eleven_multilingual_v2"
# 3) List voices (elevenlabs, edge-tts, deepgram, inworld, local-device). Optional ?lang=vi
curl "$NINEROUTER_URL/v1/audio/voices?provider=edge-tts&lang=vi" | jq '.data[].model'

model field in /v1/audio/speech = voice ID directly (e.g. edge-tts/vi-VN-HoaiMyNeural, el/<voice_id>, or openai/tts-1 model+default voice).

Endpoint

POST $NINEROUTER_URL/v1/audio/speech

FieldRequiredNotes
modelyesvoice ID from /v1/models/tts
inputyestext to speak

Query ?response_format=mp3 (default, raw bytes) or ?response_format=json ({audio: base64, format}).

Examples

Save MP3:

bash
curl -X POST "$NINEROUTER_URL/v1/audio/speech" \
  -H "Authorization: Bearer $NINEROUTER_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"openai/tts-1","input":"Hello world"}' \
  --output speech.mp3

JS (save file):

js
import { writeFile } from "node:fs/promises";
const r = await fetch(`${process.env.NINEROUTER_URL}/v1/audio/speech`, {
  method: "POST",
  headers: { "Authorization": `Bearer ${process.env.NINEROUTER_KEY}`, "Content-Type": "application/json" },
  body: JSON.stringify({ model: "el/eleven_multilingual_v2", input: "Xin chào" }),
});
await writeFile("speech.mp3", Buffer.from(await r.arrayBuffer()));

Response shape

Default → raw audio bytes (Content-Type audio/mp3).

?response_format=json:

json
{ "audio": "SUQzBAAAA...", "format": "mp3" }

Provider quirks (model format)

Providermodel formatNotes
openaitts-1/alloy (model/voice) or just voiceDefault model gpt-4o-mini-tts
elevenlabs<model_id>/<voice_id> or <voice_id>Default model eleven_flash_v2_5; list voices in Dashboard
openrouteropenai/gpt-4o-mini-tts/alloyStreamed via chat-completions audio modality
edge-ttsvoice id e.g. vi-VN-HoaiMyNeuralnoAuth; default vi-VN-HoaiMyNeural
google-ttslanguage code e.g. en, vinoAuth
local-deviceOS voice name (say -v ? / SAPI)noAuth; needs ffmpeg
deepgramaura-asteria-en etcToken auth
nvidia, inworld, cartesia, playhtmodel/voiceProvider-specific auth header
coqui, tortoisespeaker / voice idLocalhost noAuth
hyperbolicmodel idBody = {text} only

© decolua, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/9router-tts of decolua/9router.

Open the folder on GitHubat commit ce4460e

Compare with similar skills

9Router Text to Speech next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

9Router Text to Speech compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
9Router Text to Speech this skilldecolua/9router31k—~765Automated safety check: PassMIT
Keirouter Ttsmydisha/keirouter147—~599Automated safety check: PassMIT
Voice AI Developmentdavila7/claude-code-templates33k5 repos~2.1kAutomated safety check: PassMIT
Video Translatorshang-zhu/violin1.1k—~1kAutomated safety check: NotesMIT
Super Video MakerBomx/super-video-maker-skill310—~11kAutomated safety check: NotesNone
Motion Videobestagentkits/motion-video-skill118—~1.5kAutomated safety check: PassMIT

Similar skills

  • Keirouter Tts

    mydisha/keirouter

    Text-to-speech via KeiRouter /v1/audio/speech using OpenAI / ElevenLabs / Deepgram / Edge TTS / Google TTS / Inworld voices.

    147 GitHub stars~599 tokensUpdated 29 days ago
    Media & CreativeAuto-check passed
  • Voice AI Development

    davila7/claude-code-templates

    Expert in building voice AI applications - from real-time voice agents to voice-enabled apps.

    33k GitHub starsUsed in 5 repos~2.1k tokens
    Media & CreativeAuto-check passed
  • Video Translator

    shang-zhu/violin

    Dub a video into another language and generate subtitles using the default Together + Cartesia stack.

    1.1k GitHub stars~1k tokensUpdated 1 mo ago
    Media & CreativeAuto-check: notes
  • Super Video Maker

    Bomx/super-video-maker-skill

    End-to-end AI video production skill for agentic frameworks.

    310 GitHub stars~11k tokensUpdated 2 mo ago
    Media & CreativeAuto-check: notes
  • Motion Video

    bestagentkits/motion-video-skill

    Produce beat-synced 1080p motion-graphic videos in HyperFrames (HTML + GSAP) with an AI voice-over, Vietnamese karaoke captions, SFX and generated music, in one of two proven styles (glass keynote…

    118 GitHub stars~1.5k tokensUpdated 11 days ago
    Media & CreativeAuto-check passed
  • Local AI Use

    amd/skills

    Makes this agent generate images, transcribe audio, and synthesize speech on the user's own machine through a local Lemonade Server instead of a paid cloud API.

    408 GitHub stars~5k tokensUpdated yesterday
    Media & CreativeAuto-check: notes

More from decolua/9router

All 9 skills in this repo
  • Sets up access to the 9Router AI gateway, an OpenAI-compatible REST endpoint for chat, images, speech, embeddings, web search and web fetch, and indexes its capability skills.

    31k GitHub stars~744 tokensUpdated 2 days ago
    Auto-check passed
  • Sends chat and code-generation requests through a 9Router gateway using OpenAI or Anthropic message formats, with streaming and auto-fallback combos.

    31k GitHub stars~635 tokensUpdated 2 days ago
    Auto-check passed
  • Embeddings via 9Router

    decolua/9router

    Generates vector embeddings through the 9Router /v1/embeddings endpoint, using models from providers such as OpenAI, Gemini, Mistral and Voyage for RAG and semantic search.

    31k GitHub stars~604 tokensUpdated 2 days ago
    Auto-check passed
  • Generates images through a 9Router gateway's image endpoint, with model discovery, the request fields and per-provider quirks for OpenAI, Gemini, MiniMax and others.

    31k GitHub stars~830 tokensUpdated 2 days ago
    Auto-check passed
  • 9Router Speech-to-Text

    decolua/9router

    Transcribes audio files into text or subtitles through 9Router's Whisper-compatible endpoint, using models from OpenAI, Groq, Gemini, Deepgram and others.

    31k GitHub stars~914 tokensUpdated 2 days ago
    Auto-check passed
  • Submits text-to-video or image-to-video jobs to xAI Grok Imagine through 9Router, then polls the job and downloads the finished MP4.

    31k GitHub stars~992 tokensUpdated 2 days ago
    Auto-check passed

Questions about 9Router Text to Speech

What does 9Router Text to Speech do?

Turns text into spoken audio through a 9Router server, choosing among voices from OpenAI, ElevenLabs, Deepgram, Edge TTS and other providers. The agent calls a running 9Router instance to produce speech. It first lists the available voices from the `/v1/models/tts` endpoint, then posts the text to `/v1/audio/speech` with the chosen voice ID in the `model` field.

When should I use 9Router Text to Speech?

9Router Text to Speech fits situations like: converting a script or article into an MP3 voiceover; reading a block of text aloud with a chosen provider voice; listing which text-to-speech voices a 9Router server offers.

How do I install 9Router Text to Speech in Claude Code?

Run `npx skills add decolua/9router --skill 9router-tts -a claude-code`. Or copy the skill folder (skills/9router-tts in decolua/9router) into .claude/skills/9router-tts in your project. Claude Code loads it when a task matches its description.

How do I install 9Router Text to Speech in Codex?

Run `npx skills add decolua/9router --skill 9router-tts -a codex`. Or copy the skill folder (skills/9router-tts in decolua/9router) into .agents/skills/9router-tts in your project. Codex loads it when a task matches its description.

Can I use 9Router Text to Speech in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add decolua/9router --skill 9router-tts -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/9router-tts, .gemini/skills/9router-tts, .github/skills/9router-tts and .opencode/skills/9router-tts in your project.

What does 9Router Text to Speech need to run?

Going by SKILL.md and its folder, 9Router Text to Speech needs the command-line tools its instructions call (curl and jq) and credentials named NINEROUTER_KEY. Our summary lists: A running 9Router server, with its address in `NINEROUTER_URL`; `NINEROUTER_KEY` when the server has authentication enabled; `ffmpeg` for the local-device voice option.

Does 9Router Text to Speech access the network?

SKILL.md contains no URLs. Its commands use curl, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is 9Router Text to Speech safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does 9Router Text to Speech use?

9Router Text to Speech is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does 9Router Text to Speech use?

About 765 tokens (SKILL.md is roughly 3.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to 9Router Text to Speech?

Skills that share tags, products or a category with 9Router Text to Speech: Keirouter Tts (mydisha/keirouter, 147 stars), Voice AI Development (davila7/claude-code-templates, 33k stars), Video Translator (shang-zhu/violin, 1.1k stars) and Super Video Maker (Bomx/super-video-maker-skill, 310 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains 9Router Text to Speech?

decolua (a GitHub user) maintains it in decolua/9router, which has 30,536 GitHub stars. The repository holds 9 skills in this directory. The repository was last updated on October 8, 2026.

Source: decolua/9router on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.