Agent skill

Text To Speech

by dmccreary in dmccreary/ibook-skills

Convert text to speech using ElevenLabs voice AI. An agent skill from dmccreary/ibook-skills.

CC-BY-NC-4.0Auto-check passedMedia & Creative

Install Text To Speech

skills CLI
$ npx skills add dmccreary/ibook-skills --skill text-to-speech -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install dmccreary/ibook-skills text-to-speech --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/dmccreary/ibook-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/archived/text-to-speech .claude/skills/text-to-speech && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
text-to-speech
GitHub stars
105
Used in
2 other repos
Token cost
~1.8k tokens
SKILL.md length
419 words
Files
4 (incl. references)
Skills in repo
32
Repo updated
First seen
Licence
CC-BY-NC-4.0

At a glance

Convert text to speech using ElevenLabs voice AI. An agent skill from dmccreary/ibook-skills.

  • Generating audio from text
  • SKILL.md covers Quick Start, Models, Voice IDs and Voice Settings, plus 8 more sections
  • Calls curl; reaches api.elevenlabs.io; needs ELEVENLABS_API_KEY
  • Creating voiceovers

What it does

Text To Speech is an agent skill from dmccreary/ibook-skills. Convert text to speech using ElevenLabs voice AI. Use when generating audio from text, creating voiceovers, building voice apps, or synthesizing speech in 70+ languages.

Its SKILL.md is about 1.8k tokens, which your agent loads only when the skill is triggered. The skill folder holds 4 other files, including reference files (for example `references/installation.md`, `references/streaming.md` and `references/voice-settings.md`). Compatibility notes: Requires internet access and an ElevenLabs API key (ELEVENLABSAPIKEY).

It sits in Media & Creative, covering Text to speech and voice. It works with ElevenLabs.

When your agent uses it

  • Generating audio from text
  • Creating voiceovers
  • Building voice apps
  • Synthesizing speech in 70+ languages

Example prompts

  • “/text-to-speech”

Requirements

  • Python 3
  • A credential in ELEVENLABS_API_KEY
  • Compatibility (from SKILL.md): Requires internet access and an ElevenLabs API key (ELEVENLABS_API_KEY).

What it can do on your machine

Read from SKILL.md and the folder at commit 4c58491. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • curl

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • api.elevenlabs.io

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • ELEVENLABS_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

  • Compatibility

    Requires internet access and an ElevenLabs API key (ELEVENLABS_API_KEY).

    From compatibility in the SKILL.md frontmatter.

Context cost

Text To Speech loads about 1.8k tokens when it runs, and up to ~5.3k if it reads all its reference files. Until then it costs about 46 tokens; SKILL.md has 419 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~46
When it runs · the whole SKILL.md, loaded when a task matches
~1.8k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~5.3k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (CC-BY-NC-4.0) doesn't allow us to republish the file, so here is its outline and opening line. It has 419 words (~1,826 tokens).

“Generate natural speech from text - supports 70+ languages, multiple models for quality vs latency tradeoffs.”

— opening of SKILL.md by dmccreary, CC-BY-NC-4.0
name
text-to-speech
compatibility
Requires internet access and an ElevenLabs API key (ELEVENLABS_API_KEY).
model
sonnet
license
Creative Commons Attribution-NonCommercial 4.0 International (CC BY-NC 4.0)

Read the full SKILL.md on GitHub

Files

SKILL.md and 3 other files (references) in skills/archived/text-to-speech of dmccreary/ibook-skills.

  • SKILL.md
  • references/installation.md
  • references/streaming.md
  • references/voice-settings.md

Open the folder on GitHubat commit 4c58491

Used in 2 other repositories

We found 2 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 2 other GitHub owners. This page covers the copy in dmccreary/ibook-skills, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Text To Speech next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Text To Speech compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Text To Speech this skilldmccreary/ibook-skills1052 repos~1.8kAutomated safety check: PassCC-BY-NC-4.0
Musictadaspetra/loop2962 repos~827Automated safety check: PassMIT
Sound Effectstadaspetra/loop2962 repos~1.1kAutomated safety check: PassMIT
Elevenlabs Transcribeqdhenry/Claude-Command-Suite1.3k—~1.5kAutomated safety check: NotesNone
Dubbingelevenlabs/skills482—~3.3kAutomated safety check: PassMIT
Sagtrpc-group/trpc-agent-go1.9k15 repos~574Automated safety check: PassApache-2.0

Similar skills

  • Music

    tadaspetra/loop

    Generate music using ElevenLabs Music API. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 2 repos~827 tokens
    Media & CreativeAuto-check passed
  • Sound Effects

    tadaspetra/loop

    Generate sound effects from text descriptions using ElevenLabs.

    296 GitHub starsUsed in 2 repos~1.1k tokens
    Media & CreativeAuto-check passed
  • Elevenlabs Transcribe

    qdhenry/Claude-Command-Suite

    Transcribes audio/video files using ElevenLabs Scribe v2 API.

    1.3k GitHub stars~1.5k tokensUpdated 7 mo ago
    Media & CreativeAuto-check: notes
  • Dubbing

    elevenlabs/skills

    Dub audio and video into other languages using the ElevenLabs Dubbing API (dubbingv2), preserving the original speakers' voices.

    482 GitHub stars~3.3k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Sag

    trpc-group/trpc-agent-go

    ElevenLabs text-to-speech with mac-style say UX. An agent skill from trpc-group/trpc-agent-go.

    1.9k GitHub starsUsed in 15 repos~574 tokens
    Media & CreativeAuto-check passed
  • Music

    bozhouDev/video-skills-toolkit

    Generate music using ElevenLabs Music API. An agent skill from bozhouDev/video-skills-toolkit.

    150 GitHub stars~3.6k tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed

More from dmccreary/ibook-skills

All 32 skills in this repo
  • Book Publisher

    dmccreary/ibook-skills

    Publishes and promotes a finished intelligent textbook - GitHub README with badges and site statistics, LinkedIn announcement posts, LinkedIn carousel document posts (PPTX/PDF slideshows), and…

    105 GitHub stars~1.3k tokensUpdated today
    Auto-check passed
  • Linkedin Carousel Generator

    dmccreary/ibook-skills

    Generates a 13-slide LinkedIn carousel (a "document post" — PPTX/PDF) that showcases an intelligent textbook's key features, with real screenshots, mascot art, and metrics pulled from the project.

    105 GitHub stars~3.7k tokensUpdated today
    Auto-check passed
  • Add Xapi Events To Microsim

    dmccreary/ibook-skills

    A skill your agent uses when the user wants an existing MicroSim or chapter quiz to record what students do with it.

    105 GitHub stars~5.6k tokensUpdated today
    Auto-check passed
  • Diagram Reports Generator

    dmccreary/ibook-skills

    Generates a status report of all diagrams and MicroSims across an intelligent textbook's chapters, including difficulty, Bloom's level, and UI complexity.

    105 GitHub starsUsed in 1 repo~1.9k tokens
    Auto-check passed
  • DOCX To Web Publisher

    dmccreary/ibook-skills

    Converts .docx files (papers, briefs, reports) into styled React component pages in a Next.js content-catalog site.

    105 GitHub stars~2.2k tokensUpdated today
    Auto-check passed
  • Book Media Generator

    dmccreary/ibook-skills

    Generates media for intelligent textbooks - slide decks and presentations (MARP web decks in docs/slides/ or PowerPoint .pptx lecture downloads), illustrated stories and graphic novels…

    105 GitHub stars~1.7k tokensUpdated today
    Auto-check passed

Works with

Questions about Text To Speech

What does Text To Speech do?

Convert text to speech using ElevenLabs voice AI. An agent skill from dmccreary/ibook-skills. Text To Speech is an agent skill from dmccreary/ibook-skills. Convert text to speech using ElevenLabs voice AI.

When should I use Text To Speech?

Text To Speech fits situations like: generating audio from text; creating voiceovers; building voice apps; synthesizing speech in 70+ languages.

How do I install Text To Speech in Claude Code?

Run `npx skills add dmccreary/ibook-skills --skill text-to-speech -a claude-code`. Or copy the skill folder (skills/archived/text-to-speech in dmccreary/ibook-skills) into .claude/skills/text-to-speech in your project. Claude Code loads it when a task matches its description.

How do I install Text To Speech in Codex?

Run `npx skills add dmccreary/ibook-skills --skill text-to-speech -a codex`. Or copy the skill folder (skills/archived/text-to-speech in dmccreary/ibook-skills) into .agents/skills/text-to-speech in your project. Codex loads it when a task matches its description.

Can I use Text To Speech in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add dmccreary/ibook-skills --skill text-to-speech -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/text-to-speech, .gemini/skills/text-to-speech, .github/skills/text-to-speech and .opencode/skills/text-to-speech in your project.

What does Text To Speech need to run?

Going by SKILL.md and its folder, Text To Speech needs the command-line tools its instructions call (curl) and credentials named ELEVENLABS_API_KEY. Our summary lists: Python 3; A credential in ELEVENLABS_API_KEY. Compatibility (from SKILL.md): Requires internet access and an ElevenLabs API key (ELEVENLABS_API_KEY)..

Does Text To Speech access the network?

SKILL.md names 1 domain. In commands or code: api.elevenlabs.io; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.

Is Text To Speech safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Text To Speech use?

Text To Speech is published under the CC-BY-NC-4.0 licence (declared in SKILL.md).

How many tokens does Text To Speech use?

About 1.8k tokens (SKILL.md is roughly 7.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 3.4k tokens, read only when the agent opens those files.

What are the alternatives to Text To Speech?

Skills that share tags, products or a category with Text To Speech: Music (tadaspetra/loop, 296 stars), Sound Effects (tadaspetra/loop, 296 stars), Elevenlabs Transcribe (qdhenry/Claude-Command-Suite, 1.3k stars) and Dubbing (elevenlabs/skills, 482 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Text To Speech?

dmccreary (a GitHub user) maintains it in dmccreary/ibook-skills, which has 105 GitHub stars. The repository holds 32 skills in this directory. The repository was last updated on October 10, 2026.

Source: dmccreary/ibook-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.