Official agent skill

Add Read Aloud

by cursor in cursor/plugins

Wires xAI's Grok voice synthesis into an app you already have, so assistant replies get a play button, optional automatic playback, or narration of any text.

OfficialNo licenceAuto-check passedAI & LLM Engineering

Install Add Read Aloud

skills CLI
$ npx skills add cursor/plugins --skill add-read-aloud -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install cursor/plugins add-read-aloud --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/cursor/plugins.git skills-src && mkdir -p .claude/skills && cp -r skills-src/grok-voice/skills/add-read-aloud .claude/skills/add-read-aloud && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
add-read-aloud
GitHub stars
10k
Token cost
~3.6k tokens
SKILL.md length
1,291 words
Files
1
Skills in repo
99
Repo updated
First seen
Licence
None found

At a glance

Wires xAI's Grok voice synthesis into an app you already have, so assistant replies get a play button, optional automatic playback, or narration of any text.

  • Works in 5 steps: Map the app → Prepare the text → Batch path (default) → …
  • Adding a read-aloud button to the assistant replies in a chat app
  • SKILL.md covers Docs, Pick the path, Auth and Steps, plus 1 more section
  • Calls curl; reaches api.x.ai; needs XAI_API_KEY

What it does

Instead of changing the IDE, the skill wires the app itself. It begins by mapping where assistant messages render, where per-message actions such as copy live, how the reply stream ends, and which server framework and package manager are in use. It then adds a ghost speaker button to the message action row, showing a spinner while loading and a stop square while playing, with one utterance at a time and the button appearing only once the reply has finished streaming.

Two integration paths are described. Batch is the default: one POST request to the xAI /v1/tts endpoint returns one MP3 that can be cached and keeps the key on the server. Streaming over a websocket suits audio that must start while the model is still writing, barge-in, or text over 15,000 characters, and browsers reach it through a backend relay. The XAI_API_KEY goes out as a bearer token from the server only and must never end up in a client bundle or in chat.

The speaker icon is reserved for read aloud. A waveform signals voice mode, handled by /add-voice, and a microphone signals dictation, handled by /add-dictation.

When your agent uses it

  • Adding a read-aloud button to the assistant replies in a chat app
  • Auto-speaking new replies or narrating a page with text-to-speech
  • Generating audio files from text, such as IVR prompts or narration

Example prompts

  • “/add-read-aloud for our React chat app, with a speaker button on each assistant message.”
  • “Make the assistant speak its replies automatically once they finish streaming.”
  • “Add narration to the docs page with Grok text-to-speech and cache the MP3 on the server.”

Requirements

  • An xAI API key (XAI_API_KEY) kept on the server
  • An existing app that renders assistant replies

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. Map the app
  2. Prepare the text
  3. Batch path (default)
  4. Streaming path
  5. Options (JSON fields for batch, query params for streaming)

What it can do on your machine

Read from SKILL.md and the folder at commit ccb5507. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • curl

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • api.x.ai

    Also links to:

    • docs.x.ai

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • XAI_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Add Read Aloud loads about 3.6k tokens when it runs. Until then it costs about 77 tokens; SKILL.md has 1,291 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~77
When it runs · the whole SKILL.md, loaded when a task matches
~3.6k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 1,291 words (~3,600 tokens).

“Add Grok Text to Speech to an existing app: a speaker button on assistant replies, auto-speak, or narration of any text. Run on /add-read-aloud, typed Read aloud, or clear “speak this” / “TTS” intent. Cursor has no speaker; wire the…”

— opening of SKILL.md by cursor
name
add-read-aloud

Read the full SKILL.md on GitHub

Files

Just SKILL.md in grok-voice/skills/add-read-aloud of cursor/plugins.

Open the folder on GitHubat commit ccb5507

Compare with similar skills

Add Read Aloud next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Add Read Aloud compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Add Read Aloud this skillcursor/plugins10k—~3.6kAutomated safety check: PassNone
Elevenlabs Agentsjezweb/claude-skills1.1k1 repos~3.3kAutomated safety check: PassMIT
PiDeck Usage Probe Helperayuayue/PiDeck1k—~1.4kAutomated safety check: PassMIT
Agentstadaspetra/loop2961 repos~2.5kAutomated safety check: PassMIT
Lilly Community Researchssaaffaakk/Lilly171—~1.4kAutomated safety check: PassMIT
Xybrid Initxybrid-ai/xybrid466—~3kAutomated safety check: PassApache-2.0

Similar skills

  • Elevenlabs Agents

    jezweb/claude-skills

    Build conversational AI voice agents on the ElevenLabs platform.

    1.1k GitHub starsUsed in 1 repo~3.3k tokens
    AI & LLM EngineeringAuto-check passed
  • Helps show a model provider's usage, balance or quota in PiDeck: checks built-in support, points to the dialog templates, or writes a custom probe entry.

    1k GitHub stars~1.4k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Agents

    tadaspetra/loop

    Build voice AI agents with ElevenLabs. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 1 repo~2.5k tokens
    AI & LLM EngineeringAuto-check passed
  • Lilly Community Research

    ssaaffaakk/Lilly

    Lilly community-research skill. An agent skill from ssaaffaakk/Lilly.

    171 GitHub stars~1.4k tokensUpdated yesterday
    AI & LLM EngineeringAuto-check passed
  • Xybrid Init

    xybrid-ai/xybrid

    Generate model metadata for an ML model so it works with xybrid.

    466 GitHub stars~3k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Deepgram Python Management API

    deepgram/deepgram-python-sdk

    Guides Python code that calls the Deepgram Management APIs to administer projects, keys, members, usage, billing and stored Voice Agent configurations.

    469 GitHub stars~2.3k tokensUpdated today
    AI & LLM EngineeringAuto-check passed

More from cursor/plugins

All 99 skills in this repo
  • Official

    Keeps a TSV decision log for long or unattended agent runs, one row per decision with what, why, evidence and result, so a reviewer can check the work later.

    10k GitHub starsUsed in 9 repos~1.6k tokens
    Auto-check passed
  • Official

    Digs into why code is shaped the way it is by checking git history, pull requests and connected tools in parallel, then reporting a cited read on the tradeoffs.

    10k GitHub starsUsed in 9 repos~2.6k tokens
    Auto-check passed
  • Official

    Starts three parallel reviewer subagents over the current conversation transcript, then turns their findings into concrete edits to existing skills.

    10k GitHub starsUsed in 5 repos~1.2k tokens
    Auto-check passed
  • Official

    Applies four layers of technical-writing rules to docs, RFCs, readmes, PR descriptions and commit messages so a tired engineer follows them on the first read.

    10k GitHub starsUsed in 10 repos~2.4k tokens
    Auto-check passed
  • Advisor Mode

    cursor/plugins

    Official

    Adds a second, stronger model that the main agent consults before major decisions, when stuck and before finishing, controlled by /advisor commands.

    10k GitHub stars~2.6k tokensUpdated today
    Auto-check: notes
  • Official

    Prepare PRs for review by cleaning noisy history, improving PR descriptions, and adding reviewer guidance without changing code behavior.

    10k GitHub starsUsed in 3 repos~569 tokens
    Auto-check passed

Works with

Questions about Add Read Aloud

What does Add Read Aloud do?

Wires xAI's Grok voice synthesis into an app you already have, so assistant replies get a play button, optional automatic playback, or narration of any text. Instead of changing the IDE, the skill wires the app itself. It begins by mapping where assistant messages render, where per-message actions such as copy live, how the reply stream ends, and which server framework and package manager are in use.

When should I use Add Read Aloud?

Add Read Aloud fits situations like: adding a read-aloud button to the assistant replies in a chat app; auto-speaking new replies or narrating a page with text-to-speech; generating audio files from text, such as IVR prompts or narration.

How do I install Add Read Aloud in Claude Code?

Run `npx skills add cursor/plugins --skill add-read-aloud -a claude-code`. Or copy the skill folder (grok-voice/skills/add-read-aloud in cursor/plugins) into .claude/skills/add-read-aloud in your project. Claude Code loads it when a task matches its description.

How do I install Add Read Aloud in Codex?

Run `npx skills add cursor/plugins --skill add-read-aloud -a codex`. Or copy the skill folder (grok-voice/skills/add-read-aloud in cursor/plugins) into .agents/skills/add-read-aloud in your project. Codex loads it when a task matches its description.

Can I use Add Read Aloud in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add cursor/plugins --skill add-read-aloud -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/add-read-aloud, .gemini/skills/add-read-aloud, .github/skills/add-read-aloud and .opencode/skills/add-read-aloud in your project.

What does Add Read Aloud need to run?

Going by SKILL.md and its folder, Add Read Aloud needs the command-line tools its instructions call (curl) and credentials named XAI_API_KEY. Our summary lists: An xAI API key (XAI_API_KEY) kept on the server; An existing app that renders assistant replies.

Does Add Read Aloud access the network?

SKILL.md names 2 domains. In commands or code: api.x.ai; the agent is likely to contact it when it follows the instructions. As links in the text: docs.x.ai. This is read from the text; nothing was executed.

Is Add Read Aloud safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Add Read Aloud use?

No licence was found for Add Read Aloud or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Add Read Aloud use?

About 3.6k tokens (SKILL.md is roughly 14k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Add Read Aloud?

Skills that share tags, products or a category with Add Read Aloud: Elevenlabs Agents (jezweb/claude-skills, 1.1k stars), PiDeck Usage Probe Helper (ayuayue/PiDeck, 1k stars), Agents (tadaspetra/loop, 296 stars) and Lilly Community Research (ssaaffaakk/Lilly, 171 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Add Read Aloud?

cursor (a GitHub organization, an official publisher) maintains it in cursor/plugins, which has 10,278 GitHub stars. The repository holds 99 skills in this directory. The repository was last updated on October 8, 2026.

Source: cursor/plugins on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.