Agent skill

Read Aloud

by Ckokoski in Ckokoski/AuthorAgent

Text-to-speech read-aloud for manuscripts using free open-source Piper TTS with natural-sounding voices

MITAuto-check passedMedia & Creative

Install Read Aloud

skills CLI
$ npx skills add Ckokoski/AuthorAgent --skill read-aloud -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Ckokoski/AuthorAgent read-aloud --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Ckokoski/AuthorAgent.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/_archived/premium/read-aloud .claude/skills/read-aloud && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
read-aloud
GitHub stars
124
Token cost
~1.3k tokens
SKILL.md length
513 words
Files
1
Skills in repo
23
Repo updated
First seen
Licence
MIT

At a glance

Text-to-speech read-aloud for manuscripts using free open-source Piper TTS with natural-sounding voices

  • Works in 5 steps: Listen — Read Aloud plays your chapter → Flag — Mark spots that sound wrong… → Review — After playback, see all flagged… → …
  • Tasks that involve Text to speech and voice
  • SKILL.md covers Why Read Aloud?, Engine: Piper TTS, Read Modes and Ear-Edit Workflow, plus 4 more sections
  • Calls pip

What it does

Read Aloud is an agent skill from Ckokoski/AuthorAgent. Text-to-speech read-aloud for manuscripts using free open-source Piper TTS with natural-sounding voices

Its SKILL.md is about 1.3k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering Text to speech and voice. The repository describes itself as: The Autonomous AI Writing Agent — a secure, author-focused AI for fiction and nonfiction authors (Planning, Revision, Promotion, and more). The licence is MIT.

When your agent uses it

  • Tasks that involve Text to speech and voice

Example prompts

  • “/read-aloud”

Requirements

  • Python 3

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. Listen — Read Aloud plays your chapter
  2. Flag — Mark spots that sound wrong (keyboard shortcut or voice command)
  3. Review — After playback, see all flagged locations with context
  4. Fix — Edit each flagged passage
  5. Re-listen — Hear just the fixed sections to verify

What it can do on your machine

Read from SKILL.md and the folder at commit 47e9570. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • pip

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use pip, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Read Aloud loads about 1.3k tokens when it runs. Until then it costs about 29 tokens; SKILL.md has 513 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~29
When it runs · the whole SKILL.md, loaded when a task matches
~1.3k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from Ckokoski/AuthorAgent at commit 47e9570, republished under its MIT licence (© Ckokoski). 513 words, ~1,265 tokens.

Download SKILL.mdSave it as .claude/skills/read-aloud/SKILL.md (or your agent's skills folder).
name
read-aloud
description
Text-to-speech read-aloud for manuscripts using free open-source Piper TTS with natural-sounding voices
author
Writing Secrets
version
1.0.0
triggers
read aloud, read this, speak, text to speech, tts, listen, audio preview, read chapter, hear it
permissions
file:read, file:write, system:exec

Read Aloud — Premium Skill

Turn your manuscript into speech using free, open-source text-to-speech. Hear your prose the way a reader would — because the ear catches what the eye misses.

Why Read Aloud?

Every professional editor will tell you: read your work aloud. It reveals:

  • Awkward phrasing that looks fine on screen
  • Sentences that are too long to speak in one breath
  • Repeated words your eye skips but your ear catches
  • Dialogue that sounds unnatural when spoken
  • Pacing problems that only rhythm reveals
  • Tongue-twisters and consonant clusters

This skill automates that process with natural-sounding AI voices.

Engine: Piper TTS

AuthorClaw uses Piper TTS — a fast, free, open-source text-to-speech engine:

  • MIT licensed — completely free, no API costs, no subscriptions
  • Runs locally — your manuscript never leaves your machine
  • CPU-only — no GPU required (works on any computer)
  • Natural voices — neural network voices, not robotic
  • Fast — generates audio faster than real-time
  • Offline — works without internet
Setup (One-Time)
bash
# Install Piper (included in AuthorClaw setup wizard)
pip install piper-tts

# Download a voice model (~100MB each)
# AuthorClaw will prompt you to pick one on first use
Available Voice Styles
  • en_US-lessac-medium — Clear American narrator (recommended for fiction)
  • en_US-libritts-high — High quality, natural cadence
  • en_GB-alba-medium — British narrator
  • en_US-amy-medium — Female American voice
  • Multiple languages available for translated works

Read Modes

Chapter Read

Read an entire chapter from your project:

read chapter 7
  • Generates audio file saved to workspace/audio/
  • Plays through system audio
  • Shows word-by-word highlighting in dashboard (if connected)
Selection Read

Read a specific passage:

read aloud [paste or select text]
  • Quick listen for a specific section
  • Great for testing dialogue flow
Dialogue Mode

Reads with distinct pausing for dialogue vs. narration:

  • Slight pause before and after quoted speech
  • Different cadence for dialogue vs. description
  • Helps you hear if your dialogue sounds natural
Revision Mode

Reads slowly with pauses between paragraphs:

  • Gives you time to note issues
  • Automatically marks the timestamp when you say "flag" or press a key
  • Generates a revision note file with flagged locations
Show full SKILL.md (223 more words)Show less

Ear-Edit Workflow

A structured process for audio-based revision:

  1. Listen — Read Aloud plays your chapter
  2. Flag — Mark spots that sound wrong (keyboard shortcut or voice command)
  3. Review — After playback, see all flagged locations with context
  4. Fix — Edit each flagged passage
  5. Re-listen — Hear just the fixed sections to verify
Ear-Edit Report: Chapter 7
───────────────────────────
Duration: 18:42
Flags: 6

Flag 1 — 02:14 (Paragraph 4)
"She walked through the door and walked across the room and sat down."
Issue: Triple action chain, repeated "walked"
Suggestion: Vary the verbs, combine actions

Flag 2 — 05:38 (Paragraph 11)
"The simultaneously spectacular and spectacularly simultaneous..."
Issue: Tongue-twister / consonant cluster
Suggestion: Simplify

[... remaining flags ...]

Audio Export

Generate audio files from your manuscript:

  • WAV — Uncompressed, highest quality
  • MP3 — Compressed, smaller files
  • Chapter-by-chapter — Separate files per chapter
  • Full manuscript — Single continuous audio file

Useful for:

  • Sending audio versions to beta readers
  • Creating audiobook demos
  • Personal review while commuting/walking
  • Accessibility for readers with visual impairments

Integration

  • Voice Profile — Adjusts reading speed and emphasis to match your genre
  • Ghostwriter Pro — Read AI-generated scenes aloud immediately
  • Dictation Cleanup — Listen to cleaned-up dictation to verify it sounds right
  • Book Bible — Consistent pronunciation of character/location names

Performance

  • Generates ~10 minutes of audio per minute of processing (CPU)
  • A full chapter (~4,000 words) takes ~30 seconds to generate
  • Audio quality: Near-professional narration quality
  • Storage: ~1MB per minute of audio (MP3)

Commands

  • read aloud — Read selected/pasted text
  • read chapter [number] — Read a full chapter
  • read dialogue [chapter] — Read with dialogue emphasis
  • ear edit [chapter] — Start an ear-edit revision session
  • export audio [chapter/all] — Generate audio files
  • set voice [voice-name] — Change the TTS voice
  • list voices — Show available voice models
  • read speed [slow/normal/fast] — Adjust reading speed

© Ckokoski, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/_archived/premium/read-aloud of Ckokoski/AuthorAgent.

Open the folder on GitHubat commit 47e9570

Compare with similar skills

Read Aloud next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Read Aloud compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Read Aloud this skillCkokoski/AuthorAgent124—~1.3kAutomated safety check: PassMIT
MoneyPrinterTurbo Video Generatorharry0703/MoneyPrinterTurbo129k—~2.1kAutomated safety check: WarnMIT
HyperFrames Media Useheygen-com/hyperframes60k—~2.4kAutomated safety check: PassApache-2.0
Openspec OnboardSAP/e-mobility-charging-stations-simulator22725 repos~3.5kAutomated safety check: PassMIT
Blog AudioAgriciDaniel/claude-blog2.3k1 repos~2.2kAutomated safety check: NotesMIT
Musictadaspetra/loop2962 repos~827Automated safety check: PassMIT

Similar skills

  • MoneyPrinterTurbo Video Generator

    harry0703/MoneyPrinterTurbo

    Installs and runs MoneyPrinterTurbo to turn a topic or script into a finished short video with voice-over, subtitles, stock footage and music.

    129k GitHub stars~2.1k tokensUpdated today
    Media & CreativeAuto-check: warnings
  • HyperFrames Media Use

    heygen-com/hyperframes

    Finds, generates and edits media for HyperFrames video projects: music, sound effects, images, icons, logos, voiceovers, captions and color grades.

    60k GitHub stars~2.4k tokensUpdated today
    Media & CreativeAuto-check passed
  • Openspec Onboard

    SAP/e-mobility-charging-stations-simulator

    Official

    Guided onboarding for OpenSpec - walk through a complete workflow cycle with narration and real codebase work.

    227 GitHub starsUsed in 25 repos~3.5k tokens
    Media & CreativeAuto-check passed
  • Blog Audio

    AgriciDaniel/claude-blog

    Generate audio narration of blog posts using Google Gemini TTS.

    2.3k GitHub starsUsed in 1 repo~2.2k tokens
    Media & CreativeAuto-check: notes
  • Music

    tadaspetra/loop

    Generate music using ElevenLabs Music API. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 2 repos~827 tokens
    Media & CreativeAuto-check passed
  • Create News Video

    hoquanghai/Auto-Create-Video

    Tạo video tin tức ngắn 9:16 (~60s) từ URL bài báo hoặc file .txt tiếng Việt.

    319 GitHub starsUsed in 1 repo~3.7k tokens
    Media & CreativeAuto-check passed

More from Ckokoski/AuthorAgent

All 23 skills in this repo
  • After Action Review

    Ckokoski/AuthorAgent

    Structured post-goal reflection that extracts lessons, evaluates quality, and feeds the self-improvement loop

    124 GitHub stars~1.6k tokensUpdated 3 mo ago
    Auto-check passed
  • Comp Title Finder

    Ckokoski/AuthorAgent

    Find and analyze comparable titles for query letters, marketing, and positioning strategy

    124 GitHub stars~1.4k tokensUpdated 3 mo ago
    Auto-check passed
  • Continuity Check

    Ckokoski/AuthorAgent

    Scan manuscript for inconsistencies in characters, timeline, settings, and names

    124 GitHub stars~1.5k tokensUpdated 3 mo ago
    Auto-check passed
  • Deep Voice Analysis

    Ckokoski/AuthorAgent

    Advanced 47-marker voice analysis engine — analyzes your writing to build a comprehensive Voice Profile for AuthorClaw

    124 GitHub stars~1.4k tokensUpdated 3 mo ago
    Auto-check passed
  • Dictation Cleanup

    Ckokoski/AuthorAgent

    Transform raw speech-to-text dictation into polished prose while preserving the author's natural voice

    124 GitHub stars~1.2k tokensUpdated 3 mo ago
    Auto-check passed
  • Error Recovery

    Ckokoski/AuthorAgent

    Intelligent error diagnosis, automatic recovery strategies, and prevention of recurring failures

    124 GitHub stars~1.7k tokensUpdated 3 mo ago
    Auto-check passed

Questions about Read Aloud

What does Read Aloud do?

Text-to-speech read-aloud for manuscripts using free open-source Piper TTS with natural-sounding voices. Read Aloud is an agent skill from Ckokoski/AuthorAgent.

When should I use Read Aloud?

Read Aloud fits situations like: tasks that involve Text to speech and voice.

How do I install Read Aloud in Claude Code?

Run `npx skills add Ckokoski/AuthorAgent --skill read-aloud -a claude-code`. Or copy the skill folder (skills/_archived/premium/read-aloud in Ckokoski/AuthorAgent) into .claude/skills/read-aloud in your project. Claude Code loads it when a task matches its description.

How do I install Read Aloud in Codex?

Run `npx skills add Ckokoski/AuthorAgent --skill read-aloud -a codex`. Or copy the skill folder (skills/_archived/premium/read-aloud in Ckokoski/AuthorAgent) into .agents/skills/read-aloud in your project. Codex loads it when a task matches its description.

Can I use Read Aloud in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Ckokoski/AuthorAgent --skill read-aloud -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/read-aloud, .gemini/skills/read-aloud, .github/skills/read-aloud and .opencode/skills/read-aloud in your project.

What does Read Aloud need to run?

Going by SKILL.md and its folder, Read Aloud needs the command-line tools its instructions call (pip). Our summary lists: Python 3.

Does Read Aloud access the network?

SKILL.md contains no URLs. Its commands use pip, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Read Aloud safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Read Aloud use?

Read Aloud is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Read Aloud use?

About 1.3k tokens (SKILL.md is roughly 5.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Read Aloud?

Skills that share tags, products or a category with Read Aloud: MoneyPrinterTurbo Video Generator (harry0703/MoneyPrinterTurbo, 129k stars), HyperFrames Media Use (heygen-com/hyperframes, 60k stars), Openspec Onboard (SAP/e-mobility-charging-stations-simulator, 227 stars) and Blog Audio (AgriciDaniel/claude-blog, 2.3k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Read Aloud?

Ckokoski (a GitHub user) maintains it in Ckokoski/AuthorAgent, which has 124 GitHub stars. The repository holds 23 skills in this directory. The repository was last updated on July 11, 2026.

Source: Ckokoski/AuthorAgent on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.