Agent skill

Speech Adaptation

by jwynia in jwynia/agent-skills

Transform comprehensive written content into purposeful spoken guidance.

MITAuto-check passedMedia & Creative

Install Speech Adaptation

skills CLI
$ npx skills add jwynia/agent-skills --skill speech-adaptation -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install jwynia/agent-skills speech-adaptation --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/jwynia/agent-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/general/communication/speech-adaptation .claude/skills/speech-adaptation && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
speech-adaptation
GitHub stars
170
Token cost
~2.1k tokens
SKILL.md length
939 words
Files
1
Skills in repo
111
Repo updated
First seen
Licence
MIT

At a glance

Transform comprehensive written content into purposeful spoken guidance.

  • Works in 11 steps: Hierarchical Restructuring → Front-Load Value → Compress Conceptual Space → …
  • Adapting for speech
  • SKILL.md covers Purpose, Core Principle, Functional Intent Detection and Context Signals, plus 8 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Speech Adaptation is an agent skill from jwynia/agent-skills. Transform comprehensive written content into purposeful spoken guidance. Use when adapting for speech, converting to spoken format, optimizing for listening, or creating audio content from written material. Keywords: speech, audio, spoken, listening, adaptation, podcast.

Its SKILL.md is about 2.1k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative. The licence is MIT.

When your agent uses it

  • Adapting for speech
  • Converting to spoken format
  • Optimizing for listening
  • Creating audio content from written material

Example prompts

  • “/speech-adaptation”

Workflow steps

11 steps, taken from the step headings in SKILL.md.

  1. Hierarchical Restructuring
  2. Front-Load Value
  3. Compress Conceptual Space
  4. Context-Dependent Detail
  5. Eliminate Structural Artifacts
  6. Progressive Revelation Strategy
  7. Uniform Compression
  8. Written Sentences Spoken
  9. Exhaustive Completeness
  10. Missing Signposts
  11. Buried Action

What it can do on your machine

Read from SKILL.md and the folder at commit e02ec7e. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Speech Adaptation loads about 2.1k tokens when it runs. Until then it costs about 72 tokens; SKILL.md has 939 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~72
When it runs · the whole SKILL.md, loaded when a task matches
~2.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from jwynia/agent-skills at commit e02ec7e, republished under its MIT licence (© jwynia). 939 words, ~2,087 tokens.

Download SKILL.mdSave it as .claude/skills/speech-adaptation/SKILL.md (or your agent's skills folder).
name
speech-adaptation
description
Transform comprehensive written content into purposeful spoken guidance. Use when adapting for speech, converting to spoken format, optimizing for listening, or creating audio content from written material. Keywords: speech, audio, spoken, listening, adaptation, podcast.
license
MIT
metadata.author
jwynia
metadata.version
1.0
metadata.type
utility
metadata.mode
generative
metadata.domain
writing

Speech Adaptation

Purpose

Transform comprehensive written content into purposeful spoken guidance. Speech requires 3-5x compression while maintaining functional value. Apply when converting written content to audio, podcasts, presentations, or voice assistant responses.

Core Principle

Lead with value, earn attention. Listeners can't skim. Front-load what matters and offer expansion rather than exhaustive delivery.


Functional Intent Detection

Parse the original question/content for intent:

Intent TypeSignalsFocus
Problem-solving"How do I..."Actionable steps
Learning"What is..."Core concepts + examples
Decision-making"Should I..."Key considerations + recommendation
Troubleshooting"Why isn't..."Likely causes + solutions

Context Signals

Signal TypeExamplesAdaptation
Urgency"today", "now", "urgent"Compress to immediate next steps
Scope"huge", "complex", "overwhelming"Lead with simplification
Experience"beginner", "new to"Increase explanation, decrease jargon
Personal stakes"I", "my project"Increase specificity, decrease abstraction

Content Transformation Principles

1. Hierarchical Restructuring

Written: Lists methods 1-7 equally Spoken: "There are three main approaches. Start with [most relevant]. If that doesn't work, try [backup]."

2. Front-Load Value

Written: Builds up to key insights Spoken: Lead with core insight, then supporting details if needed

3. Compress Conceptual Space

Written: Seven distinct frameworks Spoken: "Basically three strategies: sort by importance, limit your focus, or batch similar work"

4. Context-Dependent Detail

Written: Explains everything at same depth Spoken: Start simple, indicate where more detail is available

  • "Use a priority matrix - urgent versus important"
  • Optional expansion cue: "I can break down those four categories if helpful"
5. Eliminate Structural Artifacts

Remove in Speech:

  • Section headers read verbatim
  • Bullet point enumeration
  • Visual formatting cues
  • Redundant category labels

Add for Speech:

  • Transition phrases between ideas
  • Purpose statements before methods
  • Summary/recap statements
6. Progressive Revelation Strategy
  1. Core insight (one sentence)
  2. Primary recommendation (actionable step)
  3. Backup approach (if primary doesn't fit)
  4. Availability cue for additional methods

Implementation Guidelines

Pre-Processing Steps
  1. Parse original question for functional intent and context signals
  2. Identify 1-2 most relevant pieces for their specific need
  3. Determine appropriate compression ratio based on urgency/complexity
Content Selection Rules
ContextSelection
High urgency1 primary method + 1 backup
Learning focusedCore concept + 1 detailed example + availability of more
Decision supportKey considerations + clear recommendation
Complex topicSimplify conceptual framework first, offer detail expansion
Speech-Specific Adaptations
  • Replace structural language with functional language
  • Add explicit transitions between ideas
  • Use pronouns and referential terms to avoid repetition
  • Include "escape valves" for different user needs
  • End with clear next step or summary

Quality Checks

TestQuestion
CompressionIs this 30-50% of original length?
CompletenessDoes this answer their core question?
FlowWould this make sense heard linearly?
ActionDo they know what to do next?

Example Transformation

Question Type: Immediate problem-solving with overwhelm signals

Written Response: 7 methods with full explanations

Spoken Adaptation:

  1. Acknowledge state: "When facing a huge list..."
  2. Core insight: "The key is separating what needs doing from what feels urgent"
  3. Primary action: "Try this: scan for things both urgent AND important"
  4. Boundary setting: "Pick just 3 - more than that sets you up to feel behind"
  5. Escape valve: "Other approaches available if this doesn't click"

Success Metrics

  • User can act immediately after listening
  • Cognitive load feels manageable
  • Key insights retained after single hearing
  • Optional detail access feels natural when needed

Integration Points

Inbound:

  • From written documentation or articles
  • From comprehensive analysis outputs
  • From detailed framework content

Outbound:

  • To audio content production
  • To presentation delivery
  • To voice assistant responses

Complementary:

  • presentation-design: For visual + spoken coordination
  • dialogue: For conversational delivery patterns
Show full SKILL.md (368 more words)Show less

Anti-Patterns

1. Uniform Compression

Pattern: Reducing all content by the same ratio regardless of importance. Why it fails: Not all content is equal. Some ideas need full explanation; others can be summarized in a phrase. Equal compression buries critical insights and pads trivial ones. Fix: Identify the 1-2 most important points. Protect those while ruthlessly compressing supporting material. Lead with what matters most.

2. Written Sentences Spoken

Pattern: Reading written prose aloud without restructuring for speech patterns. Why it fails: Written and spoken language have different rhythms, sentence structures, and information density. Written sentences spoken sound formal, awkward, and hard to follow. Fix: Restructure for oral delivery. Shorter sentences. More personal pronouns. Explicit transitions. Repetition for emphasis. Natural breathing points.

3. Exhaustive Completeness

Pattern: Including all information from the written source because "it might be important." Why it fails: Listeners can't skim, reread, or control pace. Information overload in speech creates immediate cognitive overload and retention collapse. Fix: Accept that spoken content is selective. Provide escape valves: "More on this if helpful." Trust that listeners can ask for expansion rather than front-loading everything.

4. Missing Signposts

Pattern: Moving between ideas without explicit verbal transitions. Why it fails: Listeners can't see paragraph breaks or headings. Without verbal signposts, ideas blur together. The structure becomes invisible. Fix: Add explicit transitions: "First..." "More importantly..." "Here's the key point..." "Moving on to..." Make the structure audible.

5. Buried Action

Pattern: Leaving actionable recommendations for the end after extensive context. Why it fails: Listeners who zone out during context miss the action items. Those still engaged have forgotten the details by the time recommendations arrive. Fix: Front-load action with context to follow. "Do X. Here's why..." rather than "Here's all the context, therefore do X."

Integration

Inbound (feeds into this skill)
SkillWhat it provides
prose-styleWritten content quality to work from
(written documentation)Source material for adaptation
Outbound (this skill enables)
SkillWhat this provides
presentation-designSpoken content structure for slide coordination
(audio production)Scripts ready for recording
(voice assistants)Responses optimized for spoken delivery
Complementary
SkillRelationship
presentation-designSpeech-adaptation handles the spoken component; presentation-design coordinates visual and spoken elements
dialogueSpeech-adaptation for informational delivery; dialogue for conversational and dramatic speech patterns

© jwynia, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/general/communication/speech-adaptation of jwynia/agent-skills.

Open the folder on GitHubat commit e02ec7e

Compare with similar skills

Speech Adaptation next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Speech Adaptation compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Speech Adaptation this skilljwynia/agent-skills170—~2.1kAutomated safety check: PassMIT
Guizang Social Cardsop7418/guizang-social-card-skill7.4k1 repos~7.8kAutomated safety check: PassAGPL-3.0
Weekly Changelog Videoheygen-com/hyperframes60k—~3.3kAutomated safety check: PassApache-2.0
Anthropic Brand Stylinganthropics/skills180k30 repos~559Automated safety check: PassApache-2.0
MoneyPrinterTurbo Video Generatorharry0703/MoneyPrinterTurbo129k—~2.1kAutomated safety check: WarnMIT
HyperFrames Media Useheygen-com/hyperframes60k—~2.4kAutomated safety check: PassApache-2.0

Similar skills

  • Guizang Social Cards

    op7418/guizang-social-card-skill

    Produces social card sets for Xiaohongshu and WeChat: carousels, Live Photo motion cards and puzzle layouts, and WeChat cover pairs, rendered from single-file HTML.

    7.4k GitHub starsUsed in 1 repo~7.8k tokens
    Media & CreativeAuto-check passed
  • Weekly Changelog Video

    heygen-com/hyperframes

    Turns a weekly changelog markdown file into a branded HyperFrames video with voiceover, animated mock-UI scenes and captions, using fonts, background and scripts bundled in the skill.

    60k GitHub stars~3.3k tokensUpdated today
    Media & CreativeAuto-check passed
  • Anthropic Brand Styling

    anthropics/skills

    Official

    Applies Anthropic's brand colors and fonts to artifacts such as PowerPoint slides, using fixed hex values for text and accents, Poppins headings and Lora body text.

    180k GitHub starsUsed in 30 repos~559 tokens
    Media & CreativeAuto-check passed
  • MoneyPrinterTurbo Video Generator

    harry0703/MoneyPrinterTurbo

    Installs and runs MoneyPrinterTurbo to turn a topic or script into a finished short video with voice-over, subtitles, stock footage and music.

    129k GitHub stars~2.1k tokensUpdated today
    Media & CreativeAuto-check: warnings
  • HyperFrames Media Use

    heygen-com/hyperframes

    Finds, generates and edits media for HyperFrames video projects: music, sound effects, images, icons, logos, voiceovers, captions and color grades.

    60k GitHub stars~2.4k tokensUpdated today
    Media & CreativeAuto-check passed
  • Holo Card Studio

    EverettFish/holo-card-studio

    Create collectible holographic foil cards and two-image lenticular flip cards with AI-generated full-color ukiyo-e and colored sumi-e anime artwork, layered Blender scenes, renders, GLB export, and…

    1.9k GitHub stars~1.4k tokensUpdated 18 days ago
    Media & CreativeAuto-check passed

More from jwynia/agent-skills

All 111 skills in this repo
  • Devcontainer

    jwynia/agent-skills

    Diagnose devcontainer configuration problems and guide development environment setup.

    170 GitHub stars~1.2k tokensUpdated 7 mo ago
    Auto-check: notes
  • Frontend Design

    jwynia/agent-skills

    Create distinctive, production-grade frontend interfaces with high design quality.

    170 GitHub stars~3.2k tokensUpdated 7 mo ago
    Auto-check passed
  • Gitea Workflow

    jwynia/agent-skills

    Orchestrate agile development workflows for Gitea repositories using the tea CLI.

    170 GitHub stars~3.8k tokensUpdated 7 mo ago
    Auto-check passed
  • Godot Asset Generator

    jwynia/agent-skills

    Generate game assets using AI image generation APIs (DALL-E, Replicate, fal.ai) and prepare them for Godot.

    170 GitHub stars~3.8k tokensUpdated 7 mo ago
    Auto-check passed
  • Mastra Hono

    jwynia/agent-skills

    Develop AI agents, tools, and workflows with Mastra v1 Beta and Hono servers.

    170 GitHub stars~2.9k tokensUpdated 7 mo ago
    Auto-check passed
  • PPTX Generator

    jwynia/agent-skills

    Create and manipulate PowerPoint PPTX files programmatically.

    170 GitHub stars~3.1k tokensUpdated 7 mo ago
    Auto-check passed

Questions about Speech Adaptation

What does Speech Adaptation do?

Transform comprehensive written content into purposeful spoken guidance. Speech Adaptation is an agent skill from jwynia/agent-skills. Transform comprehensive written content into purposeful spoken guidance.

When should I use Speech Adaptation?

Speech Adaptation fits situations like: adapting for speech; converting to spoken format; optimizing for listening; creating audio content from written material.

How do I install Speech Adaptation in Claude Code?

Run `npx skills add jwynia/agent-skills --skill speech-adaptation -a claude-code`. Or copy the skill folder (skills/general/communication/speech-adaptation in jwynia/agent-skills) into .claude/skills/speech-adaptation in your project. Claude Code loads it when a task matches its description.

How do I install Speech Adaptation in Codex?

Run `npx skills add jwynia/agent-skills --skill speech-adaptation -a codex`. Or copy the skill folder (skills/general/communication/speech-adaptation in jwynia/agent-skills) into .agents/skills/speech-adaptation in your project. Codex loads it when a task matches its description.

Can I use Speech Adaptation in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add jwynia/agent-skills --skill speech-adaptation -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/speech-adaptation, .gemini/skills/speech-adaptation, .github/skills/speech-adaptation and .opencode/skills/speech-adaptation in your project.

What does Speech Adaptation need to run?

SKILL.md names no scripts, command-line tools or credentials: Speech Adaptation is instructions for the agent only.

Does Speech Adaptation access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Speech Adaptation safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Speech Adaptation use?

Speech Adaptation is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Speech Adaptation use?

About 2.1k tokens (SKILL.md is roughly 8.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Speech Adaptation?

Skills that share tags, products or a category with Speech Adaptation: Guizang Social Cards (op7418/guizang-social-card-skill, 7.4k stars), Weekly Changelog Video (heygen-com/hyperframes, 60k stars), Anthropic Brand Styling (anthropics/skills, 180k stars) and MoneyPrinterTurbo Video Generator (harry0703/MoneyPrinterTurbo, 129k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Speech Adaptation?

jwynia (a GitHub user) maintains it in jwynia/agent-skills, which has 170 GitHub stars. The repository holds 111 skills in this directory. The repository was last updated on February 24, 2026.

Source: jwynia/agent-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.