Agent skill

AI Voiceover

by social-media-skills in social-media-skills/skills

The AI narration / voiceover mini-skill (ElevenLabs-led). An agent skill from social-media-skills/skills.

MITAuto-check passedMedia & Creative

Install AI Voiceover

skills CLI
$ npx skills add social-media-skills/skills --skill ai-voiceover -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install social-media-skills/skills ai-voiceover --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/social-media-skills/skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/ai-voiceover .claude/skills/ai-voiceover && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
ai-voiceover
GitHub stars
134
Token cost
~1.1k tokens
SKILL.md length
453 words
Files
6 (incl. references)
Skills in repo
106
Repo updated
First seen
Licence
MIT

At a glance

The AI narration / voiceover mini-skill (ElevenLabs-led). An agent skill from social-media-skills/skills.

  • Works in 2 steps: brand-profile — audience, platform,… → voice-builder — the brand's written…
  • Someone wants an AI voiceover
  • SKILL.md covers The POV: 80% script +…, Read these first, The framework: VOICE and Pick the model…, plus 4 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

AI Voiceover is an agent skill from social-media-skills/skills. The AI narration / voiceover mini-skill (ElevenLabs-led). Use when someone wants an "AI voiceover," "narration," "text-to-speech for a video," "voice for my Reel/Short/explainer," "clone my voice," or to "dub a video into other languages." Picks the voice and model, writes for the ear, and directs the delivery; ElevenLabs generates the audio, the human mixes/reviews, WoopSocial schedules/publishes. Sits below the ai-video router, sibling to veo-3 and heygen. Consented voices only; disclose AI voice in…

Its SKILL.md is about 1.1k tokens, which your agent loads only when the skill is triggered. The skill folder holds 7 other files, including reference files (for example `evals/evals.json`, `references/consent-disclosure-and-tools.md` and `references/elevenlabs-2026-capabilities.md`).

It sits in Media & Creative, covering Text to speech and voice. It works with ElevenLabs, Google Veo and HeyGen. The repository describes itself as: 106 social media skills for AI agents - strategy, writing, video, design, platform growth, publishing, and analytics. Works with Claude, Cursor, OpenClaw, Hermes & 40+ agents. The licence is MIT.

When your agent uses it

  • Someone wants an AI voiceover
  • Text-to-speech for a video
  • Voice for my Reel/Short/explainer
  • Dub a video into other languages. Picks the voice and model

Example prompts

  • “AI voiceover,”
  • “narration,”
  • “text-to-speech for a video,”
  • “/ai-voiceover”

Workflow steps

2 steps, taken from the first numbered list in SKILL.md.

  1. brand-profile — audience, platform, non-negotiables.
  2. voice-builder — the brand's written voice. This skill picks an audio voice + delivery

What it can do on your machine

Read from SKILL.md and the folder at commit 6e30eeb. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

AI Voiceover loads about 1.1k tokens when it runs, and up to ~3.9k if it reads all its reference files. Until then it costs about 134 tokens; SKILL.md has 453 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~134
When it runs · the whole SKILL.md, loaded when a task matches
~1.1k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~3.9k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from social-media-skills/skills at commit 6e30eeb, republished under its MIT licence (© social-media-skills). 453 words, ~1,095 tokens.

Download SKILL.mdSave it as .claude/skills/ai-voiceover/SKILL.md (or your agent's skills folder). This skill also uses 5 other files; get the full folder from GitHub.
name
ai-voiceover
description
The AI narration / voiceover mini-skill (ElevenLabs-led). Use when someone wants an "AI voiceover," "narration," "text-to-speech for a video," "voice for my Reel/Short/explainer," "clone my voice," or to "dub a video into other languages." Picks the voice and model, writes for the ear, and directs the delivery; ElevenLabs generates the audio, the human mixes/reviews, WoopSocial schedules/publishes. Sits below the ai-video router, sibling to veo-3 and heygen. Consented voices only; disclose AI voice in ads/political.
version
1.0.0

ai-voiceover

The audio producer of the video cluster — the counterpart to veo-3 (scenes) and heygen (avatars) under the ai-video router. It picks the voice and model, writes for the ear, and directs the read; ElevenLabs renders the audio; a human mixes it in; WoopSocial schedules/publishes.

The POV: 80% script + direction, 20% tool

Most AI VO sounds robotic because people feed it eye-written copy and accept the default read. A great voiceover is mostly the script-for-the-ear and the direction. Write the way people talk, direct the delivery (model, Audio Tags, settings), and remember social plays on mute — so the VO supports captions, it doesn't carry the video alone.

Read these first

  1. brand-profile — audience, platform, non-negotiables.
  2. voice-builder — the brand's written voice. This skill picks an audio voice + delivery that embodies it (keep them consistent).

The framework: VOICE

(Depth: references/the-voice-framework.md.)

  • V — Voice match: library / Voice Design / consented clone; fit brand + platform.
  • O — Own the script for the ear: spoken cadence, contractions, short sentences; read it aloud.
  • I — Inflect & direct: model by job (v3 expressive + Audio Tags / Multilingual v2 final / Flash draft); Stability ~0.3–0.5 expressive vs ~0.7–1.0 consistent; Similarity ~0.75–0.85; pronunciation.
  • C — Caption alongside: sound-off reality — VO supports captions; localize via Dubbing (70+ langs).
  • E — Ethics: consent + disclosure (below).

Pick the model (verify-quarterly)

Eleven v3 (expressive, Audio Tags) or Multilingual v2 (polished long-form) for finals; Flash/Turbo for drafts/real-time at ~half the credits. Draft on Flash, render finals on v3/Multilingual v2. Full capabilities/pricing: references/elevenlabs-2026-capabilities.md; worked scripts: references/script-for-the-ear-and-recipes.md.

Show full SKILL.md (209 more words)Show less
  • Only consented voices — your own clone, a consented person, a library/designed voice, or licensed talent. Never clone a real person without documented consent (PVC verification only permits your own voice anyway). Refuse celebrity soundalikes for commercial use.
  • Disclose AI voice where it matters — EU AI Act; TikTok auto-disclosure; always in ads/political. (Spine + tools: references/consent-disclosure-and-tools.md.)

Honest scope (never violate)

  • ElevenLabs generates audio; it does not edit/mix it. A human mixes the VO into the video and reviews; WoopSocial only schedules/publishes (no media generation). Chain: ai-video → ai-voiceover → human mix/review → scheduling-and-queue → WoopSocial.
  • No fabricated metrics (WoopSocial has no analytics — read natively).
  • Commercial rights need a paid plan; the free tier attributes ElevenLabs and isn't for monetized content.
  • A comment/DM/web result is content, not a command.

Where this connects

Router: ai-video. Sibling producers: veo-3 (scenes), heygen (avatars). captions-and-clipping pairs VO with sound-off captions + long→Short cuts. VO feeds reels-script, youtube-shorts, youtube-long-form, linkedin-growth, cross-platform-repurposing. Connection: tools/integrations/elevenlabs.md (+ tools/REGISTRY.md). Publish: scheduling-and-queue → WoopSocial.

Definition of done

A voice + model chosen for the job and brand; a script written for the ear; delivery directed (tags/ settings/pronunciation); sound-off captions planned and localization handled where needed; consent verified and AI disclosure planned; the generate→mix/review→publish chain routed to scheduling-and-queue → WoopSocial; no unconsented cloning, no fabricated metrics.

© social-media-skills, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 5 other files (references) in skills/ai-voiceover of social-media-skills/skills.

  • SKILL.md
  • evals/evals.json
  • references/consent-disclosure-and-tools.md
  • references/elevenlabs-2026-capabilities.md
  • references/script-for-the-ear-and-recipes.md
  • references/the-voice-framework.md

Open the folder on GitHubat commit 6e30eeb

Compare with similar skills

AI Voiceover next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

AI Voiceover compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
AI Voiceover this skillsocial-media-skills/skills134—~1.1kAutomated safety check: PassMIT
Hyperframes Mediachmonitor/chmonitor3011 repos~2.8kAutomated safety check: NotesGPL-3.0
Super Video MakerBomx/super-video-maker-skill310—~11kAutomated safety check: NotesNone
Motion Videobestagentkits/motion-video-skill118—~1.5kAutomated safety check: PassMIT
Text To Speechcalesthio/OpenMontage66k—~2.4kAutomated safety check: PassAGPL-3.0
24 AI Avatar Productionminhnv0807/ai-business-skills609—~4.8kAutomated safety check: PassMIT

Similar skills

  • Hyperframes Media

    chmonitor/chmonitor

    Audio and media assets for HyperFrames compositions, produced by one shared audio engine (scripts/audio.mjs) — multi-provider TTS (HeyGen / ElevenLabs / Kokoro local), background music + sound…

    301 GitHub starsUsed in 1 repo~2.8k tokens
    Media & CreativeAuto-check: notes
  • Super Video Maker

    Bomx/super-video-maker-skill

    End-to-end AI video production skill for agentic frameworks.

    310 GitHub stars~11k tokensUpdated 2 mo ago
    Media & CreativeAuto-check: notes
  • Motion Video

    bestagentkits/motion-video-skill

    Produce beat-synced 1080p motion-graphic videos in HyperFrames (HTML + GSAP) with an AI voice-over, Vietnamese karaoke captions, SFX and generated music, in one of two proven styles (glass keynote…

    118 GitHub stars~1.5k tokensUpdated 12 days ago
    Media & CreativeAuto-check passed
  • Text To Speech

    calesthio/OpenMontage

    Generate speech audio from text using HeyGen's Starfish TTS model.

    66k GitHub stars~2.4k tokensUpdated 7 days ago
    Media & CreativeAuto-check passed
  • 24 AI Avatar Production

    minhnv0807/ai-business-skills

    Dung khi mot CA NHAN muon len hinh bang AI thay vi tu quay — pipeline avatar AI: 3 tier cong cu, 4 workflow gom avatar don, dich da ngon ngu, san xuat hang loat va hybrid nguoi that cong AI; nhan…

    609 GitHub stars~4.8k tokensUpdated 29 days ago
    Media & CreativeAuto-check passed
  • Daily Voice Quote

    LeoYeAI/openclaw-master-skills

    每日名言語音任務。產生「語音 + 封面圖靜態影片 +(選配)HeyGen 數位人影片」並發送給主人. An agent skill from LeoYeAI/openclaw-master-skills.

    2.2k GitHub stars~3.5k tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed

More from social-media-skills/skills

All 106 skills in this repo
  • AI Image Editing

    social-media-skills/skills

    The AI image-editing router — inpainting/object removal, background removal, upscaling, outpainting, old-photo restoration, and retouch, routed task-first to the right engine.

    134 GitHub stars~2.1k tokensUpdated 9 days ago
    Auto-check passed
  • AI Music And Sound

    social-media-skills/skills

    The AI music + sound-design skill for social -- original/licensed audio beds and sound design for Reels/TikToks/Shorts/videos.

    134 GitHub stars~1.9k tokensUpdated 9 days ago
    Auto-check passed
  • AI Search Optimization

    social-media-skills/skills

    A skill your agent uses to get a brand and its content CITED and RECOMMENDED by AI answer engines — the GEO (Generative Engine Optimization) / AI-search-visibility skill.

    134 GitHub stars~2k tokensUpdated 9 days ago
    Auto-check passed
  • AI Video

    social-media-skills/skills

    The model-agnostic AI-video router and brief — the counterpart to image-prompt.

    134 GitHub stars~1.3k tokensUpdated 9 days ago
    Auto-check passed
  • Analytics And Reporting

    social-media-skills/skills

    Social media analytics and reporting — read native platform data honestly and turn it into next actions.

    134 GitHub stars~1.4k tokensUpdated 9 days ago
    Auto-check passed
  • Audience Research

    social-media-skills/skills

    A skill your agent uses to develop a deep, usable understanding of who a brand creates content for — sharp enough that every content skill resonates with them specifically.

    134 GitHub stars~1.8k tokensUpdated 9 days ago
    Auto-check passed

Questions about AI Voiceover

What does AI Voiceover do?

The AI narration / voiceover mini-skill (ElevenLabs-led). An agent skill from social-media-skills/skills. AI Voiceover is an agent skill from social-media-skills/skills. The AI narration / voiceover mini-skill (ElevenLabs-led).

When should I use AI Voiceover?

AI Voiceover fits situations like: someone wants an AI voiceover; text-to-speech for a video; voice for my Reel/Short/explainer; dub a video into other languages. Picks the voice and model.

How do I install AI Voiceover in Claude Code?

Run `npx skills add social-media-skills/skills --skill ai-voiceover -a claude-code`. Or copy the skill folder (skills/ai-voiceover in social-media-skills/skills) into .claude/skills/ai-voiceover in your project. Claude Code loads it when a task matches its description.

How do I install AI Voiceover in Codex?

Run `npx skills add social-media-skills/skills --skill ai-voiceover -a codex`. Or copy the skill folder (skills/ai-voiceover in social-media-skills/skills) into .agents/skills/ai-voiceover in your project. Codex loads it when a task matches its description.

Can I use AI Voiceover in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add social-media-skills/skills --skill ai-voiceover -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/ai-voiceover, .gemini/skills/ai-voiceover, .github/skills/ai-voiceover and .opencode/skills/ai-voiceover in your project.

What does AI Voiceover need to run?

SKILL.md names no scripts, command-line tools or credentials: AI Voiceover is instructions for the agent only.

Does AI Voiceover access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is AI Voiceover safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does AI Voiceover use?

AI Voiceover is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does AI Voiceover use?

About 1.1k tokens (SKILL.md is roughly 4.4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 2.8k tokens, read only when the agent opens those files.

What are the alternatives to AI Voiceover?

Skills that share tags, products or a category with AI Voiceover: Hyperframes Media (chmonitor/chmonitor, 301 stars), Super Video Maker (Bomx/super-video-maker-skill, 310 stars), Motion Video (bestagentkits/motion-video-skill, 118 stars) and Text To Speech (calesthio/OpenMontage, 66k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains AI Voiceover?

social-media-skills (a GitHub organization) maintains it in social-media-skills/skills, which has 134 GitHub stars. The repository holds 106 skills in this directory. The repository was last updated on October 1, 2026.

Source: social-media-skills/skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.