Elevenlabs Audio Prompting
nodetool-ai/nodetool
Direct ElevenLabs speech, dialogue, sound effects and music — the bracketed audio tags v3 acts on and why the voice decides whether a tag lands, stability as the delivery dial, punctuation instead…
Tired of juggling multiple audio APIs?. An agent skill from pexoai/pexo-skills.
$ npx skills add pexoai/pexo-skills --skill videoagent-audio-studio -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install pexoai/pexo-skills videoagent-audio-studio --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/pexoai/pexo-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/videoagent-audio-studio .claude/skills/videoagent-audio-studio && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "videoagent-audio-studio" agent skill from https://github.com/pexoai/pexo-skills/tree/main/skills/videoagent-audio-studio into .claude/skills/videoagent-audio-studio/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "videoagent-audio-studio", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/pexoai/pexo-skills/tree/main/skills/videoagent-audio-studioType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add pexoai/pexo-skills --skill videoagent-audio-studio -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install pexoai/pexo-skills videoagent-audio-studio --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/pexoai/pexo-skills.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/videoagent-audio-studio .agents/skills/videoagent-audio-studio && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "videoagent-audio-studio" agent skill from https://github.com/pexoai/pexo-skills/tree/main/skills/videoagent-audio-studio into .agents/skills/videoagent-audio-studio/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "videoagent-audio-studio", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add pexoai/pexo-skills --skill videoagent-audio-studio -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install pexoai/pexo-skills videoagent-audio-studio --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/pexoai/pexo-skills.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/videoagent-audio-studio .cursor/skills/videoagent-audio-studio && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "videoagent-audio-studio" agent skill from https://github.com/pexoai/pexo-skills/tree/main/skills/videoagent-audio-studio into .cursor/skills/videoagent-audio-studio/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "videoagent-audio-studio", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/pexoai/pexo-skills.git --path skills/videoagent-audio-studio--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add pexoai/pexo-skills --skill videoagent-audio-studio -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install pexoai/pexo-skills videoagent-audio-studio --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/pexoai/pexo-skills.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/videoagent-audio-studio .gemini/skills/videoagent-audio-studio && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "videoagent-audio-studio" agent skill from https://github.com/pexoai/pexo-skills/tree/main/skills/videoagent-audio-studio into .gemini/skills/videoagent-audio-studio/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "videoagent-audio-studio", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install pexoai/pexo-skills videoagent-audio-studioInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add pexoai/pexo-skills --skill videoagent-audio-studio -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/pexoai/pexo-skills.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/videoagent-audio-studio .github/skills/videoagent-audio-studio && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "videoagent-audio-studio" agent skill from https://github.com/pexoai/pexo-skills/tree/main/skills/videoagent-audio-studio into .github/skills/videoagent-audio-studio/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "videoagent-audio-studio", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add pexoai/pexo-skills --skill videoagent-audio-studio -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install pexoai/pexo-skills videoagent-audio-studio --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/pexoai/pexo-skills.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/videoagent-audio-studio .opencode/skills/videoagent-audio-studio && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "videoagent-audio-studio" agent skill from https://github.com/pexoai/pexo-skills/tree/main/skills/videoagent-audio-studio into .opencode/skills/videoagent-audio-studio/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "videoagent-audio-studio", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
videoagent-audio-studioTired of juggling multiple audio APIs?. An agent skill from pexoai/pexo-skills.
Videoagent Audio Studio is an agent skill from pexoai/pexo-skills. Tired of juggling multiple audio APIs? This skill gives you one-command access to TTS, music generation, sound effects, and voice cloning. Use when you want to generate any audio without managing multiple API keys.
Its SKILL.md is about 1.7k tokens, which your agent loads only when the skill is triggered. The skill folder holds 12 other files (for example `cli.js`, `proxy/api/audio.js` and `proxy/api/stats.js`).
It sits in Media & Creative, covering Text to speech and voice and Music and audio generation. It works with fal, Vercel and ElevenLabs. The repository describes itself as: A collection of open-source Agent Skills for content creation — images, audio, and video. The licence is MIT.
2 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit f724267. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships script files (JavaScript and Shell), which the agent can run.
Shell commands in SKILL.md call:
bashnpmvercelFrom the folder's file list and the shell code blocks in SKILL.md.
Links to these hosts (documentation or services it may open):
elevenlabs.iofal.aiFrom URLs in SKILL.md, links to its own repository left out.
Names these keys or tokens, usually read from environment variables:
ELEVENLABS_API_KEYFAL_KEYFrom names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Videoagent Audio Studio loads about 1.7k tokens when it runs. Until then it costs about 60 tokens; SKILL.md has 532 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from pexoai/pexo-skills at commit f724267, republished under its MIT licence (© pexoai). 532 words, ~1,715 tokens.
.claude/skills/videoagent-audio-studio/SKILL.md (or your agent's skills folder). This skill also uses 9 other files; get the full folder from GitHub.Use when: User asks to generate speech, narrate text, create a voice-over, compose music, or produce a sound effect.
VideoAgent Audio Studio is a smart audio dispatcher. It analyzes your request and routes it to the best available model — ElevenLabs for speech and music, fal.ai for fast SFX — and returns a ready-to-use audio URL.
| Request Type | Best Model | Latency |
|---|---|---|
| Narrate text / Voice-over | elevenlabs-tts-v3 | ~3s |
| Low-latency TTS (real-time) | elevenlabs-tts-turbo | <1s |
| Background music | cassetteai-music | ~15s |
| Sound effect | elevenlabs-sfx | ~5s |
| Clone a voice from audio | elevenlabs-voice-clone | ~10s |
bash {baseDir}/tools/start_server.shThis starts the ElevenLabs MCP server on port 8124. The skill uses it for all audio generation.
Analyze the user's request and call the appropriate tool via the MCP server:
Text-to-Speech (TTS)
When user asks to "narrate", "read aloud", "say", or "create a voice-over":
Use MCP tool: text_to_speech
text: "<the text to narrate>"
voice_id: "JBFqnCBsd6RMkjVDRZzb" # Default: "George" (professional, neutral)
model_id: "eleven_multilingual_v2" # Use "eleven_turbo_v2_5" for low latencyMusic Generation
When user asks to "compose", "create background music", or "make a soundtrack":
Use MCP tool: text_to_sound_effects (via cassetteai-music on fal.ai)
prompt: "<music description, e.g. 'upbeat lo-fi hip hop, 90 seconds'>"
duration_seconds: <duration>Sound Effect (SFX)
When user asks for a specific sound (e.g., "a door creaking", "rain on a window"):
Use MCP tool: text_to_sound_effects
text: "<sound description>"
duration_seconds: <1-22>Voice Cloning
When user provides an audio sample and wants to clone the voice:
Use MCP tool: voice_add
name: "<voice name>"
files: ["<audio_file_url>"]User: "Voice this text for me: Welcome to our product launch"
→ Route to: text_to_speech
text: "Welcome to our product launch"
voice_id: "JBFqnCBsd6RMkjVDRZzb"
model_id: "eleven_multilingual_v2"🎙️ Voiceover done! Listen here
User: "Generate 60 seconds of relaxing background music for a podcast"
→ Route to: cassetteai-music (fal.ai)
prompt: "relaxing lo-fi background music for a podcast, gentle piano and soft beats, 60 seconds"
duration_seconds: 60🎵 Background music ready! Listen here
User: "Generate a sci-fi style door opening sound effect"
→ Route to: text_to_sound_effects
text: "a futuristic sci-fi door sliding open with a hydraulic hiss"
duration_seconds: 3Set ELEVENLABS_API_KEY in ~/.openclaw/openclaw.json:
{
"skills": {
"entries": {
"videoagent-audio-studio": {
"enabled": true,
"env": {
"ELEVENLABS_API_KEY": "your_elevenlabs_key_here"
}
}
}
}
}Get your key at elevenlabs.io/app/settings/api-keys.
"FAL_KEY": "your_fal_key_here"Get your key at fal.ai/dashboard/keys.
The cli.js connects to a hosted proxy by default. If you want full control — or need to serve users in regions where vercel.app is blocked — you can deploy your own instance from the proxy/ directory.
cd proxy
npm install
vercel --prodSet these in your Vercel project (Dashboard → Settings → Environment Variables):
| Variable | Required For | Where to Get |
|---|---|---|
ELEVENLABS_API_KEY | TTS, SFX, Voice Clone | elevenlabs.io/app/settings/api-keys |
FAL_KEY | Music generation | fal.ai/dashboard/keys |
VALID_PRO_KEYS | (Optional) Restrict access | Comma-separated list of allowed client keys |
export AUDIOMIND_PROXY_URL="https://your-domain.com/api/audio"Or set it in ~/.openclaw/openclaw.json:
{
"skills": {
"entries": {
"videoagent-audio-studio": {
"env": {
"AUDIOMIND_PROXY_URL": "https://your-domain.com/api/audio"
}
}
}
}
}If your users are in mainland China, bind a custom domain in Vercel Dashboard → Settings → Domains to avoid DNS issues with vercel.app.
| Model ID | Type | Provider | Notes |
|---|---|---|---|
eleven_multilingual_v2 | TTS | ElevenLabs | Best quality, supports 29 languages |
eleven_turbo_v2_5 | TTS | ElevenLabs | Ultra-low latency, ideal for real-time |
eleven_monolingual_v1 | TTS | ElevenLabs | English only, fastest |
cassetteai-music | Music | fal.ai | Reliable, fast music generation |
elevenlabs-sfx | SFX | ElevenLabs | High-quality sound effects (up to 22s) |
elevenlabs-voice-clone | Clone | ElevenLabs | Clone any voice from a short audio sample |
ELEVENLABS_API_KEY is all you need to get started. FAL_KEY is now optional.cassetteai-music by default, which completes synchronously.cassetteai-music as a stable alternative for music generation.© pexoai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 9 other files in skills/videoagent-audio-studio of pexoai/pexo-skills.
Open the folder on GitHubat commit f724267
Videoagent Audio Studio next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Videoagent Audio Studio this skillpexoai/pexo-skills | 804 | — | ~1.7k | Automated safety check: Pass | MIT | |
| Elevenlabs Audio Promptingnodetool-ai/nodetool | 560 | — | ~1.9k | Automated safety check: Pass | AGPL-3.0 | |
| Musictadaspetra/loop | 296 | 2 repos | ~827 | Automated safety check: Pass | MIT | |
| ElevenLabs Voiceover Generatordigitalsamba/claude-code-video-toolkit | 2.2k | 1 repos | ~2.7k | Automated safety check: Notes | MIT | |
| Sound Effectstadaspetra/loop | 296 | 2 repos | ~1.1k | Automated safety check: Pass | MIT | |
| Hyperframes Mediachmonitor/chmonitor | 300 | 1 repos | ~2.8k | Automated safety check: Notes | GPL-3.0 |
nodetool-ai/nodetool
Direct ElevenLabs speech, dialogue, sound effects and music — the bracketed audio tags v3 acts on and why the voice decides whether a tag lands, stability as the delivery dial, punctuation instead…
tadaspetra/loop
Generate music using ElevenLabs Music API. An agent skill from tadaspetra/loop.
digitalsamba/claude-code-video-toolkit
Generates narration, sound effects and cloned voices through the ElevenLabs API, with model and setting choices tuned to the content's style.
tadaspetra/loop
Generate sound effects from text descriptions using ElevenLabs.
chmonitor/chmonitor
Audio and media assets for HyperFrames compositions, produced by one shared audio engine (scripts/audio.mjs) — multi-provider TTS (HeyGen / ElevenLabs / Kokoro local), background music + sound…
mikeOnBreeze/cc-crossbeam
This skill enables AI video generation from images AND text-to-speech voiceover generation using Fal.ai's API.
pexoai/pexo-skills
Generate AI video from any input — text, image, or script — with Pexo.
pexoai/pexo-skills
Create an explainer video with narration using Pexo. An agent skill from pexoai/pexo-skills.
pexoai/pexo-skills
Make a founder video with Pexo — built for solo founders and small teams.
pexoai/pexo-skills
Animate a still image into a finished, moving video with Pexo.
pexoai/pexo-skills
Make a launch video for your startup or product with Pexo. An agent skill from pexoai/pexo-skills.
pexoai/pexo-skills
Make a complete video from a simple idea with Pexo. An agent skill from pexoai/pexo-skills.
Works with
Categories
Tired of juggling multiple audio APIs?. An agent skill from pexoai/pexo-skills. Videoagent Audio Studio is an agent skill from pexoai/pexo-skills. Tired of juggling multiple audio APIs?
Videoagent Audio Studio fits situations like: you want to generate any audio without managing multiple API keys; tasks that involve Text to speech and voice; tasks that involve Music and audio generation.
Run `npx skills add pexoai/pexo-skills --skill videoagent-audio-studio -a claude-code`. Or copy the skill folder (skills/videoagent-audio-studio in pexoai/pexo-skills) into .claude/skills/videoagent-audio-studio in your project. Claude Code loads it when a task matches its description.
Run `npx skills add pexoai/pexo-skills --skill videoagent-audio-studio -a codex`. Or copy the skill folder (skills/videoagent-audio-studio in pexoai/pexo-skills) into .agents/skills/videoagent-audio-studio in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add pexoai/pexo-skills --skill videoagent-audio-studio -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/videoagent-audio-studio, .gemini/skills/videoagent-audio-studio, .github/skills/videoagent-audio-studio and .opencode/skills/videoagent-audio-studio in your project.
Going by SKILL.md and its folder, Videoagent Audio Studio needs JavaScript and a shell for the scripts in its folder, the command-line tools its instructions call (bash, npm and vercel) and credentials named ELEVENLABS_API_KEY and FAL_KEY. Our summary lists: Node.js; A Bash shell; A credential in ELEVENLABS_API_KEY; A credential in FAL_KEY.
SKILL.md names 2 domains. As links in the text: elevenlabs.io and fal.ai. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Videoagent Audio Studio is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.7k tokens (SKILL.md is roughly 6.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Videoagent Audio Studio: Elevenlabs Audio Prompting (nodetool-ai/nodetool, 560 stars), Music (tadaspetra/loop, 296 stars), ElevenLabs Voiceover Generator (digitalsamba/claude-code-video-toolkit, 2.2k stars) and Sound Effects (tadaspetra/loop, 296 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
pexoai (a GitHub user) maintains it in pexoai/pexo-skills, which has 804 GitHub stars. The repository holds 18 skills in this directory. The repository was last updated on August 20, 2026.
Source: pexoai/pexo-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.