Agent skill

Vellum Sounds

by vellum-ai in vellum-ai/vellum-assistant

Customize the desktop app's sound effects by adding sound files to the workspace, enabling sounds, setting volume, and assigning sounds to app events

MITAuto-check passedMedia & Creative

Install Vellum Sounds

skills CLI
$ npx skills add vellum-ai/vellum-assistant --skill vellum-sounds -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install vellum-ai/vellum-assistant vellum-sounds --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/vellum-ai/vellum-assistant.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/vellum-sounds .claude/skills/vellum-sounds && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
vellum-sounds
GitHub stars
1.4k
Token cost
~2.9k tokens
SKILL.md length
1,116 words
Files
2 (incl. scripts)
Skills in repo
108
Repo updated
First seen
Licence
MIT

At a glance

Customize the desktop app's sound effects by adding sound files to the workspace, enabling sounds, setting volume, and assigning sounds to app events

  • Tasks that involve Music and audio generation
  • SKILL.md covers What you're configuring, Sound events, Mode 1: Inspect current state and Mode 2: Add a sound file, plus 4 more sections
  • Runs TypeScript scripts from its folder; calls bun

What it does

Vellum Sounds is an agent skill from vellum-ai/vellum-assistant. Customize the desktop app's sound effects by adding sound files to the workspace, enabling sounds, setting volume, and assigning sounds to app events

Its SKILL.md is about 2.9k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/update-config.ts`). Compatibility notes: Designed for Vellum personal assistants

It sits in Media & Creative, covering Music and audio generation. It works with macOS. The repository describes itself as: An AI Assistant that’s easy to setup, does your work 24/7, knows your preferences and gets better over time. The licence is MIT.

When your agent uses it

  • Tasks that involve Music and audio generation

Example prompts

  • “/vellum-sounds”

Requirements

  • Node.js
  • Compatibility (from SKILL.md): Designed for Vellum personal assistants

What it can do on your machine

Read from SKILL.md and the folder at commit 33cc983. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (TypeScript), which the agent can run.

    Shell commands in SKILL.md call:

    • bun

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

  • Compatibility

    Designed for Vellum personal assistants

    From compatibility in the SKILL.md frontmatter.

Context cost

Vellum Sounds loads about 2.9k tokens when it runs. Until then it costs about 41 tokens; SKILL.md has 1,116 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~41
When it runs · the whole SKILL.md, loaded when a task matches
~2.9k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from vellum-ai/vellum-assistant at commit 33cc983, republished under its MIT licence (© vellum-ai). 1,116 words, ~2,889 tokens.

Download SKILL.mdSave it as .claude/skills/vellum-sounds/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
vellum-sounds
description
Customize the desktop app's sound effects by adding sound files to the workspace, enabling sounds, setting volume, and assigning sounds to app events
compatibility
Designed for Vellum personal assistants
metadata.emoji
🔊

You are helping the user customize the sound effects their desktop app plays. Sounds are configured in two places: a data/sounds/ directory of audio files and a data/sounds/config.json that controls what plays, at what volume, and whether it is enabled. The Settings > Sounds tab reads the same files on macOS and Windows, so changes appear there live without a restart.

All commands in this skill use the bash tool. $VELLUM_WORKSPACE_DIR is available in the sandbox environment — do not use host_bash.

What you're configuring

Two stores, both under $VELLUM_WORKSPACE_DIR/data/sounds/:

  • Sound files: .aiff, .wav, .mp3, .m4a, or .caf. No other extensions are accepted. The app scans this directory to populate the dropdown for each event. Prefer .wav, .mp3, or .m4a when the same file must play on both macOS and Windows.
  • config.json: a single JSON file that stores the global on/off switch, the master volume, and a per-event map of { enabled, sounds }. Each event's sounds is a pool of filenames; the app picks one at random on playback. An empty pool falls back to the platform's default blip.

Sound events

macOS supports all nine events below. Windows supports the eight events other than app_open. Do not configure app_open on Windows because it is not displayed or played there. Other keys are ignored by the app.

Event keyFires when
app_openApp launches (macOS only)
task_completeConversation transitions processing → idle
needs_inputConversation enters waiting-for-input
task_failedConversation enters error state
notificationA tool-triggered notification is sent
new_conversationUser creates a new conversation
message_sentUser sends a message in the composer
character_pokeUser clicks the avatar
randomAmbient timer (fires every 5–30 minutes)

Mode 1: Inspect current state

Always check current state before making changes — the user may already have things configured.

bash
ls "$VELLUM_WORKSPACE_DIR/data/sounds/" 2>/dev/null || echo "No sounds directory yet"
cat "$VELLUM_WORKSPACE_DIR/data/sounds/config.json" 2>/dev/null || echo "No config yet"

Report back what's there: whether sounds are globally enabled, current volume, which events have custom sounds assigned.

Mode 2: Add a sound file

The user either sends you an audio file or asks you to fetch/generate one. Copy it into data/sounds/ with a clean filename:

bash
mkdir -p "$VELLUM_WORKSPACE_DIR/data/sounds"
cp "<source-path>" "$VELLUM_WORKSPACE_DIR/data/sounds/<descriptive-name>.<ext>"

Rules:

  • Extension must be one of: aiff, wav, mp3, m4a, caf. If the user's file is something else (e.g. .ogg, .flac), tell them — don't try to rename.
  • Keep the filename simple (no path separators, no leading dots). Spaces are fine.
  • After adding a file, it's available in the dropdown — but nothing plays until you assign it to an event (Mode 3).

Mode 3: Configure via the helper script

Use scripts/update-config.ts to edit config.json. It validates inputs, creates the file with defaults if missing, and writes atomically so a crash can't corrupt it. If the existing file uses the legacy single-sound shape ("sound": "foo.wav"), the script normalizes it to the new pool shape ("sounds": ["foo.wav"]) on the next write.

bash
bun run scripts/update-config.ts --global-enabled true
bun run scripts/update-config.ts --volume 0.5
bun run scripts/update-config.ts --event message_sent --enabled true --sound "gentle-ding.aiff"
bun run scripts/update-config.ts --event random --enabled false
bun run scripts/update-config.ts --event task_complete --sound null   # clear the pool, revert to default blip
Mode 3a: Sound pools

Each event can hold one or more sounds. When the event fires, the app picks one entry at random from the pool. This is how you build variety, such as three different "poke" sounds that rotate when the user clicks the avatar. An empty pool falls back to the platform's default blip.

bash
# Replace the pool with three sounds
bun run scripts/update-config.ts --event character_poke --sounds "poke1.wav,poke2.wav,poke3.wav"

# Append one more sound to the existing pool
bun run scripts/update-config.ts --event character_poke --add-sound "poke4.wav"

# Drop a specific entry
bun run scripts/update-config.ts --event character_poke --remove-sound "poke2.wav"

# Empty the pool (back to the default blip)
bun run scripts/update-config.ts --event character_poke --clear-sounds

--sound is retained as a convenience for the common single-sound case: it replaces the whole pool with one entry (or clears it, when given null). Use --sounds / --add-sound / --remove-sound / --clear-sounds for pool edits.

Only one pool-mutation flag is allowed per invocation — mixing --sound and --add-sound (or any other pair) is rejected with a clear error. The one exception is --add-sound, which may be passed multiple times to append several filenames in a single run.

Flag reference:

FlagValueEffect
--global-enabledtrue or falseMaster switch. If false, NOTHING plays regardless of per-event settings.
--volume0.0–1.0 (clamped)Master volume. 0.7 is the default.
--eventa supported key aboveScopes the next flags to a single event.
--enabledtrue or falsePer-event on/off (requires --event).
--soundfilename or nullSingle-sound convenience (requires --event). Replaces the entire pool with one entry, or clears it when given null. The file must already exist in data/sounds/.
--soundscomma-separated filenamesReplaces the pool with the given list (requires --event). Every filename must already exist in data/sounds/. Use --clear-sounds to empty.
--add-soundfilenameAppends one filename to the pool (requires --event). No-op with a warning if already present. May be repeated in a single invocation.
--remove-soundfilenameRemoves one filename from the pool (requires --event). No-op with a warning if not present.
--clear-sounds—Empties the pool (requires --event).

The script prints the resulting config slice so you can confirm what changed.

Show full SKILL.md (376 more words)Show less

Mode 4: Remove a sound file

bash
rm "$VELLUM_WORKSPACE_DIR/data/sounds/<filename>"

Then remove it from any event pool that referenced it, so the config doesn't dangle:

bash
# If the file was one entry in a larger pool:
bun run scripts/update-config.ts --event <key> --remove-sound "<filename>"

# If the event only had that one sound (or you want to reset entirely):
bun run scripts/update-config.ts --event <key> --clear-sounds

(The app already falls back to the platform's default blip if every referenced file is missing, but cleaning up the config is tidier.)

UX Guidelines

  • Always check current state first. Don't ask "what do you want to do" if they already have sounds configured — summarize what's set up, then ask what to change.
  • The master switch is the #1 gotcha. globalEnabled defaults to false. If the user assigns a sound to an event and doesn't hear anything, check that flag first. When assigning the user's first sound, offer to flip the master switch on for them.
  • Per-event enabled is the #2 gotcha. Each event has its own enabled bool. Setting a sound alone doesn't enable the event.
  • Pool editing in the UI. The Settings > Sounds tab supports pool editing on macOS and Windows. Users can add and remove entries there without running this script. Power users can weight a sound more heavily by hand-editing config.json to include duplicates. For example, ["a.wav","a.wav","b.wav"] makes a.wav twice as likely. The script de-dupes on --add-sound but does not re-sort or de-dupe on read, so hand-edited duplicates survive round-trips.
  • Filename sanity. When the user sends a file named something like Screen Recording 2026-04-13 at 11.47.23.m4a, rename it to something memorable before copying — they'll have to pick it from a dropdown later.
  • Confirm after changes. Tell the user the Settings > Sounds tab will reflect changes live. Offer to open it: "You can preview it in Settings > Sounds, or I can play it for you next time that event fires."
  • Don't invent events. macOS supports the nine keys above, while Windows supports eight and excludes app_open. There is currently no event for voice-mode activation or typing indicators. If the user asks for those, tell them it needs a desktop app code change.

Config shape reference

If the user inspects config.json directly, this is the macOS superset they may see. On Windows, omit the app_open entry. Defaults otherwise match the shared desktop sound configuration.

json
{
  "globalEnabled": false,
  "volume": 0.7,
  "events": {
    "app_open": { "enabled": false, "sounds": [] },
    "task_complete": { "enabled": false, "sounds": [] },
    "needs_input": { "enabled": false, "sounds": [] },
    "task_failed": { "enabled": false, "sounds": [] },
    "notification": { "enabled": false, "sounds": [] },
    "new_conversation": { "enabled": false, "sounds": [] },
    "message_sent": { "enabled": false, "sounds": [] },
    "character_poke": { "enabled": false, "sounds": [] },
    "random": { "enabled": false, "sounds": [] }
  }
}

Legacy {"sound": "foo.wav"} entries are still accepted on read (the macOS decoder and this script both normalize them into {"sounds": ["foo.wav"]}), but new writes always use the pool shape.

© vellum-ai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (scripts) in skills/vellum-sounds of vellum-ai/vellum-assistant.

  • SKILL.md
  • scripts/update-config.ts

Open the folder on GitHubat commit 33cc983

Compare with similar skills

Vellum Sounds next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Vellum Sounds compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Vellum Sounds this skillvellum-ai/vellum-assistant1.4k—~2.9kAutomated safety check: PassMIT
Persona Character Scenesmiuuyy/persona-voice133—~953Automated safety check: PassMIT
Mlx Serveddalcu/mlx-serve1.8k—~1kAutomated safety check: PassCustom licence
Musictadaspetra/loop2962 repos~827Automated safety check: PassMIT
Sound Effectstadaspetra/loop2962 repos~1.1kAutomated safety check: PassMIT
Kkclawkk43994/kkclaw174—~576Automated safety check: PassMIT

Similar skills

  • Persona Character Scenes

    miuuyy/persona-voice

    Generate, review, and integrate consistent 2D or stylized 3D character background scenes for Codex Persona Voice session cards.

    133 GitHub stars~953 tokensUpdated 19 days ago
    Media & CreativeAuto-check passed
  • Mlx Serve

    ddalcu/mlx-serve

    Hook an app, game or script up to the local mlx-serve server for LLM chat, embeddings, image, speech, music, sound effect, video and 3D generation, and Laya/Kev/Clef typed decisions.

    1.8k GitHub stars~1k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Music

    tadaspetra/loop

    Generate music using ElevenLabs Music API. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 2 repos~827 tokens
    Media & CreativeAuto-check passed
  • Sound Effects

    tadaspetra/loop

    Generate sound effects from text descriptions using ElevenLabs.

    296 GitHub starsUsed in 2 repos~1.1k tokens
    Media & CreativeAuto-check passed
  • Kkclaw

    kk43994/kkclaw

    给你的 AI Agent 一个桌面身体 — Setup Wizard、14情绪球体、语音克隆、歌词窗、Doctor 自检、跨平台支持(Windows + macOS)

    174 GitHub stars~576 tokensUpdated 6 mo ago
    Media & CreativeAuto-check passed
  • Bench

    ddalcu/mlx-serve

    mlx-serve benchmarking methodology — bench.sh/llmprobe usage, comparison-trap rules (same-methodology cells only, spec-decode variance, thermal lies, engine naming), perf-claim etiquette.

    1.8k GitHub starsUsed in 1 repo~1.1k tokens
    Media & CreativeAuto-check passed

More from vellum-ai/vellum-assistant

All 108 skills in this repo
  • Vellum GitHub App Setup

    vellum-ai/vellum-assistant

    Create and configure a GitHub App so the assistant can push commits, open PRs, and comment under its own bot identity.

    1.4k GitHub stars~3.1k tokensUpdated yesterday
    Auto-check passed
  • Discord App Setup

    vellum-ai/vellum-assistant

    Connect a Discord bot to the assistant via the Discord Gateway with guided application creation and intent configuration

    1.4k GitHub stars~4.2k tokensUpdated yesterday
    Auto-check passed
  • Sentry App Setup

    vellum-ai/vellum-assistant

    Create and configure a Sentry internal integration so the assistant can manage issues, alerts, and releases under its own identity

    1.4k GitHub stars~1.3k tokensUpdated yesterday
    Auto-check passed
  • Memory Corpus Ingest

    vellum-ai/vellum-assistant

    Ingest a large dataset into memory as a skimmed map. An agent skill from vellum-ai/vellum-assistant.

    1.4k GitHub stars~3k tokensUpdated yesterday
    Auto-check: notes
  • Plugin Builder

    vellum-ai/vellum-assistant

    A skill your agent uses when the user wants to build, scaffold, ship, or edit a Vellum plugin that bundles multiple surfaces (hooks, tools, skills, and more) into one installable package.

    1.4k GitHub stars~3.1k tokensUpdated yesterday
    Auto-check passed
  • Slack App Setup

    vellum-ai/vellum-assistant

    Connect a Slack app to the Vellum Assistant via Socket Mode.

    1.4k GitHub stars~2.5k tokensUpdated yesterday
    Auto-check: warnings

Works with

Questions about Vellum Sounds

What does Vellum Sounds do?

Customize the desktop app's sound effects by adding sound files to the workspace, enabling sounds, setting volume, and assigning sounds to app events. Vellum Sounds is an agent skill from vellum-ai/vellum-assistant.

When should I use Vellum Sounds?

Vellum Sounds fits situations like: tasks that involve Music and audio generation.

How do I install Vellum Sounds in Claude Code?

Run `npx skills add vellum-ai/vellum-assistant --skill vellum-sounds -a claude-code`. Or copy the skill folder (skills/vellum-sounds in vellum-ai/vellum-assistant) into .claude/skills/vellum-sounds in your project. Claude Code loads it when a task matches its description.

How do I install Vellum Sounds in Codex?

Run `npx skills add vellum-ai/vellum-assistant --skill vellum-sounds -a codex`. Or copy the skill folder (skills/vellum-sounds in vellum-ai/vellum-assistant) into .agents/skills/vellum-sounds in your project. Codex loads it when a task matches its description.

Can I use Vellum Sounds in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add vellum-ai/vellum-assistant --skill vellum-sounds -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/vellum-sounds, .gemini/skills/vellum-sounds, .github/skills/vellum-sounds and .opencode/skills/vellum-sounds in your project.

What does Vellum Sounds need to run?

Going by SKILL.md and its folder, Vellum Sounds needs TypeScript for the scripts in its folder and the command-line tools its instructions call (bun). Our summary lists: Node.js. Compatibility (from SKILL.md): Designed for Vellum personal assistants.

Does Vellum Sounds access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Vellum Sounds safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Vellum Sounds use?

Vellum Sounds is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Vellum Sounds use?

About 2.9k tokens (SKILL.md is roughly 12k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Vellum Sounds?

Skills that share tags, products or a category with Vellum Sounds: Persona Character Scenes (miuuyy/persona-voice, 133 stars), Mlx Serve (ddalcu/mlx-serve, 1.8k stars), Music (tadaspetra/loop, 296 stars) and Sound Effects (tadaspetra/loop, 296 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Vellum Sounds?

vellum-ai (a GitHub organization) maintains it in vellum-ai/vellum-assistant, which has 1,408 GitHub stars. The repository holds 108 skills in this directory. The repository was last updated on October 9, 2026.

Source: vellum-ai/vellum-assistant on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.