Agent skill

Scenario Sonilo

by scenario-labs in scenario-labs/skills

A skill your agent uses when adding sound to a video or generating standalone audio with Sonilo models on Scenario via MCP: text-to-sound-effects, text-to-music, video-to-sound-effects or…

MITAuto-check passedMedia & Creative

Install Scenario Sonilo

skills CLI
$ npx skills add scenario-labs/skills --skill scenario-sonilo -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install scenario-labs/skills scenario-sonilo --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/scenario-labs/skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/scenario-sonilo .claude/skills/scenario-sonilo && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
scenario-sonilo
GitHub stars
946
Token cost
~1.6k tokens
SKILL.md length
793 words
Files
1
Skills in repo
146
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses when adding sound to a video or generating standalone audio with Sonilo models on Scenario via MCP: text-to-sound-effects, text-to-music, video-to-sound-effects or…

  • Works in 6 steps: search with target="models",… → model_schema_get with that id: fields… → upload_asset the clip (see the scenario… → …
  • Adding sound to a video
  • SKILL.md covers Overview, Quick reference, The video is the clock and What survives a music pass, plus 2 more sections
  • Calls npx

What it does

Scenario Sonilo is an agent skill from scenario-labs/skills. Use when adding sound to a video or generating standalone audio with Sonilo models on Scenario via MCP: text-to-sound-effects, text-to-music, video-to-sound-effects or video-to-music from a clip, scoring silent AI video, foley, ambience, muxed video-to-video variants returning the clip with the new track mixed in, keeping speech while replacing music, or per-segment sound control. Keywords: Sonilo V1.1, SFX, soundtrack, score, keepSpeechVocal, segments, royalty-free.

Its SKILL.md is about 1.6k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering Music and audio generation and AI video generation. It works with Model Context Protocol. The repository describes itself as: Get production-ready images, video, audio, and 3D from any AI agent: skills that pick the right model, price before spending, and keep characters and brands consistent through… The licence is MIT.

When your agent uses it

  • Adding sound to a video
  • Generating standalone audio with Sonilo models on Scenario via MCP: text-to-sound-effects
  • Video-to-sound-effects
  • Video-to-music from a clip

Example prompts

  • “/scenario-sonilo”

Requirements

  • Node.js

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. search with target="models", query="sonilo", public=true. Members list txt2audio, video2audio, or video2video capabilities; e.g…
  2. model_schema_get with that id: fields and caps before anything else.
  3. upload_asset the clip (see the scenario skill) to get its asset id.
  4. model_run with that model_id, dry_run=true, and parameters={"video": "asset_x", "segments": [{"start": 0, "end": 4, "prompt": "footsteps…
  5. Repeat model_run with wait=false, then jobs_wait with the returned job id, re-called with pending_job_ids on timeout, never a second…
  6. asset_display the muxed video, then asset_download it and the separate track.

What it can do on your machine

Read from SKILL.md and the folder at commit f6f8ab7. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • npx

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use npx, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Scenario Sonilo loads about 1.6k tokens when it runs. Until then it costs about 122 tokens; SKILL.md has 793 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~122
When it runs · the whole SKILL.md, loaded when a task matches
~1.6k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from scenario-labs/skills at commit f6f8ab7, republished under its MIT licence (© scenario-labs). 793 words, ~1,566 tokens.

Download SKILL.mdSave it as .claude/skills/scenario-sonilo/SKILL.md (or your agent's skills folder).
name
scenario-sonilo
description
Use when adding sound to a video or generating standalone audio with Sonilo models on Scenario via MCP: text-to-sound-effects, text-to-music, video-to-sound-effects or video-to-music from a clip, scoring silent AI video, foley, ambience, muxed video-to-video variants returning the clip with the new track mixed in, keeping speech while replacing music, or per-segment sound control. Keywords: Sonilo V1.1, SFX, soundtrack, score, keepSpeechVocal, segments, royalty-free.
license
MIT

Scenario Sonilo Audio

Overview

Sonilo, an audio family on Scenario, splits along two axes: what the sound is (sound effects or instrumental music) and what comes back (a standalone track, or the source video with the generated track muxed in, visuals untouched). Picking the member whose return matches the delivery matters more than any prompt. Discover with search and treat model_schema_get as the contract.

Connection and the core loop: see the scenario skill in this repo; model-agnostic audio work: the scenario-audio skill. If a sibling skill named here is missing from your available skills, ask the user to install it (npx skills add scenario-labs/skills --skill <name>); unattended, proceed from tool schemas and flag the gap.

Quick reference

Member names state the mode; input names come from the live schemas:

MemberKey inputsReturns
Text to SFXprompt, duration, audioFormatone sound effect clip
Text to Musicprompt, duration, numSamplesinstrumental track
Video to SFXvideo, optional prompt or segmentsaudio track matching the video length
Video to Musicvideo, optional prompt, numSamples, startOffset, durationmusic track
Video to Video SFXas Video to SFXthe video with SFX mixed in, plus the track
Video to Video (music)as Video to Music, plus keepSpeechVocalthe video with music mixed in

At authoring time: text SFX ran 1 to 180 seconds (default 8), text music 1 to 600 (default 90), video inputs took up to 360 seconds, segments up to 50, numSamples 1 to 3, and startOffset moved in steps of 10 with startOffset plus duration capped at the video length. audioFormat (aac default, mp3, wav, flac; wav or flac for editing pipelines) exists on the SFX members only, and on the muxed one it formats the separate track: the video's own audio stays AAC. Two music mux members were live with the same schema; prefer the newest hit. Every duration knob carries cost, so dry_run before a batch.

The video is the clock

On video-conditioned members the footage decides when sound happens and the prompt only steers what it sounds like; the clip's own audio never steers generation, visuals alone are read. An empty prompt is valid and often best: the model captions the clip and covers each scene itself. When one description cannot fit the whole clip, segments gives per-range prompts, contiguous by contract: the first start is 0, each end equals the next start, and the last stays within the video. Text to SFX has no timing control at all, so describe one sound event per run, physical words over moods (hollow, muffled, punchy), source then material then space then intensity, with the duration intent in the wording as well as the parameter. No member takes a seed, so archive the takes you like; on the music members numSamples buys up to three takes in one run instead of re-rolls.

Show full SKILL.md (322 more words)Show less

What survives a music pass

The music mux members replace the whole original track by default. keepSpeechVocal: true isolates human voice (dialogue, narration, singing, crowds) and ducks the music under it; everything non-voice (engines, footsteps, ambience, prior foley) is replaced regardless. So never score a clip after muxing foley into it: when a video needs both, take standalone tracks from Video to SFX and Video to Music and mix in post. Music comes back instrumental.

Worked example: foley for a silent gameplay clip

  1. search with target="models", query="sonilo", public=true. Members list txt2audio, video2audio, or video2video capabilities; e.g. model_sonilo-v1-1-video-to-video-sound-effects (a live hit at authoring time: re-discover each session).
  2. model_schema_get with that id: fields and caps before anything else.
  3. upload_asset the clip (see the scenario skill) to get its asset id.
  4. model_run with that model_id, dry_run=true, and parameters={"video": "asset_x", "segments": [{"start": 0, "end": 4, "prompt": "footsteps on wet metal, close and sharp"}, {"start": 4, "end": 12, "prompt": "plasma rifle shots, hollow hangar reverb"}]}: cost scales with clip length.
  5. Repeat model_run with wait=false, then jobs_wait with the returned job id, re-called with pending_job_ids on timeout, never a second model_run.
  6. asset_display the muxed video, then asset_download it and the separate track.

Common mistakes

  • Muxing when the edit needs a bare track, or the reverse: Video to SFX returns audio only; Video to Video SFX returns the finished clip. Pick by delivery.
  • Scoring a clip that already carries foley with a music mux member: non-voice sound is replaced and the foley is gone.
  • Leaving keepSpeechVocal off (the default) on a talking-head clip: the narration vanishes with the rest of the track.
  • Gapped or overlapping segments: the contract wants contiguous ranges from 0.
  • Packing sequenced events into one Text to SFX prompt: it has no timing control; one event per run.
  • Asking any member for speech or vocals: no member produces dialogue, narration, or singing; SFX and instrumental music are the whole surface.

© scenario-labs, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/scenario-sonilo of scenario-labs/skills.

Open the folder on GitHubat commit f6f8ab7

Compare with similar skills

Scenario Sonilo next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Scenario Sonilo compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Scenario Sonilo this skillscenario-labs/skills946—~1.6kAutomated safety check: PassMIT
BlockrunBlockRunAI/blockrun-mcp391—~2.7kAutomated safety check: PassMIT
Videoguaardvark/guaardvark257—~1.2kAutomated safety check: PassMIT
Motion Graphics Music Videomakevoid/motion-graphics-music-video-skill143—~3.3kAutomated safety check: PassMIT
Music VideoNoizAI/skills526—~1.5kAutomated safety check: PassNone
WorkrallyTencent/workrally166—~3.7kAutomated safety check: PassMIT-0

Similar skills

  • Blockrun

    BlockRunAI/blockrun-mcp

    Pay-per-call access to AI models, real-time data, media generation and multi-chain RPC over x402 micropayments (USDC on Base or Solana), or a BlockRun account API key.

    391 GitHub stars~2.7k tokensUpdated 2 days ago
    Media & CreativeAuto-check passed
  • Video

    guaardvark/guaardvark

    Generate video clips on the user's own GPU through Guaardvark: text-to-video, image-to-video, first+last frame animation, clips with their own soundtrack and dialogue (MiniMax H3), short looping…

    257 GitHub stars~1.2k tokensUpdated today
    Media & CreativeAuto-check passed
  • Motion Graphics Music Video

    makevoid/motion-graphics-music-video-skill

    Create and revise animated character music videos from a supplied song and creative prompt, with researched storyboards, Fal character and video generation, p5 motion graphics, audio editing, and…

    143 GitHub stars~3.3k tokensUpdated 5 days ago
    Media & CreativeAuto-check passed
  • Music Video

    NoizAI/skills

    Build a music video from an existing song and lyrics, either as p5.js/p5.brush animation or by assembling local images and clips.

    526 GitHub stars~1.5k tokensUpdated 12 days ago
    Media & CreativeAuto-check passed
  • Workrally

    Tencent/workrally

    WorkRally CLI (workrally) — 面向 AI Agent 的 AIGC 漫剧视频创作全流程工具集。支持 AI 生图、AI 生视频、视频提示词优化、画布生音频/音乐、混元 3D 模型生成、AI 生音频、项目/剧集/场次/分镜的完整 CRUD、资产库、媒资管理、无限画布、文件上传下载等。Use when user asks to generate images…

    166 GitHub stars~3.7k tokensUpdated 11 days ago
    Media & CreativeAuto-check passed
  • Runninghub

    HM-RunningHub/OpenClaw_RH_Skills

    Generate images, videos, audio, and 3D models via RunningHub API (420 endpoints) and run any RunningHub AI Application (custom ComfyUI workflow) by webappId.

    142 GitHub stars~1.6k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed

More from scenario-labs/skills

All 146 skills in this repo
  • Scenario Blender Grease Pencil

    scenario-labs/skills

    A skill your agent uses when drawing or animating with Grease Pencil in Blender 5.x from Python: 2D or 2.5D illustration, frame-by-frame animation, a cutout or part-based 2D character, strokes with…

    946 GitHub stars~4.5k tokensUpdated today
    Auto-check passed
  • Scenario Blender Hair

    scenario-labs/skills

    A skill your agent uses when grooming hair or fur in Blender with hair curves, such as a character hairstyle, animal fur, procedural fur in geometry nodes, or hair cards and mesh hair for games.

    946 GitHub stars~4.7k tokensUpdated today
    Auto-check passed
  • A skill your agent uses when lighting, rendering or compositing in Blender: light a character, product or hero shot, interior at dusk or night, three-point or motivated lighting, sun and sky, HDRI…

    946 GitHub stars~5k tokensUpdated today
    Auto-check passed
  • Scenario Chatgpt Pet Create

    scenario-labs/skills

    A skill your agent uses when creating a ChatGPT pet or Codex pet with Scenario: hatching an animated companion from a text idea, a character, mascot or brand cue, or reference photos and art; making…

    946 GitHub stars~3.6k tokensUpdated today
    Auto-check passed
  • Scenario Godot Animation

    scenario-labs/skills

    A skill your agent uses when animating characters or scenes in Godot 4.7: AnimationPlayer clips and RESET, AnimationTree state machines and blend spaces built in code, Mixamo or glTF import, loop…

    946 GitHub stars~4.7k tokensUpdated today
    Auto-check passed
  • Scenario Godot Audio

    scenario-labs/skills

    A skill your agent uses when adding or fixing sound in Godot 4.7: audio buses and effects, volume sliders, 'too many sounds', combat audio with hundreds of enemies, sounds clipping or distorting, 3D…

    946 GitHub stars~4.7k tokensUpdated today
    Auto-check passed

Questions about Scenario Sonilo

What does Scenario Sonilo do?

A skill your agent uses when adding sound to a video or generating standalone audio with Sonilo models on Scenario via MCP: text-to-sound-effects, text-to-music, video-to-sound-effects or…. Scenario Sonilo is an agent skill from scenario-labs/skills. Use when adding sound to a video or generating standalone audio with Sonilo models on Scenario via MCP: text-to-sound-effects, text-to-music, video-to-sound-effects or video-to-music from a clip, scoring silent AI video, foley, ambience, muxed video-to-video variants returning the clip with the new track mixed in, keeping speech while replacing music, or per-segment sound control.

When should I use Scenario Sonilo?

Scenario Sonilo fits situations like: adding sound to a video; generating standalone audio with Sonilo models on Scenario via MCP: text-to-sound-effects; video-to-sound-effects; video-to-music from a clip.

How do I install Scenario Sonilo in Claude Code?

Run `npx skills add scenario-labs/skills --skill scenario-sonilo -a claude-code`. Or copy the skill folder (skills/scenario-sonilo in scenario-labs/skills) into .claude/skills/scenario-sonilo in your project. Claude Code loads it when a task matches its description.

How do I install Scenario Sonilo in Codex?

Run `npx skills add scenario-labs/skills --skill scenario-sonilo -a codex`. Or copy the skill folder (skills/scenario-sonilo in scenario-labs/skills) into .agents/skills/scenario-sonilo in your project. Codex loads it when a task matches its description.

Can I use Scenario Sonilo in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add scenario-labs/skills --skill scenario-sonilo -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/scenario-sonilo, .gemini/skills/scenario-sonilo, .github/skills/scenario-sonilo and .opencode/skills/scenario-sonilo in your project.

What does Scenario Sonilo need to run?

Going by SKILL.md and its folder, Scenario Sonilo needs the command-line tools its instructions call (npx). Our summary lists: Node.js.

Does Scenario Sonilo access the network?

SKILL.md contains no URLs. Its commands use npx, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Scenario Sonilo safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Scenario Sonilo use?

Scenario Sonilo is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Scenario Sonilo use?

About 1.6k tokens (SKILL.md is roughly 6.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Scenario Sonilo?

Skills that share tags, products or a category with Scenario Sonilo: Blockrun (BlockRunAI/blockrun-mcp, 391 stars), Video (guaardvark/guaardvark, 257 stars), Motion Graphics Music Video (makevoid/motion-graphics-music-video-skill, 143 stars) and Music Video (NoizAI/skills, 526 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Scenario Sonilo?

scenario-labs (a GitHub organization) maintains it in scenario-labs/skills, which has 946 GitHub stars. The repository holds 146 skills in this directory. The repository was last updated on October 10, 2026.

Source: scenario-labs/skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.