Agent skill

Render Cinematic Music Video

by gooseworks-ai in gooseworks-ai/goose-skills

Assemble a cinematic live-action-style music-video ad from a config — an original sung anthem carries the whole narrative while N 35mm-film-look i2v clips are each cut to their lyric window and…

MITAuto-check passedMedia & Creative

Install Render Cinematic Music Video

skills CLI
$ npx skills add gooseworks-ai/goose-skills --skill render-cinematic-music-video -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install gooseworks-ai/goose-skills render-cinematic-music-video --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/gooseworks-ai/goose-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/ads/capabilities/render-cinematic-music-video .claude/skills/render-cinematic-music-video && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
render-cinematic-music-video
GitHub stars
1.2k
Token cost
~1.3k tokens
SKILL.md length
613 words
Files
6 (incl. scripts)
Skills in repo
273
Repo updated
First seen
Licence
MIT

At a glance

Assemble a cinematic live-action-style music-video ad from a config — an original sung anthem carries the whole narrative while N 35mm-film-look i2v clips are each cut to their lyric window and…

  • The cinematic-music-video format
  • SKILL.md covers Run, Choices and Contract (the free assembly)
  • Tasks that involve Image generation
  • Tasks that involve Text to speech and voice

What it does

Render Cinematic Music Video is an agent skill from gooseworks-ai/goose-skills. Assemble a cinematic live-action-style music-video ad from a config — an original sung anthem carries the whole narrative while N 35mm-film-look i2v clips are each cut to their lyric window and hard-concatenated on the beat as a 3-act arc, the anthem muxed at loudnorm I=-14, cinematic lower-third serif captions built from the song's OWN word timings (never Whisper) with the hook line landing on the chorus drop, and closed on a brand end card composited from the real asset — never AI-rendered text. This is the…

Its SKILL.md is about 1.3k tokens, which your agent loads only when the skill is triggered. The skill folder holds 7 other files, including scripts (for example `scripts/PIPELINE.md`, `scripts/README.md` and `scripts/config.example.json`).

It sits in Media & Creative, covering Image generation, Text to speech and voice and Speech recognition and synthesis. It works with ElevenLabs. The repository describes itself as: Library of Growth & GTM skills + data APIs for Claude Code, Codex, Cursor to run ads, social, content, lead gen, seo and data scraping. The licence is MIT.

When your agent uses it

  • The cinematic-music-video format
  • Tasks that involve Image generation
  • Tasks that involve Text to speech and voice

Example prompts

  • “/render-cinematic-music-video”

What it can do on your machine

Read from SKILL.md and the folder at commit c650c6d. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 3 files in scripts/, which the agent can run.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Render Cinematic Music Video loads about 1.3k tokens when it runs. Until then it costs about 200 tokens; SKILL.md has 613 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~200
When it runs · the whole SKILL.md, loaded when a task matches
~1.3k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from gooseworks-ai/goose-skills at commit c650c6d, republished under its MIT licence (© gooseworks-ai). 613 words, ~1,284 tokens.

Download SKILL.mdSave it as .claude/skills/render-cinematic-music-video/SKILL.md (or your agent's skills folder). This skill also uses 5 other files; get the full folder from GitHub.
name
render-cinematic-music-video
description
Assemble a cinematic live-action-style music-video ad from a config — an original sung anthem carries the whole narrative while N 35mm-film-look i2v clips are each cut to their lyric window and hard-concatenated on the beat as a 3-act arc, the anthem muxed at loudnorm I=-14, cinematic lower-third serif captions built from the song's OWN word timings (never Whisper) with the hook line landing on the chorus drop, and closed on a brand end card composited from the real asset — never AI-rendered text. This is the FREE deterministic assembly stage (cut-to-window + hard concat + anthem mux + captions + end card); the anthem, keyframes, and clips come from create-music-elevenlabs / create-image-gpt-image-fal / create-video-fal. Use for the cinematic-music-video format.
status
active

render-cinematic-music-video

Assemble a cinematic music-video ad from a config: a live-action-STYLE short film where an original sung anthem is the score and every visual beat is a shot-on-film tableau (Kodak Portra grain, light leaks, natural light, handheld imperfection) timed to the lyrics, arranged as a 3-act arc (its shape is the user's story_arc choice) with the hook line on the chorus drop. This capability is the FREE, deterministic assembly — cut-to-window, hard-concat, anthem mux, caption burn, and the brand end card.

scripts/config.example.json is the worked example (Hype and Vice "Game Day Girls", ~28s 1080×1920 9:16, 14 tableaux) — copy its structure, never its creative values; scripts/PIPELINE.md maps every config block to its source step and scripts/README.md documents the free assembly.

Run

This is the FREE, deterministic assembly stage — it spends nothing. The paid inputs are separate capabilities: the sung anthem (create-music-elevenlabs, force_instrumental false — the lyrics ARE the script, returns mp3 + words_timestamps); one 35mm-film keyframe per beat in one look pack (create-image-gpt-image-fal); and one Kling 3.0 i2v clip per beat (create-video-fal). Given the delivered anthem + words.json + one clip per beat + the brand end-card asset, render-cinematic-music-video cuts each clip to its lyric window, hard-concats on the beat, muxes the anthem, burns the cinematic lower-third captions, and overlays the end card → the master. Re-cuts reuse the existing anthem / keyframes / clips and cost $0.

Choices

The creative content this stage assembles is decided upstream by the format's choices, asked of the user before any paid step. The worked example's values are examples, never defaults:

  • song_style — the anthem's genre, mood, BPM and arrangement. The demo used a triumphant 120-BPM indie-pop cinematic anthem.
  • vocalist — who sings it. The demo used a confident-but-warm female lead.
  • cast — who is on screen in the tableaux. The demo used four young college women.
  • setting — place + time of day (drives the look pack's light and palette). The demo used an American college town on game day, autumn, golden hour.
  • story_arc — the 3-act shape of the tableaux and lyrics. The demo used one game day, morning → stadium peak → twilight, with an origin-story wink.

This assembly reads them only through the config (tableaux[] windows + captions, the anthem, captions.accent_words, end_card.*); it hardcodes none of them.

Show full SKILL.md (254 more words)Show less

Contract (the free assembly)

  • The sung anthem carries the story — no separate VO. The generated ElevenLabs track IS the bed and the script (force_instrumental false); do not add a spoken voiceover or a second bed.
  • Plan the timeline AROUND the delivered anthem. The anthem is generated first and reshapes/overshoots length; snap every tableau boundary to the lyric-phrase edges in the returned word timings — never trim the anthem to a pre-planned grid.
  • Captions from the anthem's OWN word timings, not Whisper. Derive words.json from the music model's words_timestamps, chunk ~4 words at lyric boundaries, and burn cinematic lower-third serif captions (--placement low) with the hook line accent-treated (bold-italic). Whisper on sung audio returns "🎵 Music Playing 🎵".
  • Land the hook on the chorus drop. The hero tableau (is_hook) is timed so the load-bearing line sits on the chorus drop; accent that line in the captions.
  • One cinematic look pack + hard cuts on the beat. The look pack (named film stock + grain + flares + handheld) plus a continuity anchor holds every clip together; cut each clip to its lyric window and hard-concat (one optional match-cut into the hero reveal) — no dissolves.
  • End card from the real brand asset — never AI-render the wordmark as diffusion text. Composite the lockup (PIL/ffmpeg) from the real asset or use a designed keyframe base; diffusion garbles a wordmark.
  • FFmpeg composite, deterministic, FREE. Normalize each clip to its beat-locked window, hard-concat, mux the anthem (afade in/out + loudnorm I=-14), burn the caption ASS, overlay the end card → a 1080×1920 h264+aac master. No paid calls, no keys.

© gooseworks-ai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 5 other files (scripts) in skills/ads/capabilities/render-cinematic-music-video of gooseworks-ai/goose-skills.

  • SKILL.md
  • scripts/PIPELINE.md
  • scripts/README.md
  • scripts/config.example.json
  • skill.meta.json
  • tests/smoke-test.md

Open the folder on GitHubat commit c650c6d

Compare with similar skills

Render Cinematic Music Video next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Render Cinematic Music Video compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Render Cinematic Music Video this skillgooseworks-ai/goose-skills1.2k—~1.3kAutomated safety check: PassMIT
Local AI Useamd/skills408—~5kAutomated safety check: NotesMIT
Video Productionspeechlab0210/video-production-skill105—~4.1kAutomated safety check: NotesMIT
Voice AI Developmentdavila7/claude-code-templates33k5 repos~2.1kAutomated safety check: PassMIT
Local AI App Integrationamd/skills408—~6kAutomated safety check: PassMIT
Elevenlabs Core Workflow Bjeremylongshore/tons-of-skills-marketplace2.8k—~1.6kAutomated safety check: PassMIT

Similar skills

  • Local AI Use

    amd/skills

    Makes this agent generate images, transcribe audio, and synthesize speech on the user's own machine through a local Lemonade Server instead of a paid cloud API.

    408 GitHub stars~5k tokensUpdated yesterday
    Media & CreativeAuto-check: notes
  • Video Production

    speechlab0210/video-production-skill

    AI educational video production pipeline. An agent skill from speechlab0210/video-production-skill.

    105 GitHub stars~4.1k tokensUpdated 3 mo ago
    Media & CreativeAuto-check: notes
  • Voice AI Development

    davila7/claude-code-templates

    Expert in building voice AI applications - from real-time voice agents to voice-enabled apps.

    33k GitHub starsUsed in 5 repos~2.1k tokens
    Media & CreativeAuto-check passed
  • Integrates local AI capabilities into applications using Embeddable Lemonade.

    408 GitHub stars~6k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Elevenlabs Core Workflow B

    jeremylongshore/tons-of-skills-marketplace

    Implement ElevenLabs speech-to-speech, sound effects, audio isolation, and speech-to-text.

    2.8k GitHub stars~1.6k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Speech To Text

    tadaspetra/loop

    Transcribe audio to text using ElevenLabs Scribe v2. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 2 repos~2k tokens
    Media & CreativeAuto-check passed

More from gooseworks-ai/goose-skills

All 273 skills in this repo
  • Reddit Post Finder

    gooseworks-ai/goose-skills

    Scrape and search Reddit posts using Apify. An agent skill from gooseworks-ai/goose-skills.

    1.2k GitHub starsUsed in 1 repo~1.2k tokens
    Auto-check passed
  • Create Image Fal

    gooseworks-ai/goose-skills

    Generate or edit an image via any FAL image model (nano-banana edit, gpt-image, flux, ...), ROUTED THROUGH THE fal-proxy so it bills the Ads agent.

    1.2k GitHub stars~1.3k tokensUpdated today
    Auto-check passed
  • Render Hook Replacement

    gooseworks-ai/goose-skills

    Replace an existing video's opening with a supplied clip or free kinetic text hook while retaining and verifying every original body frame, audio, captions and ending.

    1.2k GitHub stars~2.3k tokensUpdated today
    Auto-check passed
  • Blog Feed Monitor

    gooseworks-ai/goose-skills

    Scrape blog posts via RSS feeds (free, no API key) with Apify fallback for JS-heavy sites.

    1.2k GitHub starsUsed in 1 repo~578 tokens
    Auto-check passed
  • Competitor Post Engagers

    gooseworks-ai/goose-skills

    Find leads by scraping engagers from a competitor's top LinkedIn posts.

    1.2k GitHub starsUsed in 1 repo~1.8k tokens
    Auto-check: notes
  • Render Chatgpt Chat

    gooseworks-ai/goose-skills

    Assemble a ChatGPT chat-reveal video ad from a thread + timeline JSON — one continuous Playwright recording of a ChatGPT mobile chat (user types with the iOS keyboard up → taps send → keyboard…

    1.2k GitHub stars~2.3k tokensUpdated today
    Auto-check passed

Works with

Questions about Render Cinematic Music Video

What does Render Cinematic Music Video do?

Assemble a cinematic live-action-style music-video ad from a config — an original sung anthem carries the whole narrative while N 35mm-film-look i2v clips are each cut to their lyric window and…. Render Cinematic Music Video is an agent skill from gooseworks-ai/goose-skills. Assemble a cinematic live-action-style music-video ad from a config — an original sung anthem carries the whole narrative while N 35mm-film-look i2v clips are each cut to their lyric window and hard-concatenated on the beat as a 3-act arc, the anthem muxed at loudnorm I=-14, cinematic lower-third serif captions built from the song's OWN word timings (never Whisper) with the hook line landing on the chorus drop, and closed on a brand end card composited from the real asset — never AI-rendered text.

When should I use Render Cinematic Music Video?

Render Cinematic Music Video fits situations like: the cinematic-music-video format; tasks that involve Image generation; tasks that involve Text to speech and voice.

How do I install Render Cinematic Music Video in Claude Code?

Run `npx skills add gooseworks-ai/goose-skills --skill render-cinematic-music-video -a claude-code`. Or copy the skill folder (skills/ads/capabilities/render-cinematic-music-video in gooseworks-ai/goose-skills) into .claude/skills/render-cinematic-music-video in your project. Claude Code loads it when a task matches its description.

How do I install Render Cinematic Music Video in Codex?

Run `npx skills add gooseworks-ai/goose-skills --skill render-cinematic-music-video -a codex`. Or copy the skill folder (skills/ads/capabilities/render-cinematic-music-video in gooseworks-ai/goose-skills) into .agents/skills/render-cinematic-music-video in your project. Codex loads it when a task matches its description.

Can I use Render Cinematic Music Video in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add gooseworks-ai/goose-skills --skill render-cinematic-music-video -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/render-cinematic-music-video, .gemini/skills/render-cinematic-music-video, .github/skills/render-cinematic-music-video and .opencode/skills/render-cinematic-music-video in your project.

What does Render Cinematic Music Video need to run?

SKILL.md names no scripts, command-line tools or credentials: Render Cinematic Music Video is instructions for the agent only.

Does Render Cinematic Music Video access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Render Cinematic Music Video safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Render Cinematic Music Video use?

Render Cinematic Music Video is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Render Cinematic Music Video use?

About 1.3k tokens (SKILL.md is roughly 5.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Render Cinematic Music Video?

Skills that share tags, products or a category with Render Cinematic Music Video: Local AI Use (amd/skills, 408 stars), Video Production (speechlab0210/video-production-skill, 105 stars), Voice AI Development (davila7/claude-code-templates, 33k stars) and Local AI App Integration (amd/skills, 408 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Render Cinematic Music Video?

gooseworks-ai (a GitHub organization) maintains it in gooseworks-ai/goose-skills, which has 1,240 GitHub stars. The repository holds 273 skills in this directory. The repository was last updated on October 8, 2026.

Source: gooseworks-ai/goose-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.