Agent skill

Render Cartoon Music Video

by gooseworks-ai in gooseworks-ai/goose-skills

Assemble a cartoon / animated / hand-crafted music-video ad from a config — a sung song carries the whole narrative while N per-bar i2v clips (one recurring animated character, one look pack) are…

MITAuto-check passedMedia & Creative

Install Render Cartoon Music Video

skills CLI
$ npx skills add gooseworks-ai/goose-skills --skill render-cartoon-music-video -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install gooseworks-ai/goose-skills render-cartoon-music-video --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/gooseworks-ai/goose-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/ads/capabilities/render-cartoon-music-video .claude/skills/render-cartoon-music-video && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
render-cartoon-music-video
GitHub stars
1.2k
Token cost
~1.5k tokens
SKILL.md length
762 words
Files
6 (incl. scripts)
Skills in repo
273
Repo updated
First seen
Licence
MIT

At a glance

Assemble a cartoon / animated / hand-crafted music-video ad from a config — a sung song carries the whole narrative while N per-bar i2v clips (one recurring animated character, one look pack) are…

  • The cartoon-music-video format
  • SKILL.md covers Choices, Run and Contract (the free assembly)
  • Tasks that involve Logo and visual identity
  • Tasks that involve Speech recognition and synthesis

What it does

Render Cartoon Music Video is an agent skill from gooseworks-ai/goose-skills. Assemble a cartoon / animated / hand-crafted music-video ad from a config — a sung song carries the whole narrative while N per-bar i2v clips (one recurring animated character, one look pack) are each cut to their BAR window from librosa beat-tracking and hard-concatenated on the bar, VEED-whisper white bold-sans captions in the BOTTOM third (Alignment 2, above the logo bug, no pill) burned from the song's word timings re-spelled against the locked lyrics, a persistent brand logo bug held over the body…

Its SKILL.md is about 1.5k tokens, which your agent loads only when the skill is triggered. The skill folder holds 7 other files, including scripts (for example `scripts/PIPELINE.md`, `scripts/README.md` and `scripts/config.example.json`).

It sits in Media & Creative, covering Logo and visual identity, Speech recognition and synthesis and Text to speech and voice. It works with ElevenLabs. The repository describes itself as: Library of Growth & GTM skills + data APIs for Claude Code, Codex, Cursor to run ads, social, content, lead gen, seo and data scraping. The licence is MIT.

When your agent uses it

  • The cartoon-music-video format
  • Tasks that involve Logo and visual identity
  • Tasks that involve Speech recognition and synthesis

Example prompts

  • “/render-cartoon-music-video”

What it can do on your machine

Read from SKILL.md and the folder at commit c650c6d. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 3 files in scripts/, which the agent can run.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Render Cartoon Music Video loads about 1.5k tokens when it runs. Until then it costs about 237 tokens; SKILL.md has 762 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~237
When it runs · the whole SKILL.md, loaded when a task matches
~1.5k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from gooseworks-ai/goose-skills at commit c650c6d, republished under its MIT licence (© gooseworks-ai). 762 words, ~1,538 tokens.

Download SKILL.mdSave it as .claude/skills/render-cartoon-music-video/SKILL.md (or your agent's skills folder). This skill also uses 5 other files; get the full folder from GitHub.
name
render-cartoon-music-video
description
Assemble a cartoon / animated / hand-crafted music-video ad from a config — a sung song carries the whole narrative while N per-bar i2v clips (one recurring animated character, one look pack) are each cut to their BAR window from librosa beat-tracking and hard-concatenated on the bar, VEED-whisper white bold-sans captions in the BOTTOM third (Alignment 2, above the logo bug, no pill) burned from the song's word timings re-spelled against the locked lyrics, a persistent brand logo bug held over the body (suppressed on the end card), and closed on a solid-brand-color PIL end card with the song still playing under it — never AI-rendered text. This is the FREE deterministic assembly stage (cut-to-bar + hard concat + logo bug + captions + end card + song mux); the song, character, keyframes, and clips come from create-music-elevenlabs / create-image-fal / create-video-fal. Use for the cartoon-music-video format.
status
active

render-cartoon-music-video

Assemble a cartoon music-video ad from a config: an animated / illustrated / hand-crafted story (stop-motion felt-and-foam, claymation, 2D toon, cut-out) where a sung song is the entire script and one recurring animated character carries every shot, cut to the bar with the climax line on the drop. This capability is the FREE, deterministic assembly — cut-to-bar, hard-concat, the logo bug + captions burn, and the solid-color end card.

scripts/config.example.json is the worked example (Coinbase "Bet on anything", ~55s 1080×1920 9:16, 16 body bars + a 3s end card — its felt look, character, story and song are that demo's picks, never defaults); scripts/PIPELINE.md maps every config block to its source step and scripts/README.md documents the free assembly.

Choices

The creative calls are the caller's (the video-format recipe asks the user); this assembly never picks them. The demo's value is an example only:

  • Art style / look pack (felt stop-motion, claymation, 2D toon, cut-out…) — demo: felt-and-foam diorama.
  • Character — the ONE recurring animated figure — demo: a gender-neutral felt figure in a brand-blue jersey.
  • Story arc — demo: underdog payoff.
  • Song style + lead vocal — demo: hype dance-pop / EDM-trap, male half-rapped lead.

Run

This is the FREE, deterministic assembly stage — it spends nothing. The paid inputs are separate capabilities: the sung song (create-music-elevenlabs, or a user-supplied mp3 + Whisper word timings) beat-tracked with librosa so the BAR GRID sets the timeline; one locked recurring character + one keyframe per bar in one look pack (create-image-fal, Nano Banana); and one Seedance i2v clip per bar (create-video-fal). Given the song + word-timestamps.json + bars.json + one clip per bar + the brand wordmark SVG, render-cartoon-music-video cuts each clip to its bar window, hard-concats on the bar, burns the logo bug + captions, appends the solid-color end card, and muxes the song under it → the master. Re-cuts reuse the existing song / keyframes / clips and cost $0.

Contract (the free assembly)

  • The sung song carries the narrative — no separate VO. The generated/supplied track IS the bed (no VO to duck under); do not add a spoken voiceover or a second bed.
  • Plan the timeline AROUND the delivered song's BAR GRID. Beat-track the song with librosa (assume 4/4); one bar = one shot (default). Snap every tableau window to the bar boundaries — never trim the song to a pre-planned grid.
  • Captions from the song's word timings, re-spelled against the locked lyrics. Whisper mishears shouted accents ("BETS" → "Hearts"); re-spell the timed tokens against the locked lyric file (never edit lyrics to match Whisper). VEED-whisper white bold sans in the BOTTOM third (Alignment 2, margin_v above the bottom-left logo bug), ~4–5 words per cue, NO background pill; captions STOP at the end-card boundary. Never mid-frame — it covers the character (fixed 2026-07). If the host ffmpeg lacks libass, render the cues as timed PIL PNG overlays (ffmpeg overlay=…:enable='between(t,st,en)') at the same bottom placement.
  • Land the climax line on the drop. The climax tableau is timed so the payoff line sits on the sub-bass drop; accent that line.
  • Persistent logo bug, suppressed on the end card. Burn a white brand wordmark bottom-left over the body bars (cairosvg → PIL from the real SVG), suppressed on the end card where the big wordmark dominates.
  • End card via cairosvg + PIL from the real wordmark — never AI-render brand text. Solid brand-color card + white wordmark + subhead, holding ~3s WITH the song still playing under it (afade-out over the tail — no silent tail). A diffusion model garbles a wordmark.
  • FFmpeg composite, deterministic, FREE. Cut each clip to its bar window, hard-concat, burn the logo bug + captions, append the end card, mux the song over the whole video with a 0.5s afade tail, loudnorm I=-14 → a 1080×1920 h264+aac master. No paid calls, no keys. When the host ffmpeg lacks libass/drawtext, BOTH the captions and the logo bug are timed PIL PNG overlays (overlay=x:y:enable='between(t,st,en)'), not an ASS burn.
  • QC PER SCENE, never just the master — the two failure modes are content, not assembly. The upstream keyframe/i2v steps can (a) drift the character off its look (felt → smooth-3D "man" in the demo) partway through (off the chosen look — in the felt demo, felt → smooth-3D) and (b) hallucinate hands — realistic fingers in a hand macro, a pointing finger on a "tap the phone" shot, or a disembodied hand sliding in from the frame edge. Both hide at thumbnail size. Extract a 2 fps contact sheet + a per-bar FACE crop and HAND crop (across each clip's full duration), confirm every bar holds the ONE chosen look (the demo: matte felt) with the ONE character and no human/floating hands, and re-check the served bytes after publish. A drifted bar means regenerating that bar's KEYFRAME (not re-cutting) — see the recipe STEP 3/6.

© gooseworks-ai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 5 other files (scripts) in skills/ads/capabilities/render-cartoon-music-video of gooseworks-ai/goose-skills.

  • SKILL.md
  • scripts/PIPELINE.md
  • scripts/README.md
  • scripts/config.example.json
  • skill.meta.json
  • tests/smoke-test.md

Open the folder on GitHubat commit c650c6d

Compare with similar skills

Render Cartoon Music Video next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Render Cartoon Music Video compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Render Cartoon Music Video this skillgooseworks-ai/goose-skills1.2k—~1.5kAutomated safety check: PassMIT
Local AI Useamd/skills408—~5kAutomated safety check: NotesMIT
Voice AI Developmentdavila7/claude-code-templates33k5 repos~2.1kAutomated safety check: PassMIT
Elevenlabs Core Workflow Bjeremylongshore/tons-of-skills-marketplace2.8k—~1.6kAutomated safety check: PassMIT
Speech To Texttadaspetra/loop2962 repos~2kAutomated safety check: PassMIT
Video Productionspeechlab0210/video-production-skill105—~4.1kAutomated safety check: NotesMIT

Similar skills

  • Local AI Use

    amd/skills

    Makes this agent generate images, transcribe audio, and synthesize speech on the user's own machine through a local Lemonade Server instead of a paid cloud API.

    408 GitHub stars~5k tokensUpdated 2 days ago
    Media & CreativeAuto-check: notes
  • Voice AI Development

    davila7/claude-code-templates

    Expert in building voice AI applications - from real-time voice agents to voice-enabled apps.

    33k GitHub starsUsed in 5 repos~2.1k tokens
    Media & CreativeAuto-check passed
  • Elevenlabs Core Workflow B

    jeremylongshore/tons-of-skills-marketplace

    Implement ElevenLabs speech-to-speech, sound effects, audio isolation, and speech-to-text.

    2.8k GitHub stars~1.6k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Speech To Text

    tadaspetra/loop

    Transcribe audio to text using ElevenLabs Scribe v2. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 2 repos~2k tokens
    Media & CreativeAuto-check passed
  • Video Production

    speechlab0210/video-production-skill

    AI educational video production pipeline. An agent skill from speechlab0210/video-production-skill.

    105 GitHub stars~4.1k tokensUpdated 3 mo ago
    Media & CreativeAuto-check: notes
  • Agents

    tadaspetra/loop

    Build voice AI agents with ElevenLabs. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 1 repo~2.5k tokens
    AI & LLM EngineeringAuto-check passed

More from gooseworks-ai/goose-skills

All 273 skills in this repo
  • Reddit Post Finder

    gooseworks-ai/goose-skills

    Scrape and search Reddit posts using Apify. An agent skill from gooseworks-ai/goose-skills.

    1.2k GitHub starsUsed in 1 repo~1.2k tokens
    Auto-check passed
  • Create Image Fal

    gooseworks-ai/goose-skills

    Generate or edit an image via any FAL image model (nano-banana edit, gpt-image, flux, ...), ROUTED THROUGH THE fal-proxy so it bills the Ads agent.

    1.2k GitHub stars~1.3k tokensUpdated yesterday
    Auto-check passed
  • Render Hook Replacement

    gooseworks-ai/goose-skills

    Replace an existing video's opening with a supplied clip or free kinetic text hook while retaining and verifying every original body frame, audio, captions and ending.

    1.2k GitHub stars~2.3k tokensUpdated yesterday
    Auto-check passed
  • Blog Feed Monitor

    gooseworks-ai/goose-skills

    Scrape blog posts via RSS feeds (free, no API key) with Apify fallback for JS-heavy sites.

    1.2k GitHub starsUsed in 1 repo~578 tokens
    Auto-check passed
  • Competitor Post Engagers

    gooseworks-ai/goose-skills

    Find leads by scraping engagers from a competitor's top LinkedIn posts.

    1.2k GitHub starsUsed in 1 repo~1.8k tokens
    Auto-check: notes
  • Render Chatgpt Chat

    gooseworks-ai/goose-skills

    Assemble a ChatGPT chat-reveal video ad from a thread + timeline JSON — one continuous Playwright recording of a ChatGPT mobile chat (user types with the iOS keyboard up → taps send → keyboard…

    1.2k GitHub stars~2.3k tokensUpdated yesterday
    Auto-check passed

Works with

Questions about Render Cartoon Music Video

What does Render Cartoon Music Video do?

Assemble a cartoon / animated / hand-crafted music-video ad from a config — a sung song carries the whole narrative while N per-bar i2v clips (one recurring animated character, one look pack) are…. Render Cartoon Music Video is an agent skill from gooseworks-ai/goose-skills.

When should I use Render Cartoon Music Video?

Render Cartoon Music Video fits situations like: the cartoon-music-video format; tasks that involve Logo and visual identity; tasks that involve Speech recognition and synthesis.

How do I install Render Cartoon Music Video in Claude Code?

Run `npx skills add gooseworks-ai/goose-skills --skill render-cartoon-music-video -a claude-code`. Or copy the skill folder (skills/ads/capabilities/render-cartoon-music-video in gooseworks-ai/goose-skills) into .claude/skills/render-cartoon-music-video in your project. Claude Code loads it when a task matches its description.

How do I install Render Cartoon Music Video in Codex?

Run `npx skills add gooseworks-ai/goose-skills --skill render-cartoon-music-video -a codex`. Or copy the skill folder (skills/ads/capabilities/render-cartoon-music-video in gooseworks-ai/goose-skills) into .agents/skills/render-cartoon-music-video in your project. Codex loads it when a task matches its description.

Can I use Render Cartoon Music Video in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add gooseworks-ai/goose-skills --skill render-cartoon-music-video -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/render-cartoon-music-video, .gemini/skills/render-cartoon-music-video, .github/skills/render-cartoon-music-video and .opencode/skills/render-cartoon-music-video in your project.

What does Render Cartoon Music Video need to run?

SKILL.md names no scripts, command-line tools or credentials: Render Cartoon Music Video is instructions for the agent only.

Does Render Cartoon Music Video access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Render Cartoon Music Video safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Render Cartoon Music Video use?

Render Cartoon Music Video is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Render Cartoon Music Video use?

About 1.5k tokens (SKILL.md is roughly 6.2k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Render Cartoon Music Video?

Skills that share tags, products or a category with Render Cartoon Music Video: Local AI Use (amd/skills, 408 stars), Voice AI Development (davila7/claude-code-templates, 33k stars), Elevenlabs Core Workflow B (jeremylongshore/tons-of-skills-marketplace, 2.8k stars) and Speech To Text (tadaspetra/loop, 296 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Render Cartoon Music Video?

gooseworks-ai (a GitHub organization) maintains it in gooseworks-ai/goose-skills, which has 1,240 GitHub stars. The repository holds 273 skills in this directory. The repository was last updated on October 8, 2026.

Source: gooseworks-ai/goose-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.