Agent skill

MiniMax Music Generation

by bytedance in bytedance/deer-flow

Generates songs, jingles or instrumental tracks as MP3 files from a style prompt and optional lyrics through the MiniMax music API.

MITAuto-check passedMedia & Creative

Install MiniMax Music Generation

skills CLI
$ npx skills add bytedance/deer-flow --skill music-generation -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install bytedance/deer-flow music-generation --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/bytedance/deer-flow.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/public/music-generation .claude/skills/music-generation && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
music-generation
GitHub stars
84k
Token cost
~717 tokens
SKILL.md length
271 words
Files
2 (incl. scripts)
Skills in repo
23
Repo updated
First seen
Licence
MIT

At a glance

Generates songs, jingles or instrumental tracks as MP3 files from a style prompt and optional lyrics through the MiniMax music API.

  • Works in 3 steps: Understand Requirements → Create the Spec JSON → Execute Generation
  • Composing background music for a video or presentation
  • SKILL.md covers Overview, Workflow, Environment and Output Handling, plus 1 more section
  • Runs Python scripts from its folder; calls python; reaches api.minimaxi.com; needs MINIMAX_API_KEY

What it does

The agent identifies style, mood, scene, language and whether vocals are wanted, then writes a JSON spec with a required prompt and optional title, lyrics and an instrumental flag. Lyrics can carry structure tags such as intro, verse, chorus and bridge. If lyrics are given they are sung, an instrumental flag drops the vocals, and with neither the model writes lyrics from the prompt.

A bundled script, scripts/generate.py, takes the spec file and an output path and saves an MP3, typically in the outputs folder; the instructions say to call it with its parameters rather than read the source. It needs MINIMAX_API_KEY, with optional host and model variables; the default model is a free tier, and paid users can switch to a higher-limit one. The result is shared with you and the agent offers to adjust style or lyrics; non-English songs need lyrics in the target language.

When your agent uses it

  • Composing background music for a video or presentation
  • Producing a theme song or jingle from a mood description
  • Making an instrumental track with no vocals
  • Generating a song from lyrics you already wrote

Example prompts

  • “Make a mellow indie folk song about a rainy night in a cafe, with vocals.”
  • “Create an upbeat instrumental jingle for our podcast intro.”
  • “Turn these lyrics into a pop ballad and save the MP3.”

Requirements

  • A MiniMax API key set as MINIMAX_API_KEY
  • Python to run scripts/generate.py

Workflow steps

3 steps, taken from the step headings in SKILL.md.

  1. Understand Requirements
  2. Create the Spec JSON
  3. Execute Generation

What it can do on your machine

Read from SKILL.md and the folder at commit 8a3350a. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • api.minimaxi.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • MINIMAX_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

MiniMax Music Generation loads about 717 tokens when it runs. Until then it costs about 66 tokens; SKILL.md has 271 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~66
When it runs · the whole SKILL.md, loaded when a task matches
~717

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from bytedance/deer-flow at commit 8a3350a, republished under its MIT licence (© bytedance). 271 words, ~717 tokens.

Download SKILL.mdSave it as .claude/skills/music-generation/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
music-generation
description
Use this skill when the user requests to generate, create, compose, or produce music or songs — background music, theme songs, jingles, or instrumental tracks. Generates a song from a style/mood prompt and optional lyrics via the MiniMax music API.

Music Generation Skill

Overview

This skill generates songs (vocal or instrumental) from a structured JSON spec using the MiniMax music generation API (/v1/music_generation). You describe the style/mood/scene in prompt, optionally provide lyrics, and the script returns an MP3.

Workflow

Step 1: Understand Requirements

Identify the desired style, mood, scene, language, and whether the user wants vocals or a pure instrumental track. Decide whether to supply lyrics or let the model write them.

Step 2: Create the Spec JSON

Write a JSON file in /mnt/user-data/workspace/ named {descriptive-name}.json:

json
{
  "title": "Rainy Night Cafe",
  "prompt": "indie folk, melancholic, introspective, walking alone, cafe",
  "lyrics": "[verse]\nStreetlights glow the night wind sighs\n[chorus]\nPush the wooden door warm air inside"
}

Fields:

  • title (optional): a human-readable name.
  • prompt (required): style, mood, and scene. Drives the musical character.
  • lyrics (optional): song lyrics. Use \n between lines and structure tags such as [Intro], [Verse], [Pre Chorus], [Chorus], [Bridge], [Outro].
  • is_instrumental (optional, bool): set true for a pure instrumental track (no lyrics needed).

Behavior:

  • lyrics provided → those lyrics are sung.
  • is_instrumental: true → instrumental, no vocals.
  • neither → the model auto-writes lyrics from prompt (lyrics_optimizer).
Step 3: Execute Generation
bash
python /mnt/skills/public/music-generation/scripts/generate.py \
  --prompt-file /mnt/user-data/workspace/rainy-night-cafe.json \
  --output-file /mnt/user-data/outputs/rainy-night-cafe.mp3

Parameters:

  • --prompt-file: Absolute path to the JSON spec (required).
  • --output-file: Absolute path for the output MP3 (required).

[!NOTE] Do NOT read the python file, just call it with the parameters.

Environment

  • MINIMAX_API_KEY (required): your MiniMax interface key.
  • MINIMAX_API_HOST (optional): default https://api.minimaxi.com.
  • MINIMAX_MUSIC_MODEL (optional): default music-2.6-free (works for all API-key users); paid/Token-Plan users can set music-2.6 for higher limits.

Output Handling

  • Music is saved as MP3 (typically in /mnt/user-data/outputs/).
  • Share the generated file with the user using the present_files tool.
  • Offer to iterate on style or lyrics if adjustments are needed.

Notes

  • Keep prompt focused on style/mood/scene; put the actual sung words in lyrics.
  • For non-English songs, write lyrics in the target language.

© bytedance, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (scripts) in skills/public/music-generation of bytedance/deer-flow.

  • SKILL.md
  • scripts/generate.py

Open the folder on GitHubat commit 8a3350a

Compare with similar skills

MiniMax Music Generation next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

MiniMax Music Generation compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
MiniMax Music Generation this skillbytedance/deer-flow84k—~717Automated safety check: PassMIT
VRGDG H3 Short Film Pipelinevrgamegirl19/comfyui-vrgamedevgirl765—~4.2kAutomated safety check: PassCustom licence
Media Productionleon-ai/leon18k—~999Automated safety check: PassMIT
Music Caption RewriterT8mars/T8-penguin-canvas615—~2.2kAutomated safety check: PassMIT
Music Generation0xsline/OpenChatCut2.2k—~1.1kAutomated safety check: PassAGPL-3.0
Venice Audio Musicveniceai/skills144—~3.1kAutomated safety check: PassMIT

Similar skills

  • VRGDG H3 Short Film Pipeline

    vrgamegirl19/comfyui-vrgamedevgirl

    Builds an AI short film in a local ComfyUI with the VRGDG Video Builder, MiniMax H3 scenes, reference images, a music score, QA and a final edit.

    765 GitHub stars~4.2k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Media Production

    leon-ai/leon

    Generates images, audio and video through Leon's media tools, joins them with FFmpeg, checks the output and attaches playable files for the owner.

    18k GitHub stars~999 tokensUpdated today
    Media & CreativeAuto-check passed
  • Music Caption Rewriter

    T8mars/T8-penguin-canvas

    Turn a brief music description and optional tagged lyrics into a professional MiniMax Music 3 structured caption with Global Metadata, Vocal Details, and a section-aware Arrangement.

    615 GitHub stars~2.2k tokensUpdated today
    Media & CreativeAuto-check passed
  • Music Generation

    0xsline/OpenChatCut

    Generates instrumentals, songs, soundtracks and covers through Mureka, MiniMax, Atlas Cloud or Sonilo using the `submit_music` tool.

    2.2k GitHub stars~1.1k tokensUpdated 4 days ago
    Media & CreativeAuto-check passed
  • Venice Audio Music

    veniceai/skills

    Async music, sound-effect and long-form voice generation via Venice.

    144 GitHub stars~3.1k tokensUpdated 5 days ago
    Media & CreativeAuto-check passed
  • Video

    guaardvark/guaardvark

    Generate video clips on the user's own GPU through Guaardvark: text-to-video, image-to-video, first+last frame animation, clips with their own soundtrack and dialogue (MiniMax H3), short looping…

    258 GitHub stars~1.2k tokensUpdated today
    Media & CreativeAuto-check passed

More from bytedance/deer-flow

All 23 skills in this repo
  • Vercel Deploy

    bytedance/deer-flow

    Deploys a project to Vercel with one script and no login, then returns a live preview URL and a claim link for moving the deployment into your own Vercel account.

    84k GitHub starsUsed in 10 repos~797 tokens
    Auto-check passed
  • GitHub Deep Research

    bytedance/deer-flow

    Researches a GitHub repository over four rounds using the GitHub API and web search, then writes a structured markdown report with timeline, metrics and Mermaid diagrams.

    84k GitHub starsUsed in 4 repos~1.3k tokens
    Auto-check passed
  • Chart Visualization

    bytedance/deer-flow

    Picks a suitable chart type from 26 options for your data, maps the data to that chart's parameters and generates a chart image through a JavaScript script.

    84k GitHub starsUsed in 1 repo~840 tokens
    Auto-check passed
  • Excel and CSV Data Analysis

    bytedance/deer-flow

    Analyzes uploaded Excel and CSV files with SQL through DuckDB, producing schema inspections, statistical summaries and exports to CSV, JSON or Markdown.

    84k GitHub starsUsed in 4 repos~2.2k tokens
    Auto-check passed
  • Structured Image Generation

    bytedance/deer-flow

    Turns an image request into a structured JSON prompt and runs a bundled Python script to generate the picture, optionally guided by reference images.

    84k GitHub starsUsed in 4 repos~2.9k tokens
    Auto-check passed
  • DeerFlow Smoke Test

    bytedance/deer-flow

    Walks through an end-to-end smoke test of a DeerFlow deployment: pull the latest code, deploy with Docker or locally, verify services, run health checks and write a report.

    84k GitHub stars~2.5k tokensUpdated today
    Auto-check: notes

Works with

Questions about MiniMax Music Generation

What does MiniMax Music Generation do?

Generates songs, jingles or instrumental tracks as MP3 files from a style prompt and optional lyrics through the MiniMax music API. The agent identifies style, mood, scene, language and whether vocals are wanted, then writes a JSON spec with a required prompt and optional title, lyrics and an instrumental flag. Lyrics can carry structure tags such as intro, verse, chorus and bridge.

When should I use MiniMax Music Generation?

MiniMax Music Generation fits situations like: composing background music for a video or presentation; producing a theme song or jingle from a mood description; making an instrumental track with no vocals; generating a song from lyrics you already wrote.

How do I install MiniMax Music Generation in Claude Code?

Run `npx skills add bytedance/deer-flow --skill music-generation -a claude-code`. Or copy the skill folder (skills/public/music-generation in bytedance/deer-flow) into .claude/skills/music-generation in your project. Claude Code loads it when a task matches its description.

How do I install MiniMax Music Generation in Codex?

Run `npx skills add bytedance/deer-flow --skill music-generation -a codex`. Or copy the skill folder (skills/public/music-generation in bytedance/deer-flow) into .agents/skills/music-generation in your project. Codex loads it when a task matches its description.

Can I use MiniMax Music Generation in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add bytedance/deer-flow --skill music-generation -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/music-generation, .gemini/skills/music-generation, .github/skills/music-generation and .opencode/skills/music-generation in your project.

What does MiniMax Music Generation need to run?

Going by SKILL.md and its folder, MiniMax Music Generation needs Python for the scripts in its folder, the command-line tools its instructions call (python) and credentials named MINIMAX_API_KEY. Our summary lists: A MiniMax API key set as MINIMAX_API_KEY; Python to run scripts/generate.py.

Does MiniMax Music Generation access the network?

SKILL.md names 1 domain. In commands or code: api.minimaxi.com; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.

Is MiniMax Music Generation safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does MiniMax Music Generation use?

MiniMax Music Generation is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does MiniMax Music Generation use?

About 717 tokens (SKILL.md is roughly 2.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to MiniMax Music Generation?

Skills that share tags, products or a category with MiniMax Music Generation: VRGDG H3 Short Film Pipeline (vrgamegirl19/comfyui-vrgamedevgirl, 765 stars), Media Production (leon-ai/leon, 18k stars), Music Caption Rewriter (T8mars/T8-penguin-canvas, 615 stars) and Music Generation (0xsline/OpenChatCut, 2.2k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains MiniMax Music Generation?

bytedance (a GitHub organization) maintains it in bytedance/deer-flow, which has 83,674 GitHub stars. The repository holds 23 skills in this directory. The repository was last updated on October 11, 2026.

Source: bytedance/deer-flow on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.