Topic · Media & Creative

Best music and audio generation skills, page 2

Skills #49–96 of 150, ranked by score.

Music and audio generation skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Music and audio generation skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
49

Generates text-to-speech narration and custom sound effects for a video timeline, keeping existing voiceover in sync after visual retiming edits.

0xsline/OpenChatCut2.2k—~4.4kAutomated safety check: PassAGPL-3.02 days ago
50

A skill your agent uses whenever the user wants to generate a song or music with vocals from a text description and/or lyrics, or cover an existing song in a new style.

NoizAI/skills526—~2.2kAutomated safety check: PassNo licence11 days ago
51

Generate text, have conversations, write code, reason, and call functions with Qwen models.

QianWen-AI/qianwen-ai104—~4.7kAutomated safety check: NotesApache-2.018 days ago
52

AI-orchestrated video production on @pneuma-craft. An agent skill from pandazki/pneuma-skills.

pandazki/pneuma-skills161—~7.5kAutomated safety check: NotesMITtoday
53

Narration, sound effects and music beds in the voice the studio set up (ElevenLabs, Qwen3-TTS on this machine's GPU, or the owner's own voice from the recording booth): lines from the approved…

GTKottman/mortiflix-oss476—~1.8kAutomated safety check: PassAGPL-3.0yesterday
54

Internal skill that rewrites and re-renders sound prompts in a draft PeonPing pack to follow a reroll caption, then records the change in a log.

PeonPing/peon-ping5.1k—~910Automated safety check: PassMIT3 days ago
55

Generate music using ElevenLabs Music API. An agent skill from bozhouDev/video-skills-toolkit.

bozhouDev/video-skills-toolkit150—~3.6kAutomated safety check: PassMIT2 mo ago
56

Proposes beat-synced sound effects for a rendered short from a reusable library, then mixes an audition preview under the voice once you approve the plan.

hassancs91/claude-faceless-shorts-creator271—~2.6kAutomated safety check: NotesMIT1 mo ago
57

Build a Vietnamese vertical TikTok explainer in the "MỔ XẺ PAPER AI" (AI paper dissection) format with HyperFrames (HTML/CSS/GSAP → MP4), Noti.vn style.

notivn/AIEV126—~4.8kAutomated safety check: PassMITyesterday
58

Advanced motion designer with decades of After Effects and motion graphics experience, specialized in creating engaging video specifications for Remotion.

Hainrixz/editor-pro-max264—~2.3kAutomated safety check: PassUnknown6 mo ago
59

Generate game assets with fal (fal.ai) through the fal MCP server, the um fal CLI (REST) or fal api.

rehan-remade/universal-modder5.8k—~2kAutomated safety check: NotesMITtoday
60

Generate and validate music with Google Lyria 3 through the Gemini Interactions API.

calesthio/OpenMontage66k—~2kAutomated safety check: PassAGPL-3.06 days ago
61

Turns a Strudel live-coding music track into a 1080x1080 HyperFrames video where the code itself is the picture and each line highlights on the note it plays.

heygen-com/hyperframes-community-skills183—~1.3kAutomated safety check: PassApache-2.010 days ago
62

Step 4 of the AI Video Editor pipeline — the SFX pass. An agent skill from hassancs91/claude-youtube-editor.

hassancs91/claude-youtube-editor325—~3kAutomated safety check: PassMIT1 mo ago
63

Teaches positional (3D) audio in fluttersoloud — play3d/play3dClocked/play3dScheduled, listener position/orientation/velocity, per-source attenuation and Doppler, and the per-frame update pattern.

alnitak/flutter_soloud425—~2.2kAutomated safety check: PassMIT3 days ago
64

The AI music + sound-design skill for social -- original/licensed audio beds and sound design for Reels/TikToks/Shorts/videos.

social-media-skills/skills128—~1.9kAutomated safety check: PassMIT8 days ago
65

Capability-based matrix for video production in which the agent picks a provider per capability, from storyboard and sound design to captions, render checks and packaging.

HKUDS/CLI-Anything52k—~12kAutomated safety check: PassApache-2.017 days ago
66
66.ElevenlabsOfficial

Agent-callable ElevenLabs tools — generate spoken audio from text, create sound effects and multi-speaker dialogue, re-voice and clean up audio, transcribe audio and video, design synthetic voices…

zapier/connectors176—~3.7kAutomated safety check: PassElastic-2.01 mo ago
67

Direct ElevenLabs speech, dialogue, sound effects and music — the bracketed audio tags v3 acts on and why the voice decides whether a tag lands, stability as the delivery dial, punctuation instead…

nodetool-ai/nodetool560—~1.9kAutomated safety check: PassAGPL-3.0today
68

Finding and downloading stock assets (footage, images, illustrations, 3D models, sound effects, fonts) from the asset sites the owner uses, in the owner's own Chrome with browser-harness.

GTKottman/mortiflix-oss476—~601Automated safety check: PassAGPL-3.0yesterday
69

Songwriting craft and Suno AI music prompts. An agent skill from Prismer-AI/PrismerCloud.

Prismer-AI/PrismerCloud1.6k6 repos~2.9kAutomated safety check: PassMIT9 days ago
70

Provider-independent audio mixing and mastering direction for AI agents finishing generated videos, ads, trailers, explainers, podcasts, recuts, avatar clips, music videos, documentaries, and social…

calesthio/generative-media-skills193—~7kAutomated safety check: PassMIT2 mo ago
71

A skill your agent uses when the user wants music, voiceover, narration, or a soundtrack added to a video asset, OR wants standalone generated audio for any purpose (e.g.

ucsandman/marketing-studio251—~1.1kAutomated safety check: NotesMIT1 mo ago
72

Edit a Vietnamese vertical TikTok video (9:16) with HyperFrames following the Noti.vn/GĐT standard - talking-head + kinetic typography + karaoke captions + zoom/punch-in camera + timestamp-synced…

notivn/AIEV126—~5.4kAutomated safety check: PassMITyesterday
73

Directs a lip-synced singing music video in the MiniMax H3 studio, locking the lead singer's identity, a scene master and exact audio references so the final track stays untouched.

karuvanan/MiniMax-H3-Director-Cut-Studio131—~1.8kAutomated safety check: PassUnknown23 days ago
74

Async music, sound-effect and long-form voice generation via Venice.

veniceai/skills143—~3.1kAutomated safety check: PassMIT4 days ago
75

A skill your agent uses when playing or stopping background music, sound effects, or voice lines; creating/organizing audio channels; controlling volume, mute, or pan; or wiring Tone.js audio…

DRincs-Productions/pixi-vn149—~3.3kAutomated safety check: PassLGPL-2.18 days ago
76

Build a Vietnamese landscape 16:9 YouTube video (1920×1080) with HyperFrames (HTML/CSS/GSAP → MP4), keeping the Noti.vn/GĐT branding inherited from noti-tiktok-vn.

notivn/AIEV126—~5.6kAutomated safety check: PassMITyesterday
77

Generate video clips on the user's own GPU through Guaardvark: text-to-video, image-to-video, first+last frame animation, clips with their own soundtrack and dialogue (MiniMax H3), short looping…

guaardvark/guaardvark257—~1.1kAutomated safety check: PassMITtoday
78

Game audio principles. An agent skill from xenitV1/Antigravity-Workflows.

xenitV1/Antigravity-Workflows1309 repos~1.3kAutomated safety check: PassMIT8 mo ago
79

Diagnoses HOT-Step CPP generation failures, engine crashes, hangs, and startup problems from the logs/ session folders.

scragnog/HOT-Step-CPP173—~6kAutomated safety check: NotesMIT2 days ago
80

A skill your agent uses when turning a song, track, or audio master into a finished music video with Scenario and Seedance: planning shots against beats and sections, transcribing lyrics, generating…

scenario-labs/skills931—~1.8kAutomated safety check: PassMITyesterday
81
81.Add SfxOfficial

Add a new sound effect to @remotion/sfx

remotion-dev/remotion63k—~816Automated safety check: NotesUnknowntoday
82

Create AI-powered podcasts with text-to-speech, music, and audio editing.

NeverSight/learn-skills.dev2161 repo~2kAutomated safety check: PassNo licencetoday
83

Command-line interface for Audacity - A stateful command-line interface for audio editing, following the same patterns as the GIMP and Ble...

HKUDS/CLI-Anything52k—~1.3kAutomated safety check: PassApache-2.017 days ago
84

Audio storytelling skill used by the podcast scriptwriter and show note editor.

revfactory/harness-1001.3k—~1.6kAutomated safety check: PassApache-2.06 mo ago
85

新闻素材智能粗剪——把新闻素材粗剪为一条内容完整、逻辑清晰、节奏紧凑的新闻短视频。Use when the user asks to 粗剪新闻、新闻剪辑、把新闻素材剪成短视频、智能粗剪、news rough cut、编辑新闻视频, or provides news footage (发布会/采访/现场/监控素材) to be cut into a factual news short.

0xsline/OpenChatCut2.2k—~795Automated safety check: PassAGPL-3.02 days ago
86

Generate atmospheric sound-design / SFX (NOT speech, NOT the front-of-blog BGM) via OpenRouter using Google Lyria 3 Pro.

QinghongLin/data2story-skill156—~618Automated safety check: PassMIT3 mo ago
87

Traces the full life of a music generation request (UI form → Node job queue → engine LM → synth → SQLite song row) and every place params can silently drop, especially the LM echo sideband gotcha.

scragnog/HOT-Step-CPP173—~5.4kAutomated safety check: PassMIT2 days ago
88

Route the active iPolloWork video request to its current storyboard, compose, voiceover or soundtrack work, using the engine's native workflow and existing media tools.

Devin-AXIS/iPolloWork6.8k—~382Automated safety check: PassUnknownyesterday
89

Generate text, images, video, speech, and music via the MiniMax AI platform.

agentscope-ai/OpenJudge870—~655Automated safety check: PassApache-2.028 days ago
90

Generate music from a text prompt using Sonilo — instrumental tracks, background beds, jingles, loops — when there is no video to score.

sonilo-ai/skills115—~2.7kAutomated safety check: NotesMITtoday
91

Beatoven.ai royalty-free music generation: compose tracks, poll tasks, download audio, fetch individual stems.

Anil-matcha/awesome-muse-connectors1.3k—~823Automated safety check: PassMIT4 days ago
92

Generate music (NOT speech) via OpenRouter using Google Lyria 3 Pro.

QinghongLin/data2story-skill156—~407Automated safety check: PassMIT3 mo ago
93

音频降噪:去除录音中的背景噪声、电流声、风噪、嗡嗡声,基于 ffmpeg 滤镜链(afftdn/highpass/lowpass)。

ZJU-REAL/Easel3.3k—~762Automated safety check: PassApache-2.0today
94

通用音频处理:音频剪辑/裁剪、格式转码(mp3/wav/m4a/aac)、音量归一化、从视频提取音轨、多段拼接、淡入淡出、变速(保音高)。当用户说“剪音频”“裁一段”“转成 mp3”“调音量/响度”“提取音轨/扒音频”“拼接音频”“淡入淡出”“加速/减速音频”“变速不变调”时使用。与 audio-denoise 的区别:audio-denoise 专做降噪,audio-editing…

ZJU-REAL/Easel3.3k—~753Automated safety check: PassApache-2.0today
95

音频混合 / 混音:把旁白口播 + 背景音乐 + 音效混成一轨,BGM 自动循环补足并可闪避(旁白说话时自动压低 BGM 保证人声清晰)。当用户说 混音、音频混合、旁白加背景音乐、配音加BGM、人声和音乐混一起、加音效、音频叠加、BGM 压低、闪避、ducking、把配音和bgm合起来 时使用。基于 shared/scripts/audiomix.py。与 audio-editing…

ZJU-REAL/Easel3.3k—~483Automated safety check: PassApache-2.0today
96

批量处理:对一个目录里的一批图片/视频/音频统一套用同一操作——批量压缩、加水印、转格式、缩放、转比例、音量归一化等。当用户说 批量处理、批量压缩、批量加水印、批量转格式、一批图片/视频、给这个文件夹、全部转成、批量缩放、批量转竖版、整个目录 时使用。基于 shared/scripts/batchprocess.py(委派 imageops/videoops/audioops)。与…

ZJU-REAL/Easel3.3k—~555Automated safety check: PassApache-2.0today