Topic · Media & Creative
Best music and audio generation skills, page 2
Music and audio generation skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 49 | Generates text-to-speech narration and custom sound effects for a video timeline, keeping existing voiceover in sync after visual retiming edits. | 0xsline/ | 2.2k | — | ~4.4k | Automated safety check: Pass | AGPL-3.0 | 2 days ago |
| 50 | A skill your agent uses whenever the user wants to generate a song or music with vocals from a text description and/or lyrics, or cover an existing song in a new style. | NoizAI/ | 526 | — | ~2.2k | Automated safety check: Pass | No licence | 11 days ago |
| 51 | 51.Qianwen Text Generate text, have conversations, write code, reason, and call functions with Qwen models. | QianWen-AI/ | 104 | — | ~4.7k | Automated safety check: Notes | Apache-2.0 | 18 days ago |
| 52 | AI-orchestrated video production on @pneuma-craft. An agent skill from pandazki/pneuma-skills. | pandazki/ | 161 | — | ~7.5k | Automated safety check: Notes | MIT | today |
| 53 | 53.Voiceover Narration, sound effects and music beds in the voice the studio set up (ElevenLabs, Qwen3-TTS on this machine's GPU, or the owner's own voice from the recording booth): lines from the approved… | GTKottman/ | 476 | — | ~1.8k | Automated safety check: Pass | AGPL-3.0 | yesterday |
| 54 | Internal skill that rewrites and re-renders sound prompts in a draft PeonPing pack to follow a reroll caption, then records the change in a log. | PeonPing/ | 5.1k | — | ~910 | Automated safety check: Pass | MIT | 3 days ago |
| 55 | 55.Music Generate music using ElevenLabs Music API. An agent skill from bozhouDev/video-skills-toolkit. | bozhouDev/ | 150 | — | ~3.6k | Automated safety check: Pass | MIT | 2 mo ago |
| 56 | Proposes beat-synced sound effects for a rendered short from a reusable library, then mixes an audition preview under the voice once you approve the plan. | hassancs91/ | 271 | — | ~2.6k | Automated safety check: Notes | MIT | 1 mo ago |
| 57 | Build a Vietnamese vertical TikTok explainer in the "MỔ XẺ PAPER AI" (AI paper dissection) format with HyperFrames (HTML/CSS/GSAP → MP4), Noti.vn style. | notivn/ | 126 | — | ~4.8k | Automated safety check: Pass | MIT | yesterday |
| 58 | Advanced motion designer with decades of After Effects and motion graphics experience, specialized in creating engaging video specifications for Remotion. | Hainrixz/ | 264 | — | ~2.3k | Automated safety check: Pass | Unknown | 6 mo ago |
| 59 | 59.Fal Assets Generate game assets with fal (fal.ai) through the fal MCP server, the um fal CLI (REST) or fal api. | rehan-remade/ | 5.8k | — | ~2k | Automated safety check: Notes | MIT | today |
| 60 | 60.Lyria Generate and validate music with Google Lyria 3 through the Gemini Interactions API. | calesthio/ | 66k | — | ~2k | Automated safety check: Pass | AGPL-3.0 | 6 days ago |
| 61 | Turns a Strudel live-coding music track into a 1080x1080 HyperFrames video where the code itself is the picture and each line highlights on the note it plays. | heygen-com/ | 183 | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | 10 days ago |
| 62 | 62.Suggest Sfx Step 4 of the AI Video Editor pipeline — the SFX pass. An agent skill from hassancs91/claude-youtube-editor. | hassancs91/ | 325 | — | ~3k | Automated safety check: Pass | MIT | 1 mo ago |
| 63 | Teaches positional (3D) audio in fluttersoloud — play3d/play3dClocked/play3dScheduled, listener position/orientation/velocity, per-source attenuation and Doppler, and the per-frame update pattern. | alnitak/ | 425 | — | ~2.2k | Automated safety check: Pass | MIT | 3 days ago |
| 64 | The AI music + sound-design skill for social -- original/licensed audio beds and sound design for Reels/TikToks/Shorts/videos. | social-media-skills/ | 128 | — | ~1.9k | Automated safety check: Pass | MIT | 8 days ago |
| 65 | Capability-based matrix for video production in which the agent picks a provider per capability, from storyboard and sound design to captions, render checks and packaging. | HKUDS/ | 52k | — | ~12k | Automated safety check: Pass | Apache-2.0 | 17 days ago |
| 66 | Agent-callable ElevenLabs tools — generate spoken audio from text, create sound effects and multi-speaker dialogue, re-voice and clean up audio, transcribe audio and video, design synthetic voices… | zapier/ | 176 | — | ~3.7k | Automated safety check: Pass | Elastic-2.0 | 1 mo ago |
| 67 | Direct ElevenLabs speech, dialogue, sound effects and music — the bracketed audio tags v3 acts on and why the voice decides whether a tag lands, stability as the delivery dial, punctuation instead… | nodetool-ai/ | 560 | — | ~1.9k | Automated safety check: Pass | AGPL-3.0 | today |
| 68 | 68.Assets Finding and downloading stock assets (footage, images, illustrations, 3D models, sound effects, fonts) from the asset sites the owner uses, in the owner's own Chrome with browser-harness. | GTKottman/ | 476 | — | ~601 | Automated safety check: Pass | AGPL-3.0 | yesterday |
| 69 | Songwriting craft and Suno AI music prompts. An agent skill from Prismer-AI/PrismerCloud. | Prismer-AI/ | 1.6k | 6 repos | ~2.9k | Automated safety check: Pass | MIT | 9 days ago |
| 70 | Provider-independent audio mixing and mastering direction for AI agents finishing generated videos, ads, trailers, explainers, podcasts, recuts, avatar clips, music videos, documentaries, and social… | calesthio/ | 193 | — | ~7k | Automated safety check: Pass | MIT | 2 mo ago |
| 71 | 71.Audio Track A skill your agent uses when the user wants music, voiceover, narration, or a soundtrack added to a video asset, OR wants standalone generated audio for any purpose (e.g. | ucsandman/ | 251 | — | ~1.1k | Automated safety check: Notes | MIT | 1 mo ago |
| 72 | Edit a Vietnamese vertical TikTok video (9:16) with HyperFrames following the Noti.vn/GĐT standard - talking-head + kinetic typography + karaoke captions + zoom/punch-in camera + timestamp-synced… | notivn/ | 126 | — | ~5.4k | Automated safety check: Pass | MIT | yesterday |
| 73 | Directs a lip-synced singing music video in the MiniMax H3 studio, locking the lead singer's identity, a scene master and exact audio references so the final track stays untouched. | karuvanan/ | 131 | — | ~1.8k | Automated safety check: Pass | Unknown | 23 days ago |
| 74 | Async music, sound-effect and long-form voice generation via Venice. | veniceai/ | 143 | — | ~3.1k | Automated safety check: Pass | MIT | 4 days ago |
| 75 | A skill your agent uses when playing or stopping background music, sound effects, or voice lines; creating/organizing audio channels; controlling volume, mute, or pan; or wiring Tone.js audio… | DRincs-Productions/ | 149 | — | ~3.3k | Automated safety check: Pass | LGPL-2.1 | 8 days ago |
| 76 | Build a Vietnamese landscape 16:9 YouTube video (1920×1080) with HyperFrames (HTML/CSS/GSAP → MP4), keeping the Noti.vn/GĐT branding inherited from noti-tiktok-vn. | notivn/ | 126 | — | ~5.6k | Automated safety check: Pass | MIT | yesterday |
| 77 | 77.Video Generate video clips on the user's own GPU through Guaardvark: text-to-video, image-to-video, first+last frame animation, clips with their own soundtrack and dialogue (MiniMax H3), short looping… | guaardvark/ | 257 | — | ~1.1k | Automated safety check: Pass | MIT | today |
| 78 | 78.Game Audio Game audio principles. An agent skill from xenitV1/Antigravity-Workflows. | xenitV1/ | 130 | 9 repos | ~1.3k | Automated safety check: Pass | MIT | 8 mo ago |
| 79 | Diagnoses HOT-Step CPP generation failures, engine crashes, hangs, and startup problems from the logs/ session folders. | scragnog/ | 173 | — | ~6k | Automated safety check: Notes | MIT | 2 days ago |
| 80 | A skill your agent uses when turning a song, track, or audio master into a finished music video with Scenario and Seedance: planning shots against beats and sections, transcribing lyrics, generating… | scenario-labs/ | 931 | — | ~1.8k | Automated safety check: Pass | MIT | yesterday |
| 81 | Add a new sound effect to @remotion/sfx | remotion-dev/ | 63k | — | ~816 | Automated safety check: Notes | Unknown | today |
| 82 | Create AI-powered podcasts with text-to-speech, music, and audio editing. | NeverSight/ | 216 | 1 repo | ~2k | Automated safety check: Pass | No licence | today |
| 83 | Command-line interface for Audacity - A stateful command-line interface for audio editing, following the same patterns as the GIMP and Ble... | HKUDS/ | 52k | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | 17 days ago |
| 84 | Audio storytelling skill used by the podcast scriptwriter and show note editor. | revfactory/ | 1.3k | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 85 | 新闻素材智能粗剪——把新闻素材粗剪为一条内容完整、逻辑清晰、节奏紧凑的新闻短视频。Use when the user asks to 粗剪新闻、新闻剪辑、把新闻素材剪成短视频、智能粗剪、news rough cut、编辑新闻视频, or provides news footage (发布会/采访/现场/监控素材) to be cut into a factual news short. | 0xsline/ | 2.2k | — | ~795 | Automated safety check: Pass | AGPL-3.0 | 2 days ago |
| 86 | Generate atmospheric sound-design / SFX (NOT speech, NOT the front-of-blog BGM) via OpenRouter using Google Lyria 3 Pro. | QinghongLin/ | 156 | — | ~618 | Automated safety check: Pass | MIT | 3 mo ago |
| 87 | Traces the full life of a music generation request (UI form → Node job queue → engine LM → synth → SQLite song row) and every place params can silently drop, especially the LM echo sideband gotcha. | scragnog/ | 173 | — | ~5.4k | Automated safety check: Pass | MIT | 2 days ago |
| 88 | Route the active iPolloWork video request to its current storyboard, compose, voiceover or soundtrack work, using the engine's native workflow and existing media tools. | Devin-AXIS/ | 6.8k | — | ~382 | Automated safety check: Pass | Unknown | yesterday |
| 89 | 89.Mmx CLI Generate text, images, video, speech, and music via the MiniMax AI platform. | agentscope-ai/ | 870 | — | ~655 | Automated safety check: Pass | Apache-2.0 | 28 days ago |
| 90 | Generate music from a text prompt using Sonilo — instrumental tracks, background beds, jingles, loops — when there is no video to score. | sonilo-ai/ | 115 | — | ~2.7k | Automated safety check: Notes | MIT | today |
| 91 | 91.Beatoven Beatoven.ai royalty-free music generation: compose tracks, poll tasks, download audio, fetch individual stems. | Anil-matcha/ | 1.3k | — | ~823 | Automated safety check: Pass | MIT | 4 days ago |
| 92 | Generate music (NOT speech) via OpenRouter using Google Lyria 3 Pro. | QinghongLin/ | 156 | — | ~407 | Automated safety check: Pass | MIT | 3 mo ago |
| 93 | 音频降噪:去除录音中的背景噪声、电流声、风噪、嗡嗡声,基于 ffmpeg 滤镜链(afftdn/highpass/lowpass)。 | ZJU-REAL/ | 3.3k | — | ~762 | Automated safety check: Pass | Apache-2.0 | today |
| 94 | 通用音频处理:音频剪辑/裁剪、格式转码(mp3/wav/m4a/aac)、音量归一化、从视频提取音轨、多段拼接、淡入淡出、变速(保音高)。当用户说“剪音频”“裁一段”“转成 mp3”“调音量/响度”“提取音轨/扒音频”“拼接音频”“淡入淡出”“加速/减速音频”“变速不变调”时使用。与 audio-denoise 的区别:audio-denoise 专做降噪,audio-editing… | ZJU-REAL/ | 3.3k | — | ~753 | Automated safety check: Pass | Apache-2.0 | today |
| 95 | 95.Audio Mix 音频混合 / 混音:把旁白口播 + 背景音乐 + 音效混成一轨,BGM 自动循环补足并可闪避(旁白说话时自动压低 BGM 保证人声清晰)。当用户说 混音、音频混合、旁白加背景音乐、配音加BGM、人声和音乐混一起、加音效、音频叠加、BGM 压低、闪避、ducking、把配音和bgm合起来 时使用。基于 shared/scripts/audiomix.py。与 audio-editing… | ZJU-REAL/ | 3.3k | — | ~483 | Automated safety check: Pass | Apache-2.0 | today |
| 96 | 批量处理:对一个目录里的一批图片/视频/音频统一套用同一操作——批量压缩、加水印、转格式、缩放、转比例、音量归一化等。当用户说 批量处理、批量压缩、批量加水印、批量转格式、一批图片/视频、给这个文件夹、全部转成、批量缩放、批量转竖版、整个目录 时使用。基于 shared/scripts/batchprocess.py(委派 imageops/videoops/audioops)。与… | ZJU-REAL/ | 3.3k | — | ~555 | Automated safety check: Pass | Apache-2.0 | today |
Explore related skills
Category
More topics in Media & Creative
- Video production892
- Image generation723
- AI video generation666
- Text to speech and voice632
- Transcription577
- Motion graphics384
- Design review and critique239
- Logo and visual identity229
- Comics and storyboards220
- Image editing193
- Social media graphics187
- Video scripts and shorts186
- Infographics158
- Podcasting124
- Generative and creative coding60