Search
Media & Creative · For developers
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 193 | Separate a person, product or hand from the background in a video, on the user's own computer with SAM 2. | Barty-Bart/ | 559 | — | ~2.4k | Automated safety check: Pass | Unknown | 3 days ago |
| 194 | 制作或精修不露脸商业与消费调查视频,覆盖选题、事实核查、10分钟以上叙事、真实动态素材、配音字幕、CTA,以及AI封面和各平台发布文案。按当前请求执行阶段,保留已确认稿件和品牌;不自动公开发布。 | trustfuture/ | 382 | — | ~1.3k | Automated safety check: Pass | MIT | 17 days ago |
| 195 | 195.Kkclaw 给你的 AI Agent 一个桌面身体 — Setup Wizard、14情绪球体、语音克隆、歌词窗、Doctor 自检、跨平台支持(Windows + macOS) | kk43994/ | 174 | — | ~576 | Automated safety check: Pass | MIT | 6 mo ago |
| 196 | 196.Gemini Image Reference guide for using google-genai Python library to generate images with gemini-3-pro-image-preview model. | tyrchen/ | 236 | — | ~1.1k | Automated safety check: Pass | No licence | 6 mo ago |
| 197 | MiniMax multimodal model skill — use MiniMax Multi-Modal models for speech, music, video, and image. | poco-ai/ | 1.4k | — | ~7.6k | Automated safety check: Pass | MIT | 19 days ago |
| 198 | Run an SEO page observation-action-review loop with persistent Memory by coordinating demand research, page creation, adversarial review, image generation, IndexNow submission, and performance review. | tsingyuai/ | 2k | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | 13 days ago |
| 199 | 199.Noon Authoring Author and review Noon animation scenes using capability-qualified ManimCE-style Python or the shared Rust API. | yongkyuns/ | 133 | — | ~2k | Automated safety check: Pass | No licence | today |
| 200 | 200.Srt Vox Director 把已经写好的字幕(SRT / 配音稿 / 解说词 / 旁白稿)变成 Vox 风格解释视频的分镜与提示词包——参考图提示词、图生视频提示词、分镜表、关键词台账、视觉圣经、风格选择。只交付文本提示词,不生成任何图片或视频。适用于用户已有成片旁白、想做成分镜或配画面、需要切分镜头与处理时长差值、选定视觉风格,或出图/出片失败后诊断修复(字错了、画面没动、风格跑偏等)。 | geeklee/ | 105 | — | ~1.8k | Automated safety check: Pass | No licence | 2 mo ago |
| 201 | 201.Dreamina CLI A skill your agent uses when an agent needs Dreamina(即梦) generation, task querying, account checks, or login/session operations through the packaged Python wrapper scripts around the dreamina CLI. | yuyou-dev/ | 129 | — | ~1.1k | Automated safety check: Pass | No licence | 2 mo ago |
| 202 | 202.Bench mlx-serve benchmarking methodology — bench.sh/llmprobe usage, comparison-trap rules (same-methodology cells only, spec-decode variance, thermal lies, engine naming), perf-claim etiquette. | ddalcu/ | 1.8k | 1 repo | ~1.1k | Automated safety check: Pass | Unknown | today |
| 203 | Replaces Codex's built-in image tool with provider-based generation and editing, fetching reference images first whenever visual accuracy actually matters. | yc-duan/ | 101 | — | ~4.9k | Automated safety check: Pass | MIT | 5 mo ago |
| 204 | 204.Video Analyzer Analyzes a video file with Google Gemini and returns a structured markdown report covering top-level summary, scene-by-scene breakdown, audio transcript (or honest "silent" note), visual details… | mikefutia/ | 101 | — | ~747 | Automated safety check: Notes | No licence | 5 mo ago |
| 205 | 幫 Leo 看影片。當 Leo 丟影片連結(YouTube/IG/TikTok)或本機影片檔,要摘要、分析、拆解對標時使用——Claude 不能直接吃影片,先用這個工具抽關鍵幀+逐字稿+運鏡節奏+聲音/語氣/手勢時間軸再讀。 | HUANGCHIHHUNGLeo/ | 2.2k | — | ~559 | Automated safety check: Pass | MIT | 3 days ago |
| 206 | A skill your agent uses when the user wants to create an explainer, documentary, knowledge-sharing, news-broadcast, product-introduction, or data-report video from a topic. | calcuforge/ | 100 | — | ~6.5k | Automated safety check: Pass | Apache-2.0 | 25 days ago |
| 207 | Creates, refines and exports editable OneWorks 3D geometric avatars for mascots, bots or agents, keeping one shared scene state across preview, share link and export. | oneworks-ai/ | 251 | — | ~1.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 208 | Builds podcast-style audio narration from text with Azure OpenAI's GPT Realtime Mini over WebSocket, from a Python FastAPI backend to a React player. | microsoft/ | 3.1k | 1 repo | ~947 | Automated safety check: Pass | MIT | 2 days ago |
| 209 | 将朋友圈碎片图片、截图、物件、人物与零散想法,整理成具有幽默、黑色幽默、玩梗和编辑感的社交内容;自动选择 T1 人像主体、T2 杂物拼贴或 T3 全图压字视觉模板,并输出可直接生成的视觉简报与成图提示词。Use when the user wants to turn life fragments, photos, screenshots, objects, portraits, pets… | dacnay816y62-hub/ | 120 | — | ~1.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 210 | 210.Higgsfield A skill your agent uses whenever the user asks anything about Higgsfield AI — writing or refining video/image prompts, choosing a model (Kling, Veo, Wan, Seedance, Minimax Hailuo, DoP, Soul, Nano… | OSideMedia/ | 713 | — | ~9.1k | Automated safety check: Pass | MIT | 14 days ago |
| 211 | 211.Video Assemble 合成视频解说最终成片:把旁白音频铺到源视频上,按旁白窗口压低原声,生成 SRT / ASS 字幕并可烧录, 最后做响度标准化。作为最终合成阶段使用。输入源视频、ttsmeta.json 与旁白位置; 输出 recap 成片和字幕。触发词:视频合成、混音、字幕、压字幕、assemble video、mux、ducking、subtitles、成片。 | zenstory-ai/ | 561 | — | ~1.7k | Automated safety check: Pass | MIT | 7 days ago |
| 212 | Record browser or Web UI interaction demos as optimized GIFs using the available browser-control workflow, optional Playwright Videos for higher capture frame rates, and deterministic encoding, then… | singula-ai/ | 109 | 1 repo | ~4.1k | Automated safety check: Notes | MIT | 13 days ago |
| 213 | Generate, convert, clean, and integrate audio for Three.js browser games with ElevenLabs: sound effects, looping ambience, UI sounds, impact/weapon/vehicle audio, creature and boss stingers… | valkor-ai/ | 1.2k | 1 repo | ~1.3k | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 214 | 214.Content To Video Turn arbitrary source content (README, article, story, slides, deck, data/report, product description, tutorial text, audio/transcript, or a bare topic) into a finished, high-quality MP4 video. | architectds/ | 117 | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | today |
| 215 | 215.Kinetic Reel Make kinetic-typography motion reels (MP4) in code: showreels, portfolio or work reels, product promos, intro films, "motion design"-style videos with bold condensed type, HUD micro-type… | tuzhechen2005/ | 146 | — | ~1.8k | Automated safety check: Pass | Unknown | 14 days ago |
| 216 | 抖音视频 → 下载 → 转录 → 存为 Markdown 的完整工作流. An agent skill from chubbyguan/chubbyskills. | chubbyguan/ | 1.2k | — | ~551 | Automated safety check: Notes | MIT | 3 days ago |
| 217 | 217.PR Charts Turn llmprobe reports into the charts a perf PR embeds (one line per arm across context size, every point labelled, machine and model in the caption) and host them so the PR body renders them. | ddalcu/ | 1.8k | — | ~594 | Automated safety check: Pass | Unknown | today |
| 218 | Explains how to add and tune flutter_soloud's 13 audio filters at global, per-sound and mixing-bus level, including activation order and parameter fades. | alnitak/ | 425 | — | ~2.2k | Automated safety check: Pass | MIT | yesterday |
| 219 | Shows how to add Deepgram analytics such as diarization, summaries, sentiment, topics, redaction and language detection to speech transcription in Python. | deepgram/ | 469 | — | ~2.3k | Automated safety check: Pass | MIT | 2 days ago |
| 220 | 220.Vox Explainer End-to-end pipeline for producing Vox-style explainer videos from a single topic prompt. | CK42BB/ | 109 | — | ~2.6k | Automated safety check: Pass | MIT | 3 mo ago |
| 221 | 221.Image Generation Optimizes image generation prompts using Subject-Context-Style structure. | shinpr/ | 172 | — | ~1.5k | Automated safety check: Pass | MIT | 3 days ago |
| 222 | 222.Video Wrapper 为访谈视频添加综艺特效(花字、卡片、人物条、章节标题等)。支持 4 种视觉主题,先分析字幕内容生成建议供用户审批,再渲染视频。 | op7418/ | 339 | — | ~1.2k | Automated safety check: Notes | No licence | 8 mo ago |
| 223 | 223.Agents Build voice AI agents with ElevenLabs. An agent skill from tadaspetra/loop. | tadaspetra/ | 296 | 1 repo | ~2.5k | Automated safety check: Pass | MIT | 9 days ago |
| 224 | 224.Gauntlet Loop GAME skill. An agent skill from duolahypercho/gauntlet-loop. | duolahypercho/ | 165 | — | ~689 | Automated safety check: Pass | MIT | 2 mo ago |
| 225 | 225.Video When the user wants to create, generate, or produce video content using AI tools or programmatic frameworks. | Nexus-JPF/ | 870 | 3 repos | ~3.6k | Automated safety check: Pass | MIT | 4 days ago |
| 226 | Build, install, pair, and operate the Life Recorder iPhone-to-Mac local transcription system. | browser-use/ | 323 | — | ~599 | Automated safety check: Pass | MIT | 1 mo ago |
| 227 | 227.Test Convt CLI Test convt conversions end to end through the convt CLI, the engines in convt-engines, routing in convt-core, and the Linux file-manager menus in integrations/linux. | opencoredev/ | 286 | — | ~7.7k | Automated safety check: Pass | AGPL-3.0 | yesterday |
| 228 | Extracts frames and timestamped audio segments from video files (GIF, MP4, MOV) at configurable intervals and stores them in a directory with a manifest file. | qdhenry/ | 1.3k | — | ~1.6k | Automated safety check: Pass | No licence | 7 mo ago |
| 229 | 229.Animate Make a short procedural animation in any style — an explainer, a history, a little story — as a single-file canvas video with a synthesized score, rendered to MP4. | cth9191/ | 319 | — | ~4.5k | Automated safety check: Pass | MIT | 6 days ago |
| 230 | A skill your agent uses when building a local Windows WebCode installer from this repo for machine testing, especially when the package must bundle the Kokoro or sherpa-onnx Reply TTS service, model… | shuyu-labs/ | 278 | — | ~787 | Automated safety check: Pass | Unknown | 3 mo ago |
| 231 | 231.Video A skill your agent uses whenever the user asks to create, improve, audit, or split prompts for AI video generators (Seedance, Kling, Veo, Runway, Luma, Pika, Sora, any image-to-video system). | smixs/ | 494 | — | ~2.5k | Automated safety check: Pass | CC-BY-4.0 | 25 days ago |
| 232 | 232.Nbcraft Multi-backend Image + Video Generation CLI (nb command). An agent skill from jieyefriic/nbcraft. | jieyefriic/ | 155 | — | ~2.8k | Automated safety check: Pass | MIT | 5 mo ago |
| 233 | 233.Feedgrab Universal content grabber — fetch any URL and return structured Markdown. | iBigQiang/ | 614 | — | ~2k | Automated safety check: Pass | MIT | 1 mo ago |
| 234 | 234.Voxtype Install Guide the user through installing, configuring, and launching voxtype — local on-device voice dictation (speech-to-text that types wherever the cursor is). | pchalasani/ | 2k | — | ~857 | Automated safety check: Notes | MIT | 2 days ago |
| 235 | 235.Media over QUIC Build live video, audio, and real-time data apps with Media over QUIC (MoQ). Use when adding live streaming, conferencing, voice AI, or real-time… | moq-dev/ | 1.6k | — | ~460 | Automated safety check: Pass | Apache-2.0 | today |
| 236 | 236.Travel Skill 规划、编写、审阅和修复由真实授权素材与 AI 生成镜头混合制作的电影级文旅宣传片。用于文旅短片创意简报、场景图筛选、首帧空间分析、分镜脚本、画面内容提示词、人物动作、导演级运镜、自然动机布光、转场与尾帧衔接、Seedance 2.0/可灵/Midjourney 提示词、生成结果诊断、时间线与合规风险检查;尤其适用于需要写实、辽阔、自由感和角色剧情连续性的项目。 | kangarooking/ | 181 | — | ~756 | Automated safety check: Pass | MIT | 4 days ago |
| 237 | Design, evaluate, implement, or review image and video generation connectors for Timeline Studio. | MartinDelophy/ | 905 | — | ~989 | Automated safety check: Pass | MIT | yesterday |
| 238 | Generate PNG images by building a real Three.js 3D scene and capturing one frame headlessly — no image model involved. | hassancs91/ | 102 | — | ~1.4k | Automated safety check: Pass | MIT | 1 mo ago |
| 239 | 239.Podpull A skill your agent uses when the user wants to download a podcast episode's audio file (mp3/m4a) to disk — from an Apple Podcasts show or episode link, a raw RSS feed, or a 小宇宙/xiaoyuzhou episode… | xiaoleiy/ | 146 | — | ~962 | Automated safety check: Pass | MIT | 1 mo ago |
| 240 | 240.Video Prompting Draft and refine prompts for video generation models (including text-to-video, image/keyframe-to-video, and reference-driven generation), and create character-sheet prompts for image models when the… | Square-Zero-Labs/ | 182 | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |