Search
Media & Creative · OpenAI
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Generates or edits images from text prompts through a Python script that picks an image backend based on which API keys are configured. | zhayujie/ | 47k | — | ~1.3k | Automated safety check: Pass | MIT | yesterday |
| 2 | Recommend suitable prompts from 10,000+ Nano Banana Pro image generation prompts based on user needs. | YouMind-OpenLab/ | 1.9k | 1 repo | ~4.1k | Automated safety check: Pass | No licence | today |
| 3 | Adds subtitles to video with the videocaptioner CLI: transcribes speech, tidies and translates the text, and burns styled subtitles into the video or exports SRT and ASS files. | WEIFENG2333/ | 16k | — | ~1.7k | Automated safety check: Pass | GPL-3.0 | 29 days ago |
| 4 | Gives text-only models sight by running the modlens CLI on an image path or URL and returning structured JSON evidence with transcribed text, layout and semantics. | liustack/ | 4.2k | — | ~1.3k | Automated safety check: Notes | MIT | 7 days ago |
| 5 | Generates and edits images with GPT Image 2 or 2.5 through a packaged CLI and a prompt gallery, after settling which model fits the request. | wuyoscar/ | 5.7k | — | ~2.5k | Automated safety check: Notes | MIT | 11 days ago |
| 6 | 6.Bibi BibiGPT CLI for summarizing videos, audio, and podcasts directly in the terminal. | JimmyLv/ | 6.2k | — | ~885 | Automated safety check: Pass | GPL-3.0 | 5 mo ago |
| 7 | Transcribes audio with OpenAI's Whisper: 99 languages, translation to English, language detection, six model sizes and word-level timestamps, from Python or the CLI. | Orchestra-Research/ | 13k | 7 repos | ~1.9k | Automated safety check: Notes | MIT | 3 mo ago |
| 8 | Routes agents to the right KrillinAI command for subtitles, dubbing, video rendering, covers and speech, and explains how to read its JSON and manifest output. | krillinai/ | 13k | — | ~869 | Automated safety check: Pass | Apache-2.0 | today |
| 9 | 9.Imagegen Generate or edit raster images when the task benefits from AI-created bitmap visuals such as photos, illustrations, textures, sprites, mockups, or transparent-background cutouts. | theowenyoung/ | 115 | 4 repos | ~4.8k | Automated safety check: Pass | Apache-2.0 | 14 days ago |
| 10 | Batch-generate images via OpenAI Images API. An agent skill from trpc-group/trpc-agent-go. | trpc-group/ | 1.9k | 12 repos | ~843 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 11 | Generate or edit raster images (photos, illustrations, textures, sprites, mockups, logos, infographics) using the workspace's configured image-generation provider via onyx-cli image. | onyx-dot-app/ | 32k | 1 repo | ~1.7k | Automated safety check: Pass | Unknown | yesterday |
| 12 | Transcribe and summarize a video or podcast from a URL (YouTube, TikTok, Bilibili, Apple Podcasts, SoundCloud, 30+ platforms) or from a local media/.txt file. | wendy7756/ | 3.3k | — | ~937 | Automated safety check: Notes | Apache-2.0 | 26 days ago |
| 13 | Generates or edits images through ClawRouter's local image API, with a choice of models and sizes and payment handled automatically through x402. | BlockRunAI/ | 6.6k | — | ~2.1k | Automated safety check: Pass | MIT | 6 days ago |
| 14 | 14.Yingzao Transform real Chinese architecture and place-based cultural photos into art-directed editorial posters, integrated multi-photo scenes, and optional source comparisons. | op7418/ | 496 | — | ~1.1k | Automated safety check: Pass | No licence | 1 mo ago |
| 15 | Generates images through a 9Router gateway's image endpoint, with model discovery, the request fields and per-provider quirks for OpenAI, Gemini, MiniMax and others. | decolua/ | 31k | — | ~830 | Automated safety check: Pass | MIT | 3 days ago |
| 16 | Analyzes a reference image and writes a prompt that could recreate it in an AI image generator, focusing on the visual traits that most affect similarity. | wuyoscar/ | 5.7k | — | ~1.8k | Automated safety check: Pass | MIT | 11 days ago |
| 17 | Plans and builds 2D game maps and scenes, from tilemaps and parallax backgrounds to HD-2D plates, with collision checks, a playable HTML preview and Tiled, Godot or LDtk export. | 0x0funky/ | 4.4k | — | ~2.9k | Automated safety check: Pass | MIT | 5 days ago |
| 18 | 18.Ecom Image2 由 buluslan(公众号:新西楼.AI)研发的开源电商做图 Skill:39 个电商场景模板、Campaign 套图一致性、GPT-Image-2.5 官方双模型路由(Flare/Sunburst)与平台技术预检。通过用户配置的 OpenAI 兼容端点生成图片,或导出 prompt 包手动使用。Trigger whenever the user wants product main… | buluslan/ | 410 | — | ~2.9k | Automated safety check: Pass | MIT | 24 days ago |
| 19 | Generate, edit-from-reference, or analyze images with AI via OpenRouter (Gemini, GPT Image, Seedream, Qwen, MAI, Grok, FLUX.2, Recraft, Muse, Riverflow; Cloudflare AI Gateway BYOK). | centminmod/ | 2.7k | — | ~8.1k | Automated safety check: Notes | MIT | 3 days ago |
| 20 | 给任何故事里的角色真出参考图(小说改编、自己原创的故事、单独设计一个角色都行,不需要小说原文): 一段话描述角色,拆成分层字段、补全后确认, 先出一张正面全身锚点,其余视图(大头照、90° 侧面、背面、细节、45° 大头照)都只参考这张锚点, 按需分档出图。每张图带标识、可单独重出,重出后自动标出哪些图过期。 | eternityspring/ | 4.3k | — | ~1.7k | Automated safety check: Warn | Apache-2.0 | 3 days ago |
| 21 | Dub a video into another language and generate subtitles using the default Together + Cartesia stack. | shang-zhu/ | 1.1k | — | ~1k | Automated safety check: Notes | MIT | 1 mo ago |
| 22 | Transcribes audio files into text or subtitles through 9Router's Whisper-compatible endpoint, using models from OpenAI, Groq, Gemini, Deepgram and others. | decolua/ | 31k | — | ~914 | Automated safety check: Pass | MIT | 3 days ago |
| 23 | Generates an image or an image-to-video clip through a configured provider API or a signed-in Codex or Grok CLI, and reports the route, file, hash and cost estimate. | 0x0funky/ | 4.4k | — | ~2.2k | Automated safety check: Pass | MIT | 5 days ago |
| 24 | AI image generation with OpenAI, Google, DashScope and Replicate APIs. | XY121718/ | 103 | 1 repo | ~2.3k | Automated safety check: Notes | No licence | 7 mo ago |
| 25 | 25.Mlx Serve Hook an app, game or script up to the local mlx-serve server for LLM chat, embeddings, image, speech, music, sound effect, video and 3D generation, and Laya/Kev/Clef typed decisions. | ddalcu/ | 1.8k | — | ~1k | Automated safety check: Pass | Unknown | today |
| 26 | A skill your agent uses when the user needs engineering or research-paper figures: system architecture diagrams, algorithm workflows, hardware schematics, benchmark charts, ablation plots, figure… | heyu-233/ | 307 | — | ~1.1k | Automated safety check: Pass | MIT | 4 mo ago |
| 27 | Edit an existing JPG, JPEG, PNG, or WebP portrait to rebuild physically coherent light, exposure, color, capture style, clean optical skin, and frame quality without changing the person. | moskoo/ | 216 | — | ~4.1k | Automated safety check: Pass | MIT | 2 days ago |
| 28 | Produces PR, changelog and two-week retro images for the Basic Memory repository from evidence in PR bodies, saved to fixed paths under docs/assets/infographics. | basicmachines-co/ | 4.1k | — | ~2.7k | Automated safety check: Pass | AGPL-3.0 | today |
| 29 | Generate one or more standalone Meta image-ad creatives via ChatGPT Image 2 (gpt-image-2) through the Arcads external API. | krusemediallc/ | 1.6k | — | ~2.7k | Automated safety check: Notes | MIT | 18 days ago |
| 30 | Turns text into spoken audio through a 9Router server, choosing among voices from OpenAI, ElevenLabs, Deepgram, Edge TTS and other providers. | decolua/ | 31k | — | ~765 | Automated safety check: Pass | MIT | 3 days ago |
| 31 | Develop supplied scripts, short stories, or synopses into scene understanding, director-facing art concepts, and motivated narrative keyframes; compile scene ideas, visual references, or existing… | popopo-99/ | 570 | — | ~4.6k | Automated safety check: Pass | CC-BY-NC-4.0 | 7 days ago |
| 32 | 32.Gpt Image Generate or edit raster images through the user's ChatGPT subscription and save or preview workspace PNGs. | GENEXIS-AI/ | 175 | — | ~3.6k | Automated safety check: Pass | No licence | 1 mo ago |
| 33 | Produces game-ready 2D characters, creatures, props, icons and effects as master stills, sheets or clips, and exports frames for common game engines. | 0x0funky/ | 4.4k | — | ~3.6k | Automated safety check: Pass | MIT | 5 days ago |
| 34 | Creates and edits Xiaohongshu (RedNote) cover images from a portrait photo and text, using GPT Image 2 first and a Gemini CLI script as fallback. | Vivixiao980/ | 204 | — | ~2.9k | Automated safety check: Pass | No licence | 4 mo ago |
| 35 | Transcribe audio files to text with optional diarization and known-speaker hints. | JetBrains/ | 366 | 4 repos | ~776 | Automated safety check: Pass | Apache-2.0 | 3 mo ago |
| 36 | Generate, edit, and iterate raster images from text prompts or reference images with GPT Image, Nano Banana, or Seedream. | Negai-ai/ | 330 | — | ~1.9k | Automated safety check: Notes | Apache-2.0 | 4 mo ago |
| 37 | This skill should be used when the user asks to "generate an image", "create a logo", "draw an icon", "edit this photo", "change background to transparent", "remove background", "use GPT image"… | Wangnov/ | 142 | — | ~4.5k | Automated safety check: Pass | MIT | 6 days ago |
| 38 | Generate, review, and integrate consistent 2D or stylized 3D character background scenes for Codex Persona Voice session cards. | miuuyy/ | 133 | — | ~953 | Automated safety check: Pass | MIT | 19 days ago |
| 39 | A skill your agent uses when the user wants to turn a reference image or text concept into a WeChat-ready animated Chinese meme GIF sticker pack: 16/24 entries, 240x240 GIFs, Chinese captions… | lisamsung/ | 104 | — | ~973 | Automated safety check: Pass | MIT | 1 mo ago |
| 40 | 40.HTML Design Create, redesign, repair, validate, preview, and package lightweight standalone HTML deliverables. | NimaChu/ | 101 | — | ~2.7k | Automated safety check: Pass | No licence | 1 mo ago |
| 41 | Generate images via OpenAI Images API (GPT Image, DALL-E 3, DALL-E 2). | swarmclawai/ | 689 | — | ~705 | Automated safety check: Pass | MIT | 3 mo ago |
| 42 | A skill your agent uses whenever the user asks to generate, create, render, draw, or make an image, picture, illustration, icon, logo, or any other visual asset. | NomaDamas/ | 186 | — | ~746 | Automated safety check: Pass | No licence | 22 days ago |
| 43 | End-to-end AI video production skill for agentic frameworks. | Bomx/ | 310 | — | ~11k | Automated safety check: Notes | No licence | 2 mo ago |
| 44 | 44.Draw Image Generate an image from a text prompt using an OpenAI-compatible image generation API (gpt-image-1-mini or compatible). | ai-sns/ | 332 | — | ~778 | Automated safety check: Pass | MIT | 2 mo ago |
| 45 | 45.Image When the user wants to create, generate, edit, or optimize images for marketing — blog heroes, social graphics, product mockups, profile banners, listing visuals, or brand assets. | Nexus-JPF/ | 870 | 3 repos | ~3.9k | Automated safety check: Pass | MIT | 4 days ago |
| 46 | Calls an OpenAI-compatible image generation API using PiDeck's separate image-provider config, then saves the result as a local file. | ayuayue/ | 1.1k | — | ~1.3k | Automated safety check: Pass | MIT | today |
| 47 | 47.Shots Generate, revise, translate, and manage App Store / Google Play marketing screenshots. | hypersocialinc/ | 240 | — | ~2k | Automated safety check: Pass | No licence | 5 mo ago |
| 48 | 48.Motion Video Produce beat-synced 1080p motion-graphic videos in HyperFrames (HTML + GSAP) with an AI voice-over, Vietnamese karaoke captions, SFX and generated music, in one of two proven styles (glass keynote… | bestagentkits/ | 118 | — | ~1.5k | Automated safety check: Pass | MIT | 12 days ago |