Search
Media & Creative · OpenAI · For developers
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Recommend suitable prompts from 10,000+ Nano Banana Pro image generation prompts based on user needs. | YouMind-OpenLab/ | 1.9k | 1 repo | ~4.1k | Automated safety check: Pass | No licence | today |
| 2 | Gives text-only models sight by running the modlens CLI on an image path or URL and returning structured JSON evidence with transcribed text, layout and semantics. | liustack/ | 4.2k | — | ~1.3k | Automated safety check: Notes | MIT | 7 days ago |
| 3 | 3.Bibi BibiGPT CLI for summarizing videos, audio, and podcasts directly in the terminal. | JimmyLv/ | 6.2k | — | ~885 | Automated safety check: Pass | GPL-3.0 | 5 mo ago |
| 4 | Transcribe and summarize a video or podcast from a URL (YouTube, TikTok, Bilibili, Apple Podcasts, SoundCloud, 30+ platforms) or from a local media/.txt file. | wendy7756/ | 3.3k | — | ~937 | Automated safety check: Notes | Apache-2.0 | 26 days ago |
| 5 | 5.Yingzao Transform real Chinese architecture and place-based cultural photos into art-directed editorial posters, integrated multi-photo scenes, and optional source comparisons. | op7418/ | 496 | — | ~1.1k | Automated safety check: Pass | No licence | 1 mo ago |
| 6 | Plans and builds 2D game maps and scenes, from tilemaps and parallax backgrounds to HD-2D plates, with collision checks, a playable HTML preview and Tiled, Godot or LDtk export. | 0x0funky/ | 4.4k | — | ~2.9k | Automated safety check: Pass | MIT | 4 days ago |
| 7 | Generate, edit-from-reference, or analyze images with AI via OpenRouter (Gemini, GPT Image, Seedream, Qwen, MAI, Grok, FLUX.2, Recraft, Muse, Riverflow; Cloudflare AI Gateway BYOK). | centminmod/ | 2.7k | — | ~8.1k | Automated safety check: Notes | MIT | 3 days ago |
| 8 | 给任何故事里的角色真出参考图(小说改编、自己原创的故事、单独设计一个角色都行,不需要小说原文): 一段话描述角色,拆成分层字段、补全后确认, 先出一张正面全身锚点,其余视图(大头照、90° 侧面、背面、细节、45° 大头照)都只参考这张锚点, 按需分档出图。每张图带标识、可单独重出,重出后自动标出哪些图过期。 | eternityspring/ | 4.3k | — | ~1.7k | Automated safety check: Warn | Apache-2.0 | 3 days ago |
| 9 | Dub a video into another language and generate subtitles using the default Together + Cartesia stack. | shang-zhu/ | 1.1k | — | ~1k | Automated safety check: Notes | MIT | 1 mo ago |
| 10 | Transcribes audio files into text or subtitles through 9Router's Whisper-compatible endpoint, using models from OpenAI, Groq, Gemini, Deepgram and others. | decolua/ | 31k | — | ~914 | Automated safety check: Pass | MIT | 3 days ago |
| 11 | 11.Mlx Serve Hook an app, game or script up to the local mlx-serve server for LLM chat, embeddings, image, speech, music, sound effect, video and 3D generation, and Laya/Kev/Clef typed decisions. | ddalcu/ | 1.8k | — | ~1k | Automated safety check: Pass | Unknown | today |
| 12 | Edit an existing JPG, JPEG, PNG, or WebP portrait to rebuild physically coherent light, exposure, color, capture style, clean optical skin, and frame quality without changing the person. | moskoo/ | 216 | — | ~4.1k | Automated safety check: Pass | MIT | 2 days ago |
| 13 | Produces PR, changelog and two-week retro images for the Basic Memory repository from evidence in PR bodies, saved to fixed paths under docs/assets/infographics. | basicmachines-co/ | 4.1k | — | ~2.7k | Automated safety check: Pass | AGPL-3.0 | today |
| 14 | Generate one or more standalone Meta image-ad creatives via ChatGPT Image 2 (gpt-image-2) through the Arcads external API. | krusemediallc/ | 1.6k | — | ~2.7k | Automated safety check: Notes | MIT | 18 days ago |
| 15 | Turns text into spoken audio through a 9Router server, choosing among voices from OpenAI, ElevenLabs, Deepgram, Edge TTS and other providers. | decolua/ | 31k | — | ~765 | Automated safety check: Pass | MIT | 3 days ago |
| 16 | Develop supplied scripts, short stories, or synopses into scene understanding, director-facing art concepts, and motivated narrative keyframes; compile scene ideas, visual references, or existing… | popopo-99/ | 570 | — | ~4.6k | Automated safety check: Pass | CC-BY-NC-4.0 | 7 days ago |
| 17 | Produces game-ready 2D characters, creatures, props, icons and effects as master stills, sheets or clips, and exports frames for common game engines. | 0x0funky/ | 4.4k | — | ~3.6k | Automated safety check: Pass | MIT | 4 days ago |
| 18 | Transcribe audio files to text with optional diarization and known-speaker hints. | JetBrains/ | 366 | 4 repos | ~776 | Automated safety check: Pass | Apache-2.0 | 3 mo ago |
| 19 | Generate, edit, and iterate raster images from text prompts or reference images with GPT Image, Nano Banana, or Seedream. | Negai-ai/ | 330 | — | ~1.9k | Automated safety check: Notes | Apache-2.0 | 4 mo ago |
| 20 | This skill should be used when the user asks to "generate an image", "create a logo", "draw an icon", "edit this photo", "change background to transparent", "remove background", "use GPT image"… | Wangnov/ | 142 | — | ~4.5k | Automated safety check: Pass | MIT | 6 days ago |
| 21 | Generate, review, and integrate consistent 2D or stylized 3D character background scenes for Codex Persona Voice session cards. | miuuyy/ | 133 | — | ~953 | Automated safety check: Pass | MIT | 19 days ago |
| 22 | A skill your agent uses when the user wants to turn a reference image or text concept into a WeChat-ready animated Chinese meme GIF sticker pack: 16/24 entries, 240x240 GIFs, Chinese captions… | lisamsung/ | 104 | — | ~973 | Automated safety check: Pass | MIT | 1 mo ago |
| 23 | 23.HTML Design Create, redesign, repair, validate, preview, and package lightweight standalone HTML deliverables. | NimaChu/ | 101 | — | ~2.7k | Automated safety check: Pass | No licence | 1 mo ago |
| 24 | Generate images via OpenAI Images API (GPT Image, DALL-E 3, DALL-E 2). | swarmclawai/ | 689 | — | ~705 | Automated safety check: Pass | MIT | 3 mo ago |
| 25 | End-to-end AI video production skill for agentic frameworks. | Bomx/ | 310 | — | ~11k | Automated safety check: Notes | No licence | 2 mo ago |
| 26 | 26.Draw Image Generate an image from a text prompt using an OpenAI-compatible image generation API (gpt-image-1-mini or compatible). | ai-sns/ | 332 | — | ~778 | Automated safety check: Pass | MIT | 2 mo ago |
| 27 | 27.Image When the user wants to create, generate, edit, or optimize images for marketing — blog heroes, social graphics, product mockups, profile banners, listing visuals, or brand assets. | Nexus-JPF/ | 870 | 3 repos | ~3.9k | Automated safety check: Pass | MIT | 4 days ago |
| 28 | 28.Shots Generate, revise, translate, and manage App Store / Google Play marketing screenshots. | hypersocialinc/ | 240 | — | ~2k | Automated safety check: Pass | No licence | 5 mo ago |
| 29 | 29.Motion Video Produce beat-synced 1080p motion-graphic videos in HyperFrames (HTML + GSAP) with an AI voice-over, Vietnamese karaoke captions, SFX and generated music, in one of two proven styles (glass keynote… | bestagentkits/ | 118 | — | ~1.5k | Automated safety check: Pass | MIT | 12 days ago |
| 30 | 30.Codex Image2 Generate or edit raster images through a configurable OpenAI-compatible Image API using gpt-image-2. | fengfengzhidao/ | 144 | — | ~1.2k | Automated safety check: Pass | MIT | 2 mo ago |
| 31 | Replaces Codex's built-in image tool with provider-based generation and editing, fetching reference images first whenever visual accuracy actually matters. | yc-duan/ | 101 | — | ~4.9k | Automated safety check: Pass | MIT | 5 mo ago |
| 32 | 将朋友圈碎片图片、截图、物件、人物与零散想法,整理成具有幽默、黑色幽默、玩梗和编辑感的社交内容;自动选择 T1 人像主体、T2 杂物拼贴或 T3 全图压字视觉模板,并输出可直接生成的视觉简报与成图提示词。Use when the user wants to turn life fragments, photos, screenshots, objects, portraits, pets… | dacnay816y62-hub/ | 120 | — | ~1.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 33 | 33.Agents Build voice AI agents with ElevenLabs. An agent skill from tadaspetra/loop. | tadaspetra/ | 296 | 1 repo | ~2.5k | Automated safety check: Pass | MIT | 8 days ago |
| 34 | 34.Nbcraft Multi-backend Image + Video Generation CLI (nb command). An agent skill from jieyefriic/nbcraft. | jieyefriic/ | 155 | — | ~2.8k | Automated safety check: Pass | MIT | 5 mo ago |
| 35 | 35.Codex Imagen Generate or edit raster images by calling the ChatGPT/Codex hosted imagegeneration flow with local Codex or OpenClaw OAuth credentials, then save decoded image files for OpenClaw and other agent… | darkamenosa/ | 135 | — | ~2.6k | Automated safety check: Pass | MIT | 17 days ago |
| 36 | Create, find, inspect, update, or delete Agent Storyboard projects directly through MCP. | Yuuhann1999/ | 360 | — | ~1.3k | Automated safety check: Pass | MIT | yesterday |
| 37 | 37.Bananahub Agent-native image workflow and optional prompt optimizer for /bananahub and generic agent image generation requests. | bananahub-ai/ | 118 | — | ~7.1k | Automated safety check: Pass | MIT | 4 mo ago |
| 38 | Generate one or more standalone Meta image-ad creatives via Nano Banana 2 / Nano Banana Pro (Gemini Flash Image family) through the Arcads external API. | krusemediallc/ | 1.6k | — | ~2.5k | Automated safety check: Notes | MIT | 18 days ago |
| 39 | A skill your agent uses when working on tools/imggen avatar image generation, OpenAI-compatible image API config, human or yaoguai portrait prompts, qi-refining base generation, image-to-image realm… | 4thfever/ | 2.1k | — | ~641 | Automated safety check: Pass | Unknown | 1 mo ago |
| 40 | OpenAI Audio Transcriptions API via curl; gpt-4o-transcribe, mini, diarize, or whisper-1. | openclaw/ | 392k | 1 repo | ~518 | Automated safety check: Pass | MIT | today |
| 41 | A skill your agent uses when the user asks for text-to-speech narration or voiceover, accessibility reads, audio prompts, or batch speech generation via the OpenAI Audio API; run the bundled CLI… | JetBrains/ | 366 | 3 repos | ~1.9k | Automated safety check: Pass | Apache-2.0 | 3 mo ago |
| 42 | AI Agent自动剪辑旅行Vlog的完整工作流。从原始素材到成品视频,系统级只需ffmpeg,其余在Python venv内完成。by nyx研究所 (GitHub @znyupup · B站/小红书 @nyx研究所) | znyupup/ | 150 | — | ~6.8k | Automated safety check: Pass | MIT | 5 mo ago |
| 43 | 43.Media Tools Two CLI tools for image generation + vision analysis using goclaw's provider-chain pattern. | therichardngai-code/ | 101 | — | ~1.5k | Automated safety check: Notes | MIT | 4 mo ago |
| 44 | 44.Local AI Use Makes this agent generate images, transcribe audio, and synthesize speech on the user's own machine through a local Lemonade Server instead of a paid cloud API. | amd/ | 408 | — | ~5k | Automated safety check: Notes | MIT | yesterday |
| 45 | Process pending Agent Storyboard image and video generation tasks. | Yuuhann1999/ | 360 | — | ~2.8k | Automated safety check: Pass | MIT | yesterday |
| 46 | 為使用者產生高品質的 AI 生圖、生影片、生音樂提示詞,並在需要時透過瀏覽器自動化實際送到目標平台。涵蓋 OiiOii、Kling 3.0/O-series、Seedance 2.0/2.5、Suno v5.5、Seedream 5.0/4.0、Vidu Q3、Midjourney V8.1、Flux 1.1 Pro / Kontext、Runway Gen-4.5 /… | Hao0321/ | 260 | — | ~5.6k | Automated safety check: Pass | MIT | 2 mo ago |
| 47 | 47.Image Gen Acquire images as files — generate them with an AI image model (14 providers: OpenAI/gpt-image, Gemini, Qwen, Zhipu, Volcengine, Stability, FLUX, Ideogram, MiniMax, and more), search openly-licensed… | open-octo/ | 125 | — | ~3.1k | Automated safety check: Notes | MIT | yesterday |
| 48 | Builds a realtime voice-chat app around a talking character portrait made from your photo or a text description, with mouth sprites driven by the audio. | buildfastwithai/ | 785 | — | ~1.7k | Automated safety check: Pass | MIT | 19 days ago |