Search

Media & Creative · OpenAI · For developers

135 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Recommend suitable prompts from 10,000+ Nano Banana Pro image generation prompts based on user needs.

YouMind-OpenLab/nano-banana-pro-prompts-recommend-skill1.9k1 repo~4.1kAutomated safety check: PassNo licencetoday
2

Gives text-only models sight by running the modlens CLI on an image path or URL and returning structured JSON evidence with transcribed text, layout and semantics.

liustack/modlens4.2k—~1.3kAutomated safety check: NotesMIT7 days ago
3

BibiGPT CLI for summarizing videos, audio, and podcasts directly in the terminal.

JimmyLv/BibiGPT-v16.2k—~885Automated safety check: PassGPL-3.05 mo ago
4

Transcribe and summarize a video or podcast from a URL (YouTube, TikTok, Bilibili, Apple Podcasts, SoundCloud, 30+ platforms) or from a local media/.txt file.

wendy7756/AI-Video-Transcriber3.3k—~937Automated safety check: NotesApache-2.026 days ago
5

Transform real Chinese architecture and place-based cultural photos into art-directed editorial posters, integrated multi-photo scenes, and optional source comparisons.

op7418/guizang-yingzao-skill496—~1.1kAutomated safety check: PassNo licence1 mo ago
6

Plans and builds 2D game maps and scenes, from tilemaps and parallax backgrounds to HD-2D plates, with collision checks, a playable HTML preview and Tiled, Godot or LDtk export.

0x0funky/agent-sprite-forge4.4k—~2.9kAutomated safety check: PassMIT4 days ago
7

Generate, edit-from-reference, or analyze images with AI via OpenRouter (Gemini, GPT Image, Seedream, Qwen, MAI, Grok, FLUX.2, Recraft, Muse, Riverflow; Cloudflare AI Gateway BYOK).

centminmod/my-claude-code-setup2.7k—~8.1kAutomated safety check: NotesMIT3 days ago
8

给任何故事里的角色真出参考图(小说改编、自己原创的故事、单独设计一个角色都行,不需要小说原文): 一段话描述角色,拆成分层字段、补全后确认, 先出一张正面全身锚点,其余视图(大头照、90° 侧面、背面、细节、45° 大头照)都只参考这张锚点, 按需分档出图。每张图带标识、可单独重出,重出后自动标出哪些图过期。

eternityspring/shuohao-skills4.3k—~1.7kAutomated safety check: WarnApache-2.03 days ago
9

Dub a video into another language and generate subtitles using the default Together + Cartesia stack.

shang-zhu/violin1.1k—~1kAutomated safety check: NotesMIT1 mo ago
10

Transcribes audio files into text or subtitles through 9Router's Whisper-compatible endpoint, using models from OpenAI, Groq, Gemini, Deepgram and others.

decolua/9router31k—~914Automated safety check: PassMIT3 days ago
11

Hook an app, game or script up to the local mlx-serve server for LLM chat, embeddings, image, speech, music, sound effect, video and 3D generation, and Laya/Kev/Clef typed decisions.

ddalcu/mlx-serve1.8k—~1kAutomated safety check: PassUnknowntoday
12

Edit an existing JPG, JPEG, PNG, or WebP portrait to rebuild physically coherent light, exposure, color, capture style, clean optical skin, and frame quality without changing the person.

moskoo/xxg-portrait-rebuild-light216—~4.1kAutomated safety check: PassMIT2 days ago
13

Produces PR, changelog and two-week retro images for the Basic Memory repository from evidence in PR bodies, saved to fixed paths under docs/assets/infographics.

basicmachines-co/basic-memory4.1k—~2.7kAutomated safety check: PassAGPL-3.0today
14

Generate one or more standalone Meta image-ad creatives via ChatGPT Image 2 (gpt-image-2) through the Arcads external API.

krusemediallc/arcads-claude-code1.6k—~2.7kAutomated safety check: NotesMIT18 days ago
15

Turns text into spoken audio through a 9Router server, choosing among voices from OpenAI, ElevenLabs, Deepgram, Edge TTS and other providers.

decolua/9router31k—~765Automated safety check: PassMIT3 days ago
16

Develop supplied scripts, short stories, or synopses into scene understanding, director-facing art concepts, and motivated narrative keyframes; compile scene ideas, visual references, or existing…

popopo-99/zy-cinematic-realism570—~4.6kAutomated safety check: PassCC-BY-NC-4.07 days ago
17

Produces game-ready 2D characters, creatures, props, icons and effects as master stills, sheets or clips, and exports frames for common game engines.

0x0funky/agent-sprite-forge4.4k—~3.6kAutomated safety check: PassMIT4 days ago
18
18.TranscribeOfficial

Transcribe audio files to text with optional diarization and known-speaker hints.

JetBrains/skills3664 repos~776Automated safety check: PassApache-2.03 mo ago
19

Generate, edit, and iterate raster images from text prompts or reference images with GPT Image, Nano Banana, or Seedream.

Negai-ai/AgentClaw330—~1.9kAutomated safety check: NotesApache-2.04 mo ago
20

This skill should be used when the user asks to "generate an image", "create a logo", "draw an icon", "edit this photo", "change background to transparent", "remove background", "use GPT image"…

Wangnov/gpt-image-2-skill142—~4.5kAutomated safety check: PassMIT6 days ago
21

Generate, review, and integrate consistent 2D or stylized 3D character background scenes for Codex Persona Voice session cards.

miuuyy/persona-voice133—~953Automated safety check: PassMIT19 days ago
22

A skill your agent uses when the user wants to turn a reference image or text concept into a WeChat-ready animated Chinese meme GIF sticker pack: 16/24 entries, 240x240 GIFs, Chinese captions…

lisamsung/agent-meme-forge104—~973Automated safety check: PassMIT1 mo ago
23

Create, redesign, repair, validate, preview, and package lightweight standalone HTML deliverables.

NimaChu/html-design101—~2.7kAutomated safety check: PassNo licence1 mo ago
24

Generate images via OpenAI Images API (GPT Image, DALL-E 3, DALL-E 2).

swarmclawai/swarmclaw689—~705Automated safety check: PassMIT3 mo ago
25

End-to-end AI video production skill for agentic frameworks.

Bomx/super-video-maker-skill310—~11kAutomated safety check: NotesNo licence2 mo ago
26

Generate an image from a text prompt using an OpenAI-compatible image generation API (gpt-image-1-mini or compatible).

ai-sns/ai-sns332—~778Automated safety check: PassMIT2 mo ago
27

When the user wants to create, generate, edit, or optimize images for marketing — blog heroes, social graphics, product mockups, profile banners, listing visuals, or brand assets.

Nexus-JPF/note-companion8703 repos~3.9kAutomated safety check: PassMIT4 days ago
28

Generate, revise, translate, and manage App Store / Google Play marketing screenshots.

hypersocialinc/shots240—~2kAutomated safety check: PassNo licence5 mo ago
29

Produce beat-synced 1080p motion-graphic videos in HyperFrames (HTML + GSAP) with an AI voice-over, Vietnamese karaoke captions, SFX and generated music, in one of two proven styles (glass keynote…

bestagentkits/motion-video-skill118—~1.5kAutomated safety check: PassMIT12 days ago
30

Generate or edit raster images through a configurable OpenAI-compatible Image API using gpt-image-2.

fengfengzhidao/codex-image2-skill144—~1.2kAutomated safety check: PassMIT2 mo ago
31

Replaces Codex's built-in image tool with provider-based generation and editing, fetching reference images first whenever visual accuracy actually matters.

yc-duan/api-image101—~4.9kAutomated safety check: PassMIT5 mo ago
32

将朋友圈碎片图片、截图、物件、人物与零散想法,整理成具有幽默、黑色幽默、玩梗和编辑感的社交内容;自动选择 T1 人像主体、T2 杂物拼贴或 T3 全图压字视觉模板,并输出可直接生成的视觉简报与成图提示词。Use when the user wants to turn life fragments, photos, screenshots, objects, portraits, pets…

dacnay816y62-hub/fantasy-qiqiguaiguai-skill120—~1.7kAutomated safety check: PassMIT1 mo ago
33

Build voice AI agents with ElevenLabs. An agent skill from tadaspetra/loop.

tadaspetra/loop2961 repo~2.5kAutomated safety check: PassMIT8 days ago
34

Multi-backend Image + Video Generation CLI (nb command). An agent skill from jieyefriic/nbcraft.

jieyefriic/nbcraft155—~2.8kAutomated safety check: PassMIT5 mo ago
35

Generate or edit raster images by calling the ChatGPT/Codex hosted imagegeneration flow with local Codex or OpenClaw OAuth credentials, then save decoded image files for OpenClaw and other agent…

darkamenosa/codex-imagen135—~2.6kAutomated safety check: PassMIT17 days ago
36

Create, find, inspect, update, or delete Agent Storyboard projects directly through MCP.

Yuuhann1999/agent-storyboard360—~1.3kAutomated safety check: PassMITyesterday
37

Agent-native image workflow and optional prompt optimizer for /bananahub and generic agent image generation requests.

bananahub-ai/bananahub-skill118—~7.1kAutomated safety check: PassMIT4 mo ago
38

Generate one or more standalone Meta image-ad creatives via Nano Banana 2 / Nano Banana Pro (Gemini Flash Image family) through the Arcads external API.

krusemediallc/arcads-claude-code1.6k—~2.5kAutomated safety check: NotesMIT18 days ago
39

A skill your agent uses when working on tools/imggen avatar image generation, OpenAI-compatible image API config, human or yaoguai portrait prompts, qi-refining base generation, image-to-image realm…

4thfever/cultivation-world-simulator2.1k—~641Automated safety check: PassUnknown1 mo ago
40

OpenAI Audio Transcriptions API via curl; gpt-4o-transcribe, mini, diarize, or whisper-1.

openclaw/openclaw392k1 repo~518Automated safety check: PassMITtoday
41
41.SpeechOfficial

A skill your agent uses when the user asks for text-to-speech narration or voiceover, accessibility reads, audio prompts, or batch speech generation via the OpenAI Audio API; run the bundled CLI…

JetBrains/skills3663 repos~1.9kAutomated safety check: PassApache-2.03 mo ago
42

AI Agent自动剪辑旅行Vlog的完整工作流。从原始素材到成品视频,系统级只需ffmpeg,其余在Python venv内完成。by nyx研究所 (GitHub @znyupup · B站/小红书 @nyx研究所)

znyupup/ai-video-editing-skill150—~6.8kAutomated safety check: PassMIT5 mo ago
43

Two CLI tools for image generation + vision analysis using goclaw's provider-chain pattern.

therichardngai-code/gpt-image-2-pro-max101—~1.5kAutomated safety check: NotesMIT4 mo ago
44

Makes this agent generate images, transcribe audio, and synthesize speech on the user's own machine through a local Lemonade Server instead of a paid cloud API.

amd/skills408—~5kAutomated safety check: NotesMITyesterday
45

Process pending Agent Storyboard image and video generation tasks.

Yuuhann1999/agent-storyboard360—~2.8kAutomated safety check: PassMITyesterday
46

為使用者產生高品質的 AI 生圖、生影片、生音樂提示詞,並在需要時透過瀏覽器自動化實際送到目標平台。涵蓋 OiiOii、Kling 3.0/O-series、Seedance 2.0/2.5、Suno v5.5、Seedream 5.0/4.0、Vidu Q3、Midjourney V8.1、Flux 1.1 Pro / Kontext、Runway Gen-4.5 /…

Hao0321/ai-media-generator260—~5.6kAutomated safety check: PassMIT2 mo ago
47

Acquire images as files — generate them with an AI image model (14 providers: OpenAI/gpt-image, Gemini, Qwen, Zhipu, Volcengine, Stability, FLUX, Ideogram, MiniMax, and more), search openly-licensed…

open-octo/octo-agent125—~3.1kAutomated safety check: NotesMITyesterday
48

Builds a realtime voice-chat app around a talking character portrait made from your photo or a text description, with mouth sprites driven by the audio.

buildfastwithai/gen-ai-experiments785—~1.7kAutomated safety check: PassMIT19 days ago