Search

Python · Image generation

42 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Turns an image request into a structured JSON prompt and runs a bundled Python script to generate the picture, optionally guided by reference images.

bytedance/deer-flow84k4 repos~2.9kAutomated safety check: PassMITtoday
2
2.Gemini API DevOfficial

A skill your agent uses when writing code that calls the Gemini API for text generation, multi-turn chat, multimodal understanding, image generation, video generation, speech generation (TTS), voice…

google-gemini/gemini-skills4.3k—~5.1kAutomated safety check: PassApache-2.02 days ago
3

Plans and builds 2D game maps and scenes, from tilemaps and parallax backgrounds to HD-2D plates, with collision checks, a playable HTML preview and Tiled, Godot or LDtk export.

0x0funky/agent-sprite-forge4.4k—~2.9kAutomated safety check: PassMIT2 days ago
4

A skill your agent uses when users ask for a hand-drawn or illustrated image prompt, name one of the repository's 20 numbered styles or the 3.1 stable variant, or mention triggers such as…

threerocks/hand-drawn-styles2.2k—~398Automated safety check: PassMITtoday
5

Generates an image or an image-to-video clip through a configured provider API or a signed-in Codex or Grok CLI, and reports the route, file, hash and cost estimate.

0x0funky/agent-sprite-forge4.4k—~2.2kAutomated safety check: PassMIT2 days ago
6

Produces game-ready 2D characters, creatures, props, icons and effects as master stills, sheets or clips, and exports frames for common game engines.

0x0funky/agent-sprite-forge4.4k—~3.6kAutomated safety check: PassMIT2 days ago
7

Generate smooth hand-drawn whiteboard and story videos directly inside Codex from text scripts, GPT Image 2 color storyboards, scene plans, SVGs, line art, or local images.

gnipbao/codex-whiteboard-video-skill326—~7.2kAutomated safety check: NotesMIT1 mo ago
8

Generate a game-ready 3D asset by running the local AI pipeline sequentially: Fooocus (SDXL text-to-image) - Hunyuan3D-2 (image-to-textured-GLB) - optional Blender FBX convert + Unreal import.

LaurentiuGabriel/unreal-game-assets-creation-skill148—~2.1kAutomated safety check: PassNo licence2 mo ago
9

A skill your agent uses when one or more photographs must become adaptive photo-plus-abstraction editorial compositions while source facts, spatial relationships, and strict photo preservation…

kwhi6693-web/photo-abstract-editorial113—~870Automated safety check: PassAGPL-3.01 mo ago
10

Reference guide for using google-genai Python library to generate images with gemini-3-pro-image-preview model.

tyrchen/geektime-bootcamp-ai236—~1.1kAutomated safety check: PassNo licence6 mo ago
11

A skill your agent uses when writing code that calls the Gemini API for text generation, multi-turn chat, multimodal understanding, image generation, video generation, streaming responses…

Ayuilos/Miffan208—~4.6kAutomated safety check: PassAGPL-3.0today
12

Generates and edits images with Stable Diffusion through Hugging Face Diffusers, covering text-to-image, image-to-image, inpainting, SDXL and custom pipelines.

Orchestra-Research/AI-Research-SKILLs13k5 repos~3.2kAutomated safety check: PassMIT3 mo ago
13

Designs and builds video-driven parallax hero landing pages with smooth video scrubbing, progress-locked typography, and clean navigation.

nroadley/Creating-Oneshot-Hero-Landing-Pages134—~3.5kAutomated safety check: PassMIT1 mo ago
14

A skill your agent uses when the user wants to drive Google Flow (Veo image-to-video, Veo text-to-video, Imagen / Nano Banana image generation) from the terminal or a script — including…

ffroliva/gflow-cli266—~4.8kAutomated safety check: NotesMITtoday
15

Low-level SenseNova tools for image generation, image editing, image recognition with a VLM and text optimization with an LLM, meant to be called by higher-level skills rather than directly.

OpenSenseNova/SenseNova-Skills5.7k—~3.2kAutomated safety check: PassMITtoday
16

Wire AG2 beta's shipped tools into an Agent — both provider-native server-side tools (web search, web fetch, code execution, MCP, image generation, memory) and locally-executed common toolkits…

ag2ai/build-with-ag2252—~1.3kAutomated safety check: PassApache-2.01 mo ago
17

Splits a Markdown article by heading, drafts an image prompt for each section, then generates diagrams with the Gemini API and embeds them in the note.

SpaceZephyr/design-buddy175—~1kAutomated safety check: PassNo licence3 mo ago
18

Create professional videos autonomously using claude-code-video-toolkit — AI voiceovers, image generation, music, talking heads, and Remotion rendering.

calesthio/OpenMontage66k1 repo~3.8kAutomated safety check: NotesAGPL-3.05 days ago
19

AI image generation and editing using Google Gemini models. An agent skill from kenneth-liao/ai-launchpad-marketplace.

kenneth-liao/ai-launchpad-marketplace126—~4kAutomated safety check: NotesNo licence6 mo ago
20

Generates diagrams, flowcharts and conceptual illustrations with OpenAI's gpt-image models, while leaving data plots and result figures to real plotting code.

RealSeaberry/AutoMCM-Pro257—~1.9kAutomated safety check: NotesMIT28 days ago
21

Connect this agent to a running Guaardvark (self-hosted AI studio) and check what it can do right now.

guaardvark/guaardvark257—~1.2kAutomated safety check: PassMITtoday
22

Generate, edit, and compose images using Gemini Nano Banana models via portable Python scripts.

cnemri/google-genai-skills127—~587Automated safety check: PassMIT8 mo ago
23

Generate or edit images when the user names OpenAI, GPT Image, or a gpt-image model.

feiskyer/codex-settings244—~572Automated safety check: PassMIT11 days ago
24

Generate images when the user names MiniMax or image-01/image-01-live.

feiskyer/codex-settings244—~495Automated safety check: PassMIT11 days ago
25

Generate images via Codex CLI's built-in imagegen tool (gpt-image-2).

byungjunjang/slide-master280—~401Automated safety check: PassMIT9 days ago
26

Generate or edit images using OpenAI GPT Image API (gpt-image-2, gpt-image-1, etc).

feiskyer/claude-code-settings1.7k—~1.4kAutomated safety check: PassMIT11 days ago
27

Drive the VibeComfy package to discover ComfyUI workflows, load ready Python templates, edit and compose them in a VibeWorkflow IR, validate, and execute either embedded locally, against an existing…

peteromallet/VibeComfy150—~3.5kAutomated safety check: PassMIT8 days ago
28

Generate or edit images via Google Gemini (nanobanana). An agent skill from feiskyer/claude-code-settings.

feiskyer/claude-code-settings1.7k—~1.1kAutomated safety check: PassMIT11 days ago
29

A skill your agent uses whenever the user asks for a graphical abstract, mechanism illustration, study design schematic, concept explainer, scientific cover art, or any non-data academic image that…

Citrus-bit/Anaxa120—~1.4kAutomated safety check: PassMIT1 mo ago
30

This skill should be used for Python scripting and Gemini image generation.

NikiforovAll/claude-code-rules141—~1.4kAutomated safety check: PassApache-2.06 days ago
31

Generate images, videos, speech audio, and music using the PonyFlash Python SDK.

aiskillstore/marketplace4301 repo~4.7kAutomated safety check: PassMITtoday
32

A skill your agent uses when the user wants to generate or edit images with Google's Nanobanana/Gemini image models using the official Gemini API shape, or when they need publication-style…

LeoYeAI/openclaw-master-skills2.2k—~4.5kAutomated safety check: PassMIT2 mo ago
33

Generate or edit images using Google Nano Banana (Gemini image generation API).

OpenMinis/MinisSkills444—~1.5kAutomated safety check: PassMITyesterday
34

Generate professional poster design concepts and optimized image-generation prompts, then automatically run a drawing script to produce the final poster image when a user needs a poster.

aipoch/medical-research-skills2k—~1.3kAutomated safety check: PassMIT22 days ago
35
35.Fal

A skill your agent uses when calling a fal.ai endpoint by id to generate image, audio, or video from JS/Python/curl: subscribe vs submit, queue states, ED25519 webhook signature verification…

ericrisco/rsc-harness174—~2.2kAutomated safety check: PassMITyesterday
36

A skill your agent uses when writing code that calls the Gemini API for text generation, multi-turn chat, multimodal understanding, image generation, streaming responses, background research tasks…

JetBrains/skills366—~2.5kAutomated safety check: PassNo licence3 mo ago
37

Use this repo skill for DiscoArt image generation, configuration/prompt scheduling, CLI, Jina serving, Docker runtime planning, and troubleshooting.

VectorSpaceLab/AREX-Skill330—~1.3kAutomated safety check: PassUnknown1 mo ago
38

Generate or edit images using Google Gemini API via nanobanana.

Microck/ordinary-claude-skills404—~1kAutomated safety check: PassUnknown1 mo ago
39

Generate PPTs from topics or templates, edit existing presentations via natural language, or perform local file operations (delete/reorder/merge slides).

aiskillstore/marketplace430—~1.9kAutomated safety check: PassNo licencetoday
40

Generate, edit, and upscale images; create videos from images or other videos via Venice AI.

sundial-org/awesome-openclaw-skills663—~2.1kAutomated safety check: PassNo licence7 mo ago
41

generate an image, create a picture, draw something, make an image of, text to image, paint a picture, illustrate, visualize, local image generation, AI art, image synthesis, offline image…

LeoYeAI/openclaw-master-skills2.2k—~4.4kAutomated safety check: PassMIT2 mo ago
42

本地文生图、AI画图、生成图像、画一张图、帮我画、生成图片、创作图像、制作一幅图、 图像生成、文字生成图片、AI绘画、画个XX、我想要一张XX的图、本地生图、离线生图。

LeoYeAI/openclaw-master-skills2.2k—~4.7kAutomated safety check: PassMIT2 mo ago