Search
Media & Creative · Google Gemini · For developers
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Recommend suitable prompts from 10,000+ Nano Banana Pro image generation prompts based on user needs. | YouMind-OpenLab/ | 1.9k | 1 repo | ~4.1k | Automated safety check: Pass | No licence | today |
| 2 | Gives text-only models sight by running the modlens CLI on an image path or URL and returning structured JSON evidence with transcribed text, layout and semantics. | liustack/ | 4.2k | — | ~1.3k | Automated safety check: Notes | MIT | 7 days ago |
| 3 | Generates game art from text prompts: PNG images, GLB 3D models, rigged characters, animations and sprites, with background removal. | htdt/ | 7.1k | — | ~2.8k | Automated safety check: Pass | MIT | 9 days ago |
| 4 | A skill your agent uses when writing code that calls the Gemini API for text generation, multi-turn chat, multimodal understanding, image generation, video generation, speech generation (TTS), voice… | google-gemini/ | 4.3k | — | ~5.1k | Automated safety check: Pass | Apache-2.0 | 4 days ago |
| 5 | Create publication-quality scientific diagrams using Nano Banana 2 AI with smart iterative refinement. | spacering-net/ | 3.9k | 11 repos | ~5.9k | Automated safety check: Notes | MIT | today |
| 6 | A skill your agent uses when building real-time, bidirectional streaming applications with the Gemini Live API, or migrating legacy Live models (2.0/2.5/3.1) to Gemini 3.8 Live. | google-gemini/ | 4.3k | — | ~4.6k | Automated safety check: Pass | Apache-2.0 | 4 days ago |
| 7 | 专业的Seedance 2.0平台AI视频脚本和分镜生成器。当用户要求:(1) 将文章/故事转换为视频脚本,(2) 生成Seedance 2.0分镜提示词,(3) 规划多集AI视频系列,(4) 为GPT-Image-2、Seedream、Nano Banana… | liangdabiao/ | 2.6k | — | ~2.2k | Automated safety check: Pass | No licence | 20 days ago |
| 8 | Plans and builds 2D game maps and scenes, from tilemaps and parallax backgrounds to HD-2D plates, with collision checks, a playable HTML preview and Tiled, Godot or LDtk export. | 0x0funky/ | 4.4k | — | ~2.9k | Automated safety check: Pass | MIT | 5 days ago |
| 9 | 将约 5 秒口播文稿、观点句或抽象概念做成高级 editorial halftone paper-collage / 半调纸拼贴 B-roll。用户说“collage b-roll”“纸拼贴 b-roll”“半调拼贴”“拼贴风格配画面”“用这段文稿做拼贴动画”“gbro-collage-broll”,或希望把一句文稿转成拼贴视觉隐喻时,必须使用此… | pyang5166/ | 1.3k | — | ~2.6k | Automated safety check: Notes | MIT | 2 mo ago |
| 10 | Generate, edit-from-reference, or analyze images with AI via OpenRouter (Gemini, GPT Image, Seedream, Qwen, MAI, Grok, FLUX.2, Recraft, Muse, Riverflow; Cloudflare AI Gateway BYOK). | centminmod/ | 2.7k | — | ~8.1k | Automated safety check: Notes | MIT | 3 days ago |
| 11 | 11.Nano Banana Generate or edit images via Nano Banana image models. An agent skill from nexu-io/nexu. | nexu-io/ | 3.3k | — | ~757 | Automated safety check: Pass | MIT | 5 mo ago |
| 12 | Transcribes audio files into text or subtitles through 9Router's Whisper-compatible endpoint, using models from OpenAI, Groq, Gemini, Deepgram and others. | decolua/ | 31k | — | ~914 | Automated safety check: Pass | MIT | 3 days ago |
| 13 | A skill your agent uses for generative video editing, text-to-video, image-referenced video generation, first-frame-to-video, first-and-last-frame transitions, and video extensions using Gemini Omni… | google-gemini/ | 4.3k | — | ~6.4k | Automated safety check: Pass | Apache-2.0 | 4 days ago |
| 14 | A skill your agent uses when the user mentions a video file (.mp4, .mov, .avi, .mkv, .webm), a YouTube URL, asks to watch/analyze/review a video, or references video content in conversation | jordanrendric/ | 1.4k | — | ~1.4k | Automated safety check: Pass | MIT | 4 days ago |
| 15 | Generate and verify consistent DeepSeek Whale-chan character illustrations from text, screenshots, chat logs, dialogue, or user reference images. | Neko3000/ | 237 | — | ~2k | Automated safety check: Pass | MIT | 3 days ago |
| 16 | 16.Gemini Skill 通过 Gemini 官网(gemini.google.com)执行生图、对话等操作。用户提到"生图/画图/绘图/nano banana/nanobanana/生成图片"等关键词时触发。操作方式分三级优先级:首选 MCP 工具 → 次选 Skill 脚本 → 最次连接 Skill 浏览器手动操作(需用户授权)。禁止自行启动外部浏览器访问 Gemini。 | WJZ-P/ | 832 | — | ~1.1k | Automated safety check: Pass | MIT | 23 days ago |
| 17 | Generate one or more standalone Meta image-ad creatives via ChatGPT Image 2 (gpt-image-2) through the Arcads external API. | krusemediallc/ | 1.6k | — | ~2.7k | Automated safety check: Notes | MIT | 18 days ago |
| 18 | All-in-one image generation with Gemini models. An agent skill from nexu-io/nexu. | nexu-io/ | 3.3k | — | ~870 | Automated safety check: Pass | MIT | 5 mo ago |
| 19 | Build full-stack web applications powered by Google Gemini's Nano Banana & Nano Banana Pro image generation APIs. | chongdashu/ | 138 | 1 repo | ~2.2k | Automated safety check: Notes | No licence | 9 mo ago |
| 20 | Develop supplied scripts, short stories, or synopses into scene understanding, director-facing art concepts, and motivated narrative keyframes; compile scene ideas, visual references, or existing… | popopo-99/ | 570 | — | ~4.6k | Automated safety check: Pass | CC-BY-NC-4.0 | 7 days ago |
| 21 | 21.Image Image prompting skill for Nano Banana (NBP/NB2) and GPT Image 2.5 (Flare/Sunburst). | smixs/ | 494 | — | ~2.1k | Automated safety check: Pass | CC-BY-4.0 | 25 days ago |
| 22 | Produces game-ready 2D characters, creatures, props, icons and effects as master stills, sheets or clips, and exports frames for common game engines. | 0x0funky/ | 4.4k | — | ~3.6k | Automated safety check: Pass | MIT | 5 days ago |
| 23 | Generate, edit, and iterate raster images from text prompts or reference images with GPT Image, Nano Banana, or Seedream. | Negai-ai/ | 330 | — | ~1.9k | Automated safety check: Notes | Apache-2.0 | 4 mo ago |
| 24 | A skill your agent uses when the user wants to turn a reference image or text concept into a WeChat-ready animated Chinese meme GIF sticker pack: 16/24 entries, 240x240 GIFs, Chinese captions… | lisamsung/ | 104 | — | ~973 | Automated safety check: Pass | MIT | 1 mo ago |
| 25 | Propose five DeepSeek Whale-chan comic concepts from text, screenshots, images, chat logs, or reasoning traces, then generate verified comics after the user selects concepts and separately confirms… | Neko3000/ | 237 | — | ~4.4k | Automated safety check: Pass | MIT | 3 days ago |
| 26 | 26.Image When the user wants to create, generate, edit, or optimize images for marketing — blog heroes, social graphics, product mockups, profile banners, listing visuals, or brand assets. | Nexus-JPF/ | 870 | 3 repos | ~3.9k | Automated safety check: Pass | MIT | 4 days ago |
| 27 | 27.Seecut 网感口播精剪(v3.4):口播视频加动效,自动剪成"网感口播精剪风格"竖版/横版样片,交付渲染成片+分层工程包(装了剪映引擎则再出剪映分层草稿,音效放草稿;没装则音效混进成片)。当用户给了口播粗剪(+B-roll/截图/录屏)并想要"口播动效/加卡片弹字/剪成短视频/做动效样片/网感口播精剪"时用。自循环:预检→素材理解→找真证据→编排→HyperFrames构建→硬门+两两对比自迭代(≤3版)… | YeJe-cpu/ | 261 | — | ~1.5k | Automated safety check: Pass | Unknown | 10 days ago |
| 28 | 28.Gemini Image Reference guide for using google-genai Python library to generate images with gemini-3-pro-image-preview model. | tyrchen/ | 236 | — | ~1.1k | Automated safety check: Pass | No licence | 6 mo ago |
| 29 | Analyzes a video file with Google Gemini and returns a structured markdown report covering top-level summary, scene-by-scene breakdown, audio transcript (or honest "silent" note), visual details… | mikefutia/ | 101 | — | ~747 | Automated safety check: Notes | No licence | 5 mo ago |
| 30 | 30.Higgsfield A skill your agent uses whenever the user asks anything about Higgsfield AI — writing or refining video/image prompts, choosing a model (Kling, Veo, Wan, Seedance, Minimax Hailuo, DoP, Soul, Nano… | OSideMedia/ | 713 | — | ~9.1k | Automated safety check: Pass | MIT | 14 days ago |
| 31 | Optimizes image generation prompts using Subject-Context-Style structure. | shinpr/ | 172 | — | ~1.5k | Automated safety check: Pass | MIT | 3 days ago |
| 32 | 32.Video A skill your agent uses whenever the user asks to create, improve, audit, or split prompts for AI video generators (Seedance, Kling, Veo, Runway, Luma, Pika, Sora, any image-to-video system). | smixs/ | 494 | — | ~2.5k | Automated safety check: Pass | CC-BY-4.0 | 25 days ago |
| 33 | 33.Nbcraft Multi-backend Image + Video Generation CLI (nb command). An agent skill from jieyefriic/nbcraft. | jieyefriic/ | 155 | — | ~2.8k | Automated safety check: Pass | MIT | 5 mo ago |
| 34 | 34.Cdaf Read CDAF sidecar files (.cdaf) instead of processing video with vision. | UditAkhourii/ | 121 | — | ~1.9k | Automated safety check: Pass | MIT | 1 mo ago |
| 35 | 35.Bananahub Agent-native image workflow and optional prompt optimizer for /bananahub and generic agent image generation requests. | bananahub-ai/ | 118 | — | ~7.1k | Automated safety check: Pass | MIT | 4 mo ago |
| 36 | Generate one or more standalone Meta image-ad creatives via Nano Banana 2 / Nano Banana Pro (Gemini Flash Image family) through the Arcads external API. | krusemediallc/ | 1.6k | — | ~2.5k | Automated safety check: Notes | MIT | 18 days ago |
| 37 | 37.AI Video Gen Generate AI videos from text prompts using multiple provider gateways. | calesthio/ | 66k | — | ~3k | Automated safety check: Pass | AGPL-3.0 | 7 days ago |
| 38 | A skill your agent uses when writing code that calls the Gemini API for text generation, multi-turn chat, multimodal understanding, image generation, video generation, streaming responses… | Ayuilos/ | 225 | — | ~4.6k | Automated safety check: Pass | AGPL-3.0 | today |
| 39 | Generate or edit images via Gemini 3 Pro Image (Nano Banana Pro). | swarmclawai/ | 689 | — | ~481 | Automated safety check: Pass | MIT | 3 mo ago |
| 40 | 40.Fal AI This skill enables AI video generation from images AND text-to-speech voiceover generation using Fal.ai's API. | mikeOnBreeze/ | 293 | — | ~1.9k | Automated safety check: Notes | MIT | 7 mo ago |
| 41 | 41.Gflow CLI A skill your agent uses when the user wants to drive Google Flow (Veo image-to-video, Veo text-to-video, Imagen / Nano Banana image generation) from the terminal or a script — including… | ffroliva/ | 278 | — | ~4.8k | Automated safety check: Notes | MIT | today |
| 42 | 為使用者產生高品質的 AI 生圖、生影片、生音樂提示詞,並在需要時透過瀏覽器自動化實際送到目標平台。涵蓋 OiiOii、Kling 3.0/O-series、Seedance 2.0/2.5、Suno v5.5、Seedream 5.0/4.0、Vidu Q3、Midjourney V8.1、Flux 1.1 Pro / Kontext、Runway Gen-4.5 /… | Hao0321/ | 260 | — | ~5.6k | Automated safety check: Pass | MIT | 2 mo ago |
| 43 | 43.Gemini Tts Generates spoken MP3 audio from text or Markdown with Gemini TTS. | iurysza/ | 420 | — | ~968 | Automated safety check: Pass | MIT | 3 days ago |
| 44 | 44.Transcribe Transcribe video and audio files via Gemini API. An agent skill from comol/ai_rules_1c. | comol/ | 481 | — | ~960 | Automated safety check: Notes | No licence | 3 days ago |
| 45 | 45.AI Video Gen Generate AI videos from text prompts using multiple provider gateways. | calesthio/ | 66k | — | ~2.8k | Automated safety check: Pass | AGPL-3.0 | 7 days ago |
| 46 | A skill your agent uses when generating or editing images from Flutter/Dart with Firebase AI Logic and a Gemini image model (Nano Banana), making the first call work, choosing Gemini Developer API… | evanca/ | 651 | — | ~2.3k | Automated safety check: Pass | MIT | yesterday |
| 47 | 47.Atlas Cloud Generate or edit images and videos through the Atlas Cloud gateway. | calesthio/ | 66k | — | ~1.2k | Automated safety check: Pass | AGPL-3.0 | 7 days ago |
| 48 | 48.Gemini Audio Guide for implementing Google Gemini API audio capabilities - analyze audio with transcription, summarization, and understanding (up to 9.5 hours), plus generate speech with controllable TTS. | einverne/ | 121 | — | ~2k | Automated safety check: Notes | MIT | 1 mo ago |