Search
Media & Creative · Google Gemini
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 49 | 49.Nanobanana Generate and edit images using Google Gemini 3 Pro Image (Nano Banana Pro). | ReScienceLab/ | 1.8k | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 50 | 50.Image When the user wants to create, generate, edit, or optimize images for marketing — blog heroes, social graphics, product mockups, profile banners, listing visuals, or brand assets. | Nexus-JPF/ | 870 | 3 repos | ~3.9k | Automated safety check: Pass | MIT | 4 days ago |
| 51 | 51.Seecut 网感口播精剪(v3.4):口播视频加动效,自动剪成"网感口播精剪风格"竖版/横版样片,交付渲染成片+分层工程包(装了剪映引擎则再出剪映分层草稿,音效放草稿;没装则音效混进成片)。当用户给了口播粗剪(+B-roll/截图/录屏)并想要"口播动效/加卡片弹字/剪成短视频/做动效样片/网感口播精剪"时用。自循环:预检→素材理解→找真证据→编排→HyperFrames构建→硬门+两两对比自迭代(≤3版)… | YeJe-cpu/ | 261 | — | ~1.5k | Automated safety check: Pass | Unknown | 10 days ago |
| 52 | 52.Gemini Image Reference guide for using google-genai Python library to generate images with gemini-3-pro-image-preview model. | tyrchen/ | 236 | — | ~1.1k | Automated safety check: Pass | No licence | 6 mo ago |
| 53 | Analyzes a video file with Google Gemini and returns a structured markdown report covering top-level summary, scene-by-scene breakdown, audio transcript (or honest "silent" note), visual details… | mikefutia/ | 101 | — | ~747 | Automated safety check: Notes | No licence | 5 mo ago |
| 54 | 54.Higgsfield A skill your agent uses whenever the user asks anything about Higgsfield AI — writing or refining video/image prompts, choosing a model (Kling, Veo, Wan, Seedance, Minimax Hailuo, DoP, Soul, Nano… | OSideMedia/ | 713 | — | ~9.1k | Automated safety check: Pass | MIT | 14 days ago |
| 55 | Dedicated YouTube thumbnail generator — interviews you for exactly the style elements you want (environment, text budget, extras, accent color), then renders high-contrast, vibrant, face-consistent… | hassancs91/ | 328 | — | ~2.1k | Automated safety check: Notes | MIT | 1 mo ago |
| 56 | Optimizes image generation prompts using Subject-Context-Style structure. | shinpr/ | 172 | — | ~1.5k | Automated safety check: Pass | MIT | 3 days ago |
| 57 | Generate a branded slide-by-slide LinkedIn carousel using Gemini. | charlie947/ | 3.8k | — | ~1.7k | Automated safety check: Pass | MIT | 26 days ago |
| 58 | 58.Video A skill your agent uses whenever the user asks to create, improve, audit, or split prompts for AI video generators (Seedance, Kling, Veo, Runway, Luma, Pika, Sora, any image-to-video system). | smixs/ | 494 | — | ~2.5k | Automated safety check: Pass | CC-BY-4.0 | 25 days ago |
| 59 | 59.Nbcraft Multi-backend Image + Video Generation CLI (nb command). An agent skill from jieyefriic/nbcraft. | jieyefriic/ | 155 | — | ~2.8k | Automated safety check: Pass | MIT | 5 mo ago |
| 60 | 60.Cdaf Read CDAF sidecar files (.cdaf) instead of processing video with vision. | UditAkhourii/ | 121 | — | ~1.9k | Automated safety check: Pass | MIT | 1 mo ago |
| 61 | 61.Bananahub Agent-native image workflow and optional prompt optimizer for /bananahub and generic agent image generation requests. | bananahub-ai/ | 118 | — | ~7.1k | Automated safety check: Pass | MIT | 4 mo ago |
| 62 | Generate one or more standalone Meta image-ad creatives via Nano Banana 2 / Nano Banana Pro (Gemini Flash Image family) through the Arcads external API. | krusemediallc/ | 1.6k | — | ~2.5k | Automated safety check: Notes | MIT | 19 days ago |
| 63 | Converts a report or other document into a business story and then an infographic image, pausing for your review after each step. | TyrealQ/ | 108 | — | ~814 | Automated safety check: Notes | MIT | 17 days ago |
| 64 | Turns a story, novel or myth into a retro hand-drawn animation style short-drama script, episode breakdown, asset prompts and Seedance 2.0 storyboard prompts. | liangdabiao/ | 138 | — | ~3k | Automated safety check: Pass | No licence | 3 days ago |
| 65 | 65.AI Video Gen Generate AI videos from text prompts using multiple provider gateways. | calesthio/ | 66k | — | ~3k | Automated safety check: Pass | AGPL-3.0 | 8 days ago |
| 66 | A skill your agent uses when writing code that calls the Gemini API for text generation, multi-turn chat, multimodal understanding, image generation, video generation, streaming responses… | Ayuilos/ | 225 | — | ~4.6k | Automated safety check: Pass | AGPL-3.0 | today |
| 67 | 67.Ecom Shot Generate Nano Banana Pro (Gemini 3 Pro Image) prompts for e-commerce product photography. | MagicCube/ | 516 | — | ~1.5k | Automated safety check: Pass | No licence | 5 mo ago |
| 68 | Generates or edits PNG images with Google's Nano Banana 2 model through a small script, with a choice of 512, 1K, 2K or 4K output. | LichAmnesia/ | 234 | — | ~1.1k | Automated safety check: Notes | MIT | 4 mo ago |
| 69 | Generate or edit images via Gemini 3 Pro Image (Nano Banana Pro). | swarmclawai/ | 689 | — | ~481 | Automated safety check: Pass | MIT | 3 mo ago |
| 70 | 70.Fal AI This skill enables AI video generation from images AND text-to-speech voiceover generation using Fal.ai's API. | mikeOnBreeze/ | 293 | — | ~1.9k | Automated safety check: Notes | MIT | 7 mo ago |
| 71 | 71.Gflow CLI A skill your agent uses when the user wants to drive Google Flow (Veo image-to-video, Veo text-to-video, Imagen / Nano Banana image generation) from the terminal or a script — including… | ffroliva/ | 278 | — | ~4.8k | Automated safety check: Notes | MIT | today |
| 72 | 為使用者產生高品質的 AI 生圖、生影片、生音樂提示詞,並在需要時透過瀏覽器自動化實際送到目標平台。涵蓋 OiiOii、Kling 3.0/O-series、Seedance 2.0/2.5、Suno v5.5、Seedream 5.0/4.0、Vidu Q3、Midjourney V8.1、Flux 1.1 Pro / Kontext、Runway Gen-4.5 /… | Hao0321/ | 260 | — | ~5.6k | Automated safety check: Pass | MIT | 2 mo ago |
| 73 | 73.Gemini Tts Generates spoken MP3 audio from text or Markdown with Gemini TTS. | iurysza/ | 420 | — | ~968 | Automated safety check: Pass | MIT | 3 days ago |
| 74 | 74.Transcribe Transcribe video and audio files via Gemini API. An agent skill from comol/ai_rules_1c. | comol/ | 481 | — | ~960 | Automated safety check: Notes | No licence | 4 days ago |
| 75 | 75.AI Video Gen Generate AI videos from text prompts using multiple provider gateways. | calesthio/ | 66k | — | ~2.8k | Automated safety check: Pass | AGPL-3.0 | 8 days ago |
| 76 | A skill your agent uses when generating or editing images from Flutter/Dart with Firebase AI Logic and a Gemini image model (Nano Banana), making the first call work, choosing Gemini Developer API… | evanca/ | 651 | — | ~2.3k | Automated safety check: Pass | MIT | yesterday |
| 77 | Routes visual design requests to built-in logo, corporate identity, slide, banner, social photo and icon workflows, backed by style data and AI image generation. | assafkip/ | 112 | — | ~2.9k | Automated safety check: Pass | MIT | yesterday |
| 78 | 78.Gemini Image Automate image generation on Google Gemini (gemini.google.com) using CamoFox CLI. | redf0x1/ | 412 | — | ~1.6k | Automated safety check: Pass | MIT | 19 days ago |
| 79 | Turns a research paper into a slide deck and, optionally, a narrated demo video, through script, slide generation, text-to-speech and video assembly stages you control. | OpenLAIR/ | 1.2k | — | ~1.5k | Automated safety check: Pass | Unknown | 24 days ago |
| 80 | Generate and edit images using Google's Nano Banana Pro (Gemini 3 Pro Image) API. | rtadewald/ | 185 | — | ~1.2k | Automated safety check: Notes | No licence | 12 days ago |
| 81 | 81.Generate Nano Banana (nano-banana) image generation skill. An agent skill from buildatscale-tv/claude-code-plugins. | buildatscale-tv/ | 140 | — | ~2k | Automated safety check: Notes | No licence | 5 mo ago |
| 82 | 82.Nano Banana Generate professional presentation slides and high-quality illustrations using Gemini image generation API (Nano Banana 2), with interactive browser-based review and iterative editing. | EvoScientist/ | 478 | 2 repos | ~3.5k | Automated safety check: Pass | Apache-2.0 | 11 days ago |
| 83 | 83.Nano Banana This skill enables image generation and editing using Google's Gemini Nano Banana models. | mikeOnBreeze/ | 293 | — | ~944 | Automated safety check: Notes | MIT | 7 mo ago |
| 84 | Sets up a virtual try-on agent on Google Cloud that generates image and catwalk-video try-ons with Gemini, from first setup through local testing. | google/ | 10k | — | ~3.5k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 85 | Generate and edit images using the Gemini API (Nano Banana). | kylesnowschwartz/ | 114 | — | ~1.6k | Automated safety check: Pass | No licence | today |
| 86 | Generate/edit images with Nano Banana Pro (Gemini 3 Pro Image). | CraftOS-dev/ | 392 | 5 repos | ~1.4k | Automated safety check: Pass | MIT | 2 days ago |
| 87 | Generates and edits images with Google's Nano Banana 2 (Gemini 3.1 Flash Image) through a uv script, with a draft-then-final workflow and sizes from 512 to 4K. | steipete/ | 7.3k | — | ~1.5k | Automated safety check: Pass | MIT | yesterday |
| 88 | 88.Atlas Cloud Generate or edit images and videos through the Atlas Cloud gateway. | calesthio/ | 66k | — | ~1.2k | Automated safety check: Pass | AGPL-3.0 | 8 days ago |
| 89 | 89.Gemini Audio Guide for implementing Google Gemini API audio capabilities - analyze audio with transcription, summarization, and understanding (up to 9.5 hours), plus generate speech with controllable TTS. | einverne/ | 121 | — | ~2k | Automated safety check: Notes | MIT | 1 mo ago |
| 90 | A skill your agent uses when the user wants to reverse-engineer an existing image ad into a reusable prompt template. | krusemediallc/ | 1.6k | — | ~2.4k | Automated safety check: Notes | MIT | 19 days ago |
| 91 | Generate images using Google's Nano Banana (Gemini 2.5 Flash Image). | Abilityai/ | 109 | — | ~1.2k | Automated safety check: Notes | MIT | 3 days ago |
| 92 | Generates PNG images through OpenRouter models, with transparent backgrounds and reference-image edits, and describes existing images with multimodal vision. | evolution-foundation/ | 545 | — | ~5.1k | Automated safety check: Notes | Unknown | 5 mo ago |
| 93 | AI-powered image generation for Salesforce visuals via Nano Banana Pro. | Jaganpro/ | 424 | — | ~1.6k | Automated safety check: Pass | MIT | 5 mo ago |
| 94 | Extracts pixel, video-frame, speech, music and visual-semantic features from image, video and audio files for research datasets, using local tools or the Gemini API. | TyrealQ/ | 108 | — | ~2k | Automated safety check: Notes | MIT | 17 days ago |
| 95 | 95.Infographics Create professional infographics using Nano Banana Pro AI with smart iterative refinement. | lamm-mit/ | 246 | 6 repos | ~4.4k | Automated safety check: Notes | Apache-2.0 | 1 mo ago |
| 96 | Turns a spoken line or a full topic into editorial-collage videos of halftone cut-outs assembled piece by piece, as 5-second B-roll or a narrated explainer. | imraywang/ | 159 | — | ~2.6k | Automated safety check: Pass | Unknown | 18 days ago |