Category

Best media and creative skills, page 12

Skills #529–576 of 4,318, ranked by score.

Media & Creative skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Media & Creative skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
529

制作与持续修改 SVG 页面工作台,从逐页内容稿或 video-idea-system 的 visual-plan.json 与素材出发,输出可编辑 PPT 或固定版本的视频画面快照。用于页面返修、视觉参考复用及需人工确认的模板提炼入库;复杂原始材料先交 planners-bypage 展开。

thePlannerIvan/planners-ppt-hell234—~875Automated safety check: PassAGPL-3.0today
530

Generate speech audio from text using Qwen3 TTS, or clone a voice from reference audio.

second-state/qwen3_tts_rs233—~1.6kAutomated safety check: PassNo licence4 mo ago
531

Agent skill for running registered ComfyUI workflows through a stable CLI, and for importing a user's own ComfyUI workflow into their private registry after review.

MieMieeeee/comfyui-agent-skill116—~3.9kAutomated safety check: PassApache-2.0today
532

Add or modify end-of-speech integrations in assistant-api with strict separation from VAD internals.

rapidaai/voice-ai745—~877Automated safety check: PassUnknown3 days ago
533

Generates and prompts video clips on the filmmaking canvas. An agent skill from Utopai-Research/pai-code.

Utopai-Research/pai-code357—~4.1kAutomated safety check: PassUnknown10 days ago
534

Write and critique viral hooks for short-form video: the opening 1 to 3 seconds that decide whether anything else gets seen.

vyralcontent/content-skills1341 repo~2.5kAutomated safety check: PassMIT3 mo ago
535

Renders premium 3D product films in headless Blender from JSON shot specs, building the product from a photo or procedurally, then finishing titles, music and cuts.

edenfunf/reelmimic1.9k—~2.6kAutomated safety check: PassMITtoday
536

Generate or edit raster images through a configurable OpenAI-compatible Image API using gpt-image-2.

fengfengzhidao/codex-image2-skill144—~1.2kAutomated safety check: PassMIT2 mo ago
537
537.Deprecate APIOfficial

Reference for how deprecation of public Remotion APIs is represented in TypeScript source and documentation.

remotion-dev/remotion63k—~529Automated safety check: PassUnknowntoday
538

A skill your agent uses when one or more photographs must become adaptive photo-plus-abstraction editorial compositions while source facts, spatial relationships, and strict photo preservation…

kwhi6693-web/photo-abstract-editorial113—~870Automated safety check: PassAGPL-3.01 mo ago
539

Generate images, videos, audio, and 3D models via RunningHub API (420 endpoints) and run any RunningHub AI Application (custom ComfyUI workflow) by webappId.

HM-RunningHub/OpenClaw_RH_Skills142—~1.6kAutomated safety check: PassApache-2.01 mo ago
540

Translate an existing Remotion (React-based) video composition into a HyperFrames HTML composition.

boraoztunc/skills398—~2.2kAutomated safety check: PassApache-2.01 mo ago
541
541.Comfy

Generate images, video, audio, and 3D with Comfy Cloud — search hundreds of models and workflow templates, run custom ComfyUI workflows, and manage generation jobs through the hosted Comfy Cloud MCP…

Comfy-Org/comfy-skills222—~1.4kAutomated safety check: PassMITyesterday
542

Separate a person, product or hand from the background in a video, on the user's own computer with SAM 2.

Barty-Bart/motion-graphics559—~2.4kAutomated safety check: PassUnknown2 days ago
543

制作或精修不露脸商业与消费调查视频,覆盖选题、事实核查、10分钟以上叙事、真实动态素材、配音字幕、CTA,以及AI封面和各平台发布文案。按当前请求执行阶段,保留已确认稿件和品牌;不自动公开发布。

trustfuture/simon-skills382—~1.3kAutomated safety check: PassMIT16 days ago
544

Plan, art-direct, generate, edit, and quality-check tactile minimal-zine bitmap imagery from almost any subject: landscapes, portraits, objects, products, architecture, interiors, existing…

jiahuiqu17/paper-signal109—~2.6kAutomated safety check: PassMIT1 mo ago
545
545.Kkclaw

给你的 AI Agent 一个桌面身体 — Setup Wizard、14情绪球体、语音克隆、歌词窗、Doctor 自检、跨平台支持(Windows + macOS)

kk43994/kkclaw174—~576Automated safety check: PassMIT5 mo ago
546

Generate or plan publish-grade 16:9 article illustrations using folded-paper action characters, with Chinese-first article/body-image workflows.

Alexsun1one/paper-operators109—~4.4kAutomated safety check: PassMIT3 mo ago
547

Reference guide for using google-genai Python library to generate images with gemini-3-pro-image-preview model.

tyrchen/geektime-bootcamp-ai236—~1.1kAutomated safety check: PassNo licence6 mo ago
548

A skill your agent uses when running the full idea workflow: capture a rough idea, expand it into design/UI/implementation docs, research similar products, and generate build-ready Markdown artifacts.

AkoliteZA/hermes-agent-idea-workflow272—~3.9kAutomated safety check: PassMIT5 mo ago
549

MiniMax multimodal model skill — use MiniMax Multi-Modal models for speech, music, video, and image.

poco-ai/poco-claw1.4k—~7.6kAutomated safety check: PassMIT18 days ago
550

Design, derive, optimize, explain, and diagnose cinematic Eastern xianxia camera direction, storyboards, motion choreography, pacing, continuity locks, keyframes, and per-shot image-to-video prompts.

liyue-aigc/xianxia-cinematic-video-director269—~1.4kAutomated safety check: PassNo licence1 mo ago
551

Run an SEO page observation-action-review loop with persistent Memory by coordinating demand research, page creation, adversarial review, image generation, IndexNow submission, and performance review.

tsingyuai/growth-lab2k—~1.2kAutomated safety check: PassApache-2.011 days ago
552

Canonical VideoStudio review authorization and state-transition policy.

Orkas-AI/Orkas-VideoStudio499—~2.7kAutomated safety check: PassMIT18 days ago
553

Author and review Noon animation scenes using capability-qualified ManimCE-style Python or the shared Rust API.

yongkyuns/noon133—~2kAutomated safety check: PassNo licencetoday
554

Authors or edits a custom HyperFrames video composition when no specialized workflow fits, such as multi-scene pieces, montages and brand reels.

heygen-com/hyperframes60k3 repos~5.4kAutomated safety check: PassApache-2.0today
555

Lays timed graphic cards such as lower-thirds, data callouts and quotes over an existing talking-head video, synced to the transcript.

heygen-com/hyperframes60k3 repos~16kAutomated safety check: PassApache-2.0today
556

把已经写好的字幕(SRT / 配音稿 / 解说词 / 旁白稿)变成 Vox 风格解释视频的分镜与提示词包——参考图提示词、图生视频提示词、分镜表、关键词台账、视觉圣经、风格选择。只交付文本提示词,不生成任何图片或视频。适用于用户已有成片旁白、想做成分镜或配画面、需要切分镜头与处理时长差值、选定视觉风格,或出图/出片失败后诊断修复(字错了、画面没动、风格跑偏等)。

geeklee/srt-vox-director105—~1.8kAutomated safety check: PassNo licence2 mo ago
557

Generate and edit images from the CLI using picture-it. An agent skill from geongeorge/picture-it.

geongeorge/picture-it132—~3.7kAutomated safety check: PassMIT6 mo ago
558

This skill should be used when the user asks to "make an audiobook", "narration script", "narrator", "ACX", "Findaway", "pronunciation guide", "how long is the audiobook", "adapt to a screenplay"…

danjdewhurst/story-skills2861 repo~3.5kAutomated safety check: NotesMITyesterday
559

Choose GPT-Image2 / gpt-image-2 visual styles and industrial prompt templates from the awesome-gpt-image-2 style library.

OWWZO/ai-agent1891 repo~881Automated safety check: PassNo licencetoday
560

A skill your agent uses when an agent needs Dreamina(即梦) generation, task querying, account checks, or login/session operations through the packaged Python wrapper scripts around the dreamina CLI.

yuyou-dev/dreamina-cli-skill129—~1.1kAutomated safety check: PassNo licence2 mo ago
561
561.Bench

mlx-serve benchmarking methodology — bench.sh/llmprobe usage, comparison-trap rules (same-methodology cells only, spec-decode variance, thermal lies, engine naming), perf-claim etiquette.

ddalcu/mlx-serve1.8k1 repo~1.1kAutomated safety check: PassUnknowntoday
562

Turns an idea, article, outline or audio file into a sourced, reviewable AI video, tracking whether narration uses a human, synthetic or cloned voice.

wanghui2323/ai-video-maker101—~924Automated safety check: PassMIT1 mo ago
563

Replaces Codex's built-in image tool with provider-based generation and editing, fetching reference images first whenever visual accuracy actually matters.

yc-duan/api-image101—~4.9kAutomated safety check: PassMIT5 mo ago
564

Analyzes a video file with Google Gemini and returns a structured markdown report covering top-level summary, scene-by-scene breakdown, audio transcript (or honest "silent" note), visual details…

mikefutia/claude-vision101—~747Automated safety check: NotesNo licence5 mo ago
565

Create, audit, and repair short page-native diary comics around an existing authorized recurring character, with story-directed page rhythm, exact dialogue, directional-surface proof, and…

ZSeven-W/craft-skills2251 repo~2.8kAutomated safety check: PassApache-2.017 days ago
566

幫 Leo 看影片。當 Leo 丟影片連結(YouTube/IG/TikTok)或本機影片檔,要摘要、分析、拆解對標時使用——Claude 不能直接吃影片,先用這個工具抽關鍵幀+逐字稿+運鏡節奏+聲音/語氣/手勢時間軸再讀。

HUANGCHIHHUNGLeo/claude-real-video2.2k—~559Automated safety check: PassMITyesterday
567

Generates still images through the submit_image tool, choosing among Fal.ai, gpt-image-2, nano-banana, MiniMax image-01 and Grok Imagine by configured keys.

0xsline/OpenChatCut2.2k—~1.3kAutomated safety check: PassAGPL-3.03 days ago
568

A skill your agent uses when the user wants to create an explainer, documentary, knowledge-sharing, news-broadcast, product-introduction, or data-report video from a topic.

calcuforge/explainer-video-maker100—~6.5kAutomated safety check: PassApache-2.023 days ago
569

STEM 领域 AI 图像生成 Skill。面向科研、教育、工程场景,生成信号通路、实验流程、机制图、细胞结构、概念信息图、学术海报、架构图等示意图。

liangdabiao/stem-illustration-skill126—~2.3kAutomated safety check: NotesNo licence2 mo ago
570

Creates, refines and exports editable OneWorks 3D geometric avatars for mascots, bots or agents, keeping one shared scene state across preview, share link and export.

oneworks-ai/avatar251—~1.7kAutomated safety check: PassMIT1 mo ago
571

Convert fragmented cultural visual materials into traceable asset indexes, visual genes, high-end brand visual systems, poster concepts, image-generation prompts, and refined poster/KV/ad/cover…

dacnay816y62-hub/culture-fragment-poster-engine249—~3.4kAutomated safety check: PassNo licence1 mo ago
572

A skill your agent uses whenever the user wants speech to sound more human, companion-like, or emotionally expressive.

NoizAI/skills526—~1.8kAutomated safety check: PassNo licence12 days ago
573

Builds podcast-style audio narration from text with Azure OpenAI's GPT Realtime Mini over WebSocket, from a Python FastAPI backend to a React player.

microsoft/skills3.1k1 repo~947Automated safety check: PassMITtoday
574
574.Tufte

Apply Edward Tufte's principles to any data visualization, chart, dashboard, or infographic.

aref-vc/tufte-claude-skill307—~1.3kAutomated safety check: PassMIT3 mo ago
575

Transcribes local audio into Markdown with Gemini 3.5 Transcribe, including speaker labels and provider timestamps.

iurysza/module-graph420—~1kAutomated safety check: PassMIT2 days ago
576

Generates 2D Remotion-based TSX video files for VidTSX from a shot or scene description, with style presets and rules that keep renders from crashing.

hassancs91/claude-faceless-shorts-creator271—~2.8kAutomated safety check: PassMIT1 mo ago