Search

Media & Creative · Qwen · For developers

32 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Generates game art from text prompts: PNG images, GLB 3D models, rigged characters, animations and sprites, with background removal.

htdt/godogen7.1k—~2.8kAutomated safety check: PassMIT10 days ago
2

Extract Bilibili videos and opus/article posts into readable Markdown knowledge notes.

Rimagination/bili-note322—~3.6kAutomated safety check: PassMIT19 days ago
3

Generate, edit-from-reference, or analyze images with AI via OpenRouter (Gemini, GPT Image, Seedream, Qwen, MAI, Grok, FLUX.2, Recraft, Muse, Riverflow; Cloudflare AI Gateway BYOK).

centminmod/my-claude-code-setup2.7k—~8.1kAutomated safety check: NotesMIT3 days ago
4

给任何故事里的角色真出参考图(小说改编、自己原创的故事、单独设计一个角色都行,不需要小说原文): 一段话描述角色,拆成分层字段、补全后确认, 先出一张正面全身锚点,其余视图(大头照、90° 侧面、背面、细节、45° 大头照)都只参考这张锚点, 按需分档出图。每张图带标识、可单独重出,重出后自动标出哪些图过期。

eternityspring/shuohao-skills4.3k—~1.7kAutomated safety check: WarnApache-2.03 days ago
5

DyNote: systematically and efficiently extract raw Douyin/DY video data and analyze videos, comments, accounts, hashtags, and short-video scenes into evidence-graded learning notes, summaries…

Rimagination/dy-note172—~4.5kAutomated safety check: PassMIT3 mo ago
6

Generate speech audio from text using Qwen3 TTS, or clone a voice from reference audio.

second-state/qwen3_tts_rs233—~1.6kAutomated safety check: PassNo licence4 mo ago
7

Chinese-language entry point into Alibaba Cloud Bailian's image, video and speech generation and understanding, routed through separate image, video, speech and vision commands.

modelstudioai/cli542—~2kAutomated safety check: PassApache-2.02 days ago
8

Synthesize speech from text with Qwen TTS models. An agent skill from QianWen-AI/qianwen-ai.

QianWen-AI/qianwen-ai105—~4.2kAutomated safety check: NotesApache-2.0yesterday
9

Anime/illustration text-to-image (ANIMA 1.0, ~2B Cosmos DiT).

artokun/comfyui-mcp803—~4kAutomated safety check: PassMIT6 days ago
10

Acquire images as files — generate them with an AI image model (14 providers: OpenAI/gpt-image, Gemini, Qwen, Zhipu, Volcengine, Stability, FLUX, Ideogram, MiniMax, and more), search openly-licensed…

open-octo/octo-agent125—~3.1kAutomated safety check: NotesMITyesterday
11

A skill your agent uses when creating complete generated audio with qwen-audio-3.1-tts-next, including podcasts, radio drama, advertisements, multiple speakers, reference voices, ambience, sound…

tjxj/z-skills548—~716Automated safety check: PassNo licence19 days ago
12

Generate or edit images via OpenRouter Image API across top models (Nano Banana, GPT Image, FLUX.2, Seedream, Qwen, Grok).

rtadewald/skills185—~800Automated safety check: NotesNo licence13 days ago
13

Builds and publishes a Gradio demo on Hugging Face Spaces for a LoRA, with the pipeline, UI and settings chosen to match that LoRA's task and model card.

huggingface/skills11k2 repos~8.4kAutomated safety check: PassApache-2.03 days ago
14

播客/小宇宙 → 下载 → 转录 → 存为 Markdown 的完整工作流. An agent skill from chubbyguan/chubbyskills.

chubbyguan/chubbyskills1.2k—~1.1kAutomated safety check: NotesMIT3 days ago
15

Narration, sound effects and music beds in the voice the studio set up (ElevenLabs, Qwen3-TTS on this machine's GPU, or the owner's own voice from the recording booth): lines from the approved…

GTKottman/mortiflix-oss499—~1.8kAutomated safety check: PassAGPL-3.03 days ago
16

DashScope (Alibaba Cloud Bailian / 阿里云百炼) integration — image generation (qwen-image-2.0-pro), text-to-speech (qwen3-tts-flash), and ASR with word-level timestamps (qwen3-asr-flash-filetrans).

calesthio/OpenMontage66k—~1.5kAutomated safety check: NotesAGPL-3.08 days ago
17

Cloud GPU processing via RunPod serverless. An agent skill from digitalsamba/claude-code-video-toolkit.

digitalsamba/claude-code-video-toolkit2.2k—~2.1kAutomated safety check: NotesMITyesterday
18

Generate speech from text via POST /audio/speech, and clone a voice via POST /audio/voices.

veniceai/skills144—~3.6kAutomated safety check: PassMIT5 days ago
19

A skill your agent uses when generating images with Model Studio DashScope SDK using Qwen Image generation models (qwen-image, qwen-image-plus, qwen-image-max, qwen-image-2.0 series and snapshots).

cinience/alicloud-skills397—~1.8kAutomated safety check: PassMIT2 mo ago
20

A skill your agent uses when editing images with Alibaba Cloud Model Studio Qwen Image Edit models (qwen-image-edit, qwen-image-edit-plus, qwen-image-edit-max, qwen-image-2.0 series and snapshots).

cinience/alicloud-skills397—~762Automated safety check: PassMIT2 mo ago
21

A skill your agent uses when generating human-like speech audio with Model Studio DashScope Qwen TTS models (qwen3-tts-flash, qwen3-tts-instruct-flash).

cinience/alicloud-skills397—~951Automated safety check: PassMIT2 mo ago
22

A skill your agent uses when real-time speech synthesis is needed with Alibaba Cloud Model Studio Qwen TTS Realtime models.

cinience/alicloud-skills397—~788Automated safety check: PassMIT2 mo ago
23

A skill your agent uses when cloning voices with Alibaba Cloud Model Studio Qwen TTS VC models.

cinience/alicloud-skills397—~661Automated safety check: PassMIT2 mo ago
24

A skill your agent uses when designing custom voices with Alibaba Cloud Model Studio Qwen TTS VD models.

cinience/alicloud-skills397—~675Automated safety check: PassMIT2 mo ago
25

内容创作者与影片与视频编辑在制作视频配音、有声书或无网环境朗读时,请用此技能一键生成多语言、多音色的本地高质量音频。基于Edge-TTS引擎,零配置完全离线,高效产出专业级语音,彻底告别网络限制。

anbeime/skill7.8k—~586Automated safety check: PassNo licence2 days ago
26

A skill your agent uses when routing Alibaba Cloud Model Studio requests to the right local skill (Qwen text, coder, deep research, image, video, audio, search and multimodal skills).

cinience/alicloud-skills397—~1.4kAutomated safety check: PassMIT2 mo ago
27

A skill your agent uses when running a minimal test matrix for the Model Studio skills that exist in this repo, including image/video/audio, realtime speech, omni, visual reasoning, embedding…

cinience/alicloud-skills397—~1.3kAutomated safety check: PassMIT2 mo ago
28

Full production pipeline covering story to scenes, Z-Image start frames, Qwen Edit end frames, WAN FLF video clips, ffmpeg concatenation

artokun/comfyui-mcp803—~4.5kAutomated safety check: PassMIT6 days ago
29

Build Ideogram 4 (Ideogram Ultra) txt2img and img2img workflows with the local open-weights model, dual conditional/unconditional models with DualModelGuider, Qwen3-VL text encoder, and structured…

artokun/comfyui-mcp803—~5.6kAutomated safety check: PassMIT6 days ago
30

Prompt Alibaba's Qwen-Image line for the typography work it is built for — quoting exact copy so the model stops inventing it, switching prompt expansion off before it rewrites your strings…

nodetool-ai/nodetool560—~1.5kAutomated safety check: PassAGPL-3.0yesterday
31

A skill your agent uses when authoring or validating PAIDF augmentation YAML configs, or running remote Cosmos Transfer (including Cosmos3 WSM controls), Cosmos Predict, image-edit, or…

NVIDIA/skills3.6k—~5.3kAutomated safety check: PassApache-2.02 days ago
32

Generate images with Venice. An agent skill from veniceai/skills.

veniceai/skills144—~4.7kAutomated safety check: PassMIT5 days ago