Search
Media & Creative · Alibaba Cloud · For developers
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Chinese-language entry point into Alibaba Cloud Bailian's image, video and speech generation and understanding, routed through separate image, video, speech and vision commands. | modelstudioai/ | 542 | — | ~2k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 2 | DashScope (Alibaba Cloud Bailian / 阿里云百炼) integration — image generation (qwen-image-2.0-pro), text-to-speech (qwen3-tts-flash), and ASR with word-level timestamps (qwen3-asr-flash-filetrans). | calesthio/ | 66k | — | ~1.5k | Automated safety check: Notes | AGPL-3.0 | 8 days ago |
| 3 | A skill your agent uses when generating dance or motion-transfer videos with Alibaba Cloud Model Studio AnimateAnyone (animate-anyone-gen2) using a detected character image and an action template. | cinience/ | 397 | — | ~725 | Automated safety check: Pass | MIT | 2 mo ago |
| 4 | A skill your agent uses when creating cloned voices with Alibaba Cloud Model Studio CosyVoice customization models, especially cosyvoice-v3.5-plus or cosyvoice-v3.5-flash, from reference audio and… | cinience/ | 397 | — | ~904 | Automated safety check: Pass | MIT | 2 mo ago |
| 5 | A skill your agent uses when designing custom voices with Alibaba Cloud Model Studio CosyVoice customization models, especially cosyvoice-v3.5-plus or cosyvoice-v3.5-flash, from a voice prompt plus… | cinience/ | 397 | — | ~847 | Automated safety check: Pass | MIT | 2 mo ago |
| 6 | A skill your agent uses when generating expressive portrait videos from a person image and speech audio with Alibaba Cloud Model Studio EMO (emo-v1). | cinience/ | 397 | — | ~665 | Automated safety check: Pass | MIT | 2 mo ago |
| 7 | A skill your agent uses when generating template-driven emoji videos with Alibaba Cloud Model Studio Emoji (emoji-v1) from a detected portrait image. | cinience/ | 397 | — | ~637 | Automated safety check: Pass | MIT | 2 mo ago |
| 8 | A skill your agent uses when generating lightweight talking-head portrait videos with Alibaba Cloud Model Studio LivePortrait (liveportrait) from a detected portrait image and speech audio. | cinience/ | 397 | — | ~752 | Automated safety check: Pass | MIT | 2 mo ago |
| 9 | A skill your agent uses when generating images with Model Studio DashScope SDK using Qwen Image generation models (qwen-image, qwen-image-plus, qwen-image-max, qwen-image-2.0 series and snapshots). | cinience/ | 397 | — | ~1.8k | Automated safety check: Pass | MIT | 2 mo ago |
| 10 | A skill your agent uses when editing images with Alibaba Cloud Model Studio Qwen Image Edit models (qwen-image-edit, qwen-image-edit-plus, qwen-image-edit-max, qwen-image-2.0 series and snapshots). | cinience/ | 397 | — | ~762 | Automated safety check: Pass | MIT | 2 mo ago |
| 11 | A skill your agent uses when generating human-like speech audio with Model Studio DashScope Qwen TTS models (qwen3-tts-flash, qwen3-tts-instruct-flash). | cinience/ | 397 | — | ~951 | Automated safety check: Pass | MIT | 2 mo ago |
| 12 | A skill your agent uses when real-time speech synthesis is needed with Alibaba Cloud Model Studio Qwen TTS Realtime models. | cinience/ | 397 | — | ~788 | Automated safety check: Pass | MIT | 2 mo ago |
| 13 | A skill your agent uses when cloning voices with Alibaba Cloud Model Studio Qwen TTS VC models. | cinience/ | 397 | — | ~661 | Automated safety check: Pass | MIT | 2 mo ago |
| 14 | A skill your agent uses when designing custom voices with Alibaba Cloud Model Studio Qwen TTS VD models. | cinience/ | 397 | — | ~675 | Automated safety check: Pass | MIT | 2 mo ago |
| 15 | A skill your agent uses when replacing lip sync in existing videos with Alibaba Cloud Model Studio VideoRetalk (videoretalk). | cinience/ | 397 | — | ~735 | Automated safety check: Pass | MIT | 2 mo ago |
| 16 | A skill your agent uses when generating motion videos from a person image and a reference video with DashScope Wan 2.2 animate-move model (wan2.2-animate-move). | cinience/ | 397 | — | ~1.5k | Automated safety check: Pass | MIT | 2 mo ago |
| 17 | A skill your agent uses when generating talking, singing, or presentation videos from a single character image and audio with Alibaba Cloud Model Studio digital-human model wan2.2-s2v. | cinience/ | 397 | — | ~701 | Automated safety check: Pass | MIT | 2 mo ago |
| 18 | A skill your agent uses when generating reference-based videos with Alibaba Cloud Model Studio Wan R2V models (wan2.6-r2v-flash, wan2.6-r2v). | cinience/ | 397 | — | ~686 | Automated safety check: Pass | MIT | 2 mo ago |
| 19 | A skill your agent uses when routing Alibaba Cloud Model Studio requests to the right local skill (Qwen text, coder, deep research, image, video, audio, search and multimodal skills). | cinience/ | 397 | — | ~1.4k | Automated safety check: Pass | MIT | 2 mo ago |
| 20 | A skill your agent uses when running a minimal test matrix for the Model Studio skills that exist in this repo, including image/video/audio, realtime speech, omni, visual reasoning, embedding… | cinience/ | 397 | — | ~1.3k | Automated safety check: Pass | MIT | 2 mo ago |
| 21 | A skill your agent uses when creating or managing Alibaba Cloud IMS video translation jobs via OpenAPI (subtitle/voice/face). | cinience/ | 397 | — | ~503 | Automated safety check: Pass | MIT | 2 mo ago |