Bailian Media Generation
modelstudioai/cli
Chinese-language entry point into Alibaba Cloud Bailian's image, video and speech generation and understanding, routed through separate image, video, speech and vision commands.
A skill your agent uses when understanding images with Alibaba Cloud Model Studio Qwen VL models (qwen3-vl-plus/qwen3-vl-flash and latest aliases).
$ npx skills add cinience/alicloud-skills --skill aliyun-qwen-vl -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install cinience/alicloud-skills aliyun-qwen-vl --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/cinience/alicloud-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/ai/multimodal/aliyun-qwen-vl .claude/skills/aliyun-qwen-vl && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "aliyun-qwen-vl" agent skill from https://github.com/cinience/alicloud-skills/tree/main/skills/ai/multimodal/aliyun-qwen-vl into .claude/skills/aliyun-qwen-vl/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "aliyun-qwen-vl", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/cinience/alicloud-skills/tree/main/skills/ai/multimodal/aliyun-qwen-vlType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add cinience/alicloud-skills --skill aliyun-qwen-vl -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install cinience/alicloud-skills aliyun-qwen-vl --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/cinience/alicloud-skills.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/ai/multimodal/aliyun-qwen-vl .agents/skills/aliyun-qwen-vl && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "aliyun-qwen-vl" agent skill from https://github.com/cinience/alicloud-skills/tree/main/skills/ai/multimodal/aliyun-qwen-vl into .agents/skills/aliyun-qwen-vl/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "aliyun-qwen-vl", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add cinience/alicloud-skills --skill aliyun-qwen-vl -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install cinience/alicloud-skills aliyun-qwen-vl --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/cinience/alicloud-skills.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/ai/multimodal/aliyun-qwen-vl .cursor/skills/aliyun-qwen-vl && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "aliyun-qwen-vl" agent skill from https://github.com/cinience/alicloud-skills/tree/main/skills/ai/multimodal/aliyun-qwen-vl into .cursor/skills/aliyun-qwen-vl/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "aliyun-qwen-vl", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/cinience/alicloud-skills.git --path skills/ai/multimodal/aliyun-qwen-vl--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add cinience/alicloud-skills --skill aliyun-qwen-vl -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install cinience/alicloud-skills aliyun-qwen-vl --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/cinience/alicloud-skills.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/ai/multimodal/aliyun-qwen-vl .gemini/skills/aliyun-qwen-vl && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "aliyun-qwen-vl" agent skill from https://github.com/cinience/alicloud-skills/tree/main/skills/ai/multimodal/aliyun-qwen-vl into .gemini/skills/aliyun-qwen-vl/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "aliyun-qwen-vl", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install cinience/alicloud-skills aliyun-qwen-vlInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add cinience/alicloud-skills --skill aliyun-qwen-vl -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/cinience/alicloud-skills.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/ai/multimodal/aliyun-qwen-vl .github/skills/aliyun-qwen-vl && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "aliyun-qwen-vl" agent skill from https://github.com/cinience/alicloud-skills/tree/main/skills/ai/multimodal/aliyun-qwen-vl into .github/skills/aliyun-qwen-vl/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "aliyun-qwen-vl", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add cinience/alicloud-skills --skill aliyun-qwen-vl -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install cinience/alicloud-skills aliyun-qwen-vl --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/cinience/alicloud-skills.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/ai/multimodal/aliyun-qwen-vl .opencode/skills/aliyun-qwen-vl && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "aliyun-qwen-vl" agent skill from https://github.com/cinience/alicloud-skills/tree/main/skills/ai/multimodal/aliyun-qwen-vl into .opencode/skills/aliyun-qwen-vl/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "aliyun-qwen-vl", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
aliyun-qwen-vlA skill your agent uses when understanding images with Alibaba Cloud Model Studio Qwen VL models (qwen3-vl-plus/qwen3-vl-flash and latest aliases).
Aliyun Qwen Vl is an agent skill from cinience/alicloud-skills. Use when understanding images with Alibaba Cloud Model Studio Qwen VL models (qwen3-vl-plus/qwen3-vl-flash and latest aliases). Use when building image Q&A, visual analysis, OCR-like extraction, chart/table reading, or screenshot understanding workflows.
Its SKILL.md is about 1.4k tokens, which your agent loads only when the skill is triggered. The skill folder holds 9 other files, including scripts and reference files (for example `agents/openai.yaml`, `references/api_reference.md` and `references/examples/invoice.schema.json`).
It sits in Documents & Office, covering Document parsing. It works with Qwen and Alibaba Cloud. The repository describes itself as: alibaba cloud skills,qwen ,wan and all skills. The licence is MIT.
4 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 1818263. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships 1 file in scripts/ (Python), which the agent can run.
Shell commands in SKILL.md call:
pythonpython3curlFrom the folder's file list and the shell code blocks in SKILL.md.
Hosts in commands or code, which the agent is likely to contact:
dashscope.aliyuncs.comFrom URLs in SKILL.md, links to its own repository left out.
Names these keys or tokens, usually read from environment variables:
DASHSCOPE_API_KEYFrom names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Aliyun Qwen Vl loads about 1.4k tokens when it runs, and up to ~2k if it reads all its reference files. Until then it costs about 67 tokens; SKILL.md has 404 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
The full file from cinience/alicloud-skills at commit 1818263, republished under its MIT licence (© cinience). 404 words, ~1,425 tokens.
.claude/skills/aliyun-qwen-vl/SKILL.md (or your agent's skills folder). This skill also uses 5 other files; get the full folder from GitHub.Category: provider
mkdir -p output/aliyun-qwen-vl
python -m py_compile skills/ai/multimodal/aliyun-qwen-vl/scripts/analyze_image.py && echo "py_compile_ok" > output/aliyun-qwen-vl/validate.txtPass criteria: command exits 0 and output/aliyun-qwen-vl/validate.txt is generated.
output/aliyun-qwen-vl/.Use Qwen VL models for image input + text output understanding tasks via DashScope compatible-mode API.
python3 -m venv .venv
. .venv/bin/activate
python -m pip install requestsDASHSCOPE_API_KEY in environment, or add dashscope_api_key to ~/.alibabacloud/credentials.Prefer the Qwen3 VL family:
qwen3-vl-plusqwen3-vl-flashWhen you need explicit "latest" routing or reproducible snapshots, use supported aliases/snapshots from the official model list, such as:
qwen3-vl-plus-latestqwen3-vl-plus-2025-12-19qwen3-vl-flash-2026-01-22qwen3-vl-flash-latestLegacy names still seen in some workloads:
qwen-vl-max-latestqwen-vl-plus-latestFor OCR-specialized extraction, prefer skills/ai/multimodal/aliyun-qwen-ocr/ instead of using the general VL skill.
prompt (string, required): user question/instruction about image.image (string, required): HTTPS URL, local path, or data: URL.model (string, optional): default qwen3-vl-plus.max_tokens (int, optional): default 512.temperature (float, optional): default 0.2.detail (string, optional): auto/low/high, default auto.json_mode (bool, optional): return JSON-only response when possible.schema (object, optional): JSON Schema for structured extraction.max_retries (int, optional): retry count for 429/5xx, default 2.retry_backoff_s (float, optional): exponential backoff base seconds, default 1.5.text (string): primary model answer.model (string): model actually used.usage (object): token usage if returned by backend.python skills/ai/multimodal/aliyun-qwen-vl/scripts/analyze_image.py \
--request '{"prompt":"Summarize the main content in this image","image":"https://example.com/demo.jpg"}' \
--print-responseUsing local image:
python skills/ai/multimodal/aliyun-qwen-vl/scripts/analyze_image.py \
--request '{"prompt":"Extract key information from the image","image":"./samples/invoice.png","model":"qwen3-vl-plus"}' \
--print-responseStructured extraction (JSON mode):
python skills/ai/multimodal/aliyun-qwen-vl/scripts/analyze_image.py \
--request '{"prompt":"Extract fields: title, amount, date","image":"./samples/invoice.png"}' \
--json-mode \
--print-responseStructured extraction (JSON Schema):
python skills/ai/multimodal/aliyun-qwen-vl/scripts/analyze_image.py \
--request '{"prompt":"Extract invoice fields","image":"./samples/invoice.png"}' \
--schema skills/ai/multimodal/aliyun-qwen-vl/references/examples/invoice.schema.json \
--print-responsecurl -sS https://dashscope.aliyuncs.com/compatible-mode/v1/chat/completions \
-H "Authorization: Bearer $DASHSCOPE_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model":"qwen3-vl-plus",
"messages":[
{
"role":"user",
"content":[
{"type":"image_url","image_url":{"url":"https://example.com/demo.jpg"}},
{"type":"text","text":"Describe this image and list executable actions"}
]
}
],
"max_tokens":512,
"temperature":0.2
}'--output is set, JSON response is saved to that file.output/aliyun-qwen-vl/.python tests/ai/multimodal/aliyun-qwen-vl-test/scripts/smoke_test_qwen_vl.py \
--image ./tmp/vl_test_cat.png| Error | Likely cause | Action |
|---|---|---|
| 401/403 | Missing or invalid key | Check DASHSCOPE_API_KEY and account permissions. |
| 400 | Invalid request schema or unsupported image source | Validate messages content and image URL/path format. |
| 429 | Rate limit | Retry with exponential backoff and lower concurrency. |
| 5xx | Temporary backend issue | Retry with backoff and idempotent request design. |
-latest.references/sources.mdreferences/api_reference.md© cinience, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 5 other files (scripts, references) in skills/ai/multimodal/aliyun-qwen-vl of cinience/alicloud-skills.
Open the folder on GitHubat commit 1818263
Aliyun Qwen Vl next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Aliyun Qwen Vl this skillcinience/alicloud-skills | 397 | — | ~1.4k | Automated safety check: Pass | MIT | |
| Bailian Media Generationmodelstudioai/cli | 542 | — | ~2k | Automated safety check: Pass | Apache-2.0 | |
| Dashscopecalesthio/OpenMontage | 66k | — | ~1.5k | Automated safety check: Notes | AGPL-3.0 | |
| DOCX ToolkitXiaomiMiMo/MiMo-Code | 14k | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | |
| Office File Processxstongxue/best-skills | 3k | — | ~1.8k | Automated safety check: Pass | Proprietary | |
| PDF ToolkitXiaomiMiMo/MiMo-Code | 14k | — | ~1.7k | Automated safety check: Pass | Apache-2.0 |
modelstudioai/cli
Chinese-language entry point into Alibaba Cloud Bailian's image, video and speech generation and understanding, routed through separate image, video, speech and vision commands.
calesthio/OpenMontage
DashScope (Alibaba Cloud Bailian / 阿里云百炼) integration — image generation (qwen-image-2.0-pro), text-to-speech (qwen3-tts-flash), and ASR with word-level timestamps (qwen3-asr-flash-filetrans).
XiaomiMiMo/MiMo-Code
Produces, edits and reads Microsoft Word files through python-docx and lxml, with a decision table for picking the lightest workflow for a given task.
xstongxue/best-skills
处理 Office 文档的一站式 skill:Word(.doc/.docx/.dotx)、Excel(.xls/.xlsx/.xlsm/.csv)、PowerPoint(.ppt/.pptx/.potx) 的创建、读取、编辑、提取、转换、校验。触发:『读取 word 文档』『提取 excel 内容』『看 ppt 讲了什么』、.doc 老格式打不开、生成/编辑 Word…
XiaomiMiMo/MiMo-Code
Reads, transforms, composes and fills PDFs with Python scripts for extraction, merging, watermarking, encryption, OCR and form filling.
AeternaLabsHQ/pullmd
Read any web page, document, or YouTube video as clean Markdown using PullMD.
cinience/alicloud-skills
A skill your agent uses when creating, migrating, or optimizing skills for this alicloud-skills repository.
cinience/alicloud-skills
Bootstrap, create, connect to, operate, secure, scale, upgrade, troubleshoot, inspect, and tear down Alibaba Cloud Container Compute Service (ACS) Agent Sandbox environments.
cinience/alicloud-skills
Create, inspect, connect to, inventory, and delete Alibaba Cloud Container Compute Service (ACS) clusters through the official CS OpenAPI.
cinience/alicloud-skills
A skill your agent uses when managing Alibaba Cloud AnalyticDB for MySQL (ADB) via OpenAPI/SDK, including the user needs AnalyticDB resource lifecycle and configuration operations, status checks, or…
cinience/alicloud-skills
A skill your agent uses when managing Alibaba Cloud AIContent (AiContent) via OpenAPI/SDK, including the user needs AI content generation or content workflow operations in Alibaba Cloud, including…
cinience/alicloud-skills
A skill your agent uses when managing Alibaba Cloud Quan Miao (AiMiaoBi) via OpenAPI/SDK, including the user asks for Alibaba Cloud MiaoBi content operations, including listing resources…
Works with
Categories
A skill your agent uses when understanding images with Alibaba Cloud Model Studio Qwen VL models (qwen3-vl-plus/qwen3-vl-flash and latest aliases). Aliyun Qwen Vl is an agent skill from cinience/alicloud-skills. Use when understanding images with Alibaba Cloud Model Studio Qwen VL models (qwen3-vl-plus/qwen3-vl-flash and latest aliases).
Aliyun Qwen Vl fits situations like: understanding images with Alibaba Cloud Model Studio Qwen VL models (qwen3-vl-plus/qwen3-vl-flash and latest aliases); building image Q&A; visual analysis; OCR-like extraction.
Run `npx skills add cinience/alicloud-skills --skill aliyun-qwen-vl -a claude-code`. Or copy the skill folder (skills/ai/multimodal/aliyun-qwen-vl in cinience/alicloud-skills) into .claude/skills/aliyun-qwen-vl in your project. Claude Code loads it when a task matches its description.
Run `npx skills add cinience/alicloud-skills --skill aliyun-qwen-vl -a codex`. Or copy the skill folder (skills/ai/multimodal/aliyun-qwen-vl in cinience/alicloud-skills) into .agents/skills/aliyun-qwen-vl in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add cinience/alicloud-skills --skill aliyun-qwen-vl -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/aliyun-qwen-vl, .gemini/skills/aliyun-qwen-vl, .github/skills/aliyun-qwen-vl and .opencode/skills/aliyun-qwen-vl in your project.
Going by SKILL.md and its folder, Aliyun Qwen Vl needs Python for the scripts in its folder, the command-line tools its instructions call (python, python3 and curl) and credentials named DASHSCOPE_API_KEY. Our summary lists: Python 3; A credential in DASHSCOPE_API_KEY.
SKILL.md names 1 domain. In commands or code: dashscope.aliyuncs.com; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
Aliyun Qwen Vl is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.4k tokens (SKILL.md is roughly 5.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 562 tokens, read only when the agent opens those files.
Skills that share tags, products or a category with Aliyun Qwen Vl: Bailian Media Generation (modelstudioai/cli, 542 stars), Dashscope (calesthio/OpenMontage, 66k stars), DOCX Toolkit (XiaomiMiMo/MiMo-Code, 14k stars) and Office File Process (xstongxue/best-skills, 3k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
cinience (a GitHub user) maintains it in cinience/alicloud-skills, which has 397 GitHub stars. The repository holds 96 skills in this directory. The repository was last updated on August 11, 2026.
Source: cinience/alicloud-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.