Agent skill

Sglang Diffusion Video

by LeoYeAI in LeoYeAI/openclaw-master-skills

Generate videos using a local SGLang-Diffusion server (Wan2.2, Hunyuan, FastWan, etc.).

MITAuto-check passedMedia & Creative

Install Sglang Diffusion Video

skills CLI
$ npx skills add LeoYeAI/openclaw-master-skills --skill sglang-diffusion-video -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install LeoYeAI/openclaw-master-skills sglang-diffusion-video --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/LeoYeAI/openclaw-master-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/sglang-diffusion-video .claude/skills/sglang-diffusion-video && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
sglang-diffusion-video
GitHub stars
2.2k
Token cost
~676 tokens
SKILL.md length
127 words
Files
4 (incl. scripts)
Skills in repo
1,235
Repo updated
First seen
Licence
MIT

At a glance

Generate videos using a local SGLang-Diffusion server (Wan2.2, Hunyuan, FastWan, etc.).

  • : user asks to generate
  • SKILL.md covers Prerequisites, Generate a video, Useful flags and API key (optional), plus 1 more section
  • Runs Python scripts from its folder; calls python3; needs SGLANG_DIFFUSION_API_KEY
  • Render a video with a locally running SGLang-Diffusion instance

What it does

Sglang Diffusion Video is an agent skill from LeoYeAI/openclaw-master-skills. Generate videos using a local SGLang-Diffusion server (Wan2.2, Hunyuan, FastWan, etc.). Use when: user asks to generate, create, or render a video with a locally running SGLang-Diffusion instance. NOT for: cloud-hosted video APIs or image generation (use sglang-diffusion for images). Requires a running SGLang-Diffusion server with a video model loaded.

Its SKILL.md is about 680 tokens, which your agent loads only when the skill is triggered. The skill folder holds 5 other files, including scripts (for example `.clawhub/origin.json`, `_meta.json` and `scripts/generate_video.py`).

It sits in Media & Creative, covering AI video generation and Image generation. It works with SGLang. The repository describes itself as: 🧠 Curated collection of 1209+ best OpenClaw skills — weekly updated by MyClaw.ai. The licence is MIT.

When your agent uses it

  • : user asks to generate
  • Render a video with a locally running SGLang-Diffusion instance

Example prompts

  • “/sglang-diffusion-video”

Requirements

  • Python 3
  • A credential in SGLANG_DIFFUSION_API_KEY

What it can do on your machine

Read from SKILL.md and the folder at commit e5199b5. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • SGLANG_DIFFUSION_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Sglang Diffusion Video loads about 676 tokens when it runs. Until then it costs about 94 tokens; SKILL.md has 127 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~94
When it runs · the whole SKILL.md, loaded when a task matches
~676

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from LeoYeAI/openclaw-master-skills at commit e5199b5, republished under its MIT licence (© LeoYeAI). 127 words, ~676 tokens.

Download SKILL.mdSave it as .claude/skills/sglang-diffusion-video/SKILL.md (or your agent's skills folder). This skill also uses 3 other files; get the full folder from GitHub.
name
sglang-diffusion-video
description
Generate videos using a local SGLang-Diffusion server (Wan2.2, Hunyuan, FastWan, etc.). Use when: user asks to generate, create, or render a video with a locally running SGLang-Diffusion instance. NOT for: cloud-hosted video APIs or image generation (use sglang-diffusion for images). Requires a running SGLang-Diffusion server with a video model loaded.
homepage
https://github.com/sgl-project/sglang

SGLang-Diffusion Video Generation

Generate videos via a local SGLang-Diffusion server's OpenAI-compatible API.

Video generation is asynchronous and takes several minutes. The script handles submission, polling, and download automatically.

Prerequisites

  • SGLang-Diffusion server running a video model (default: http://127.0.0.1:30000)
  • Supported models: Wan2.2-T2V, Wan2.2-I2V, FastWan, Hunyuan
  • If the server was started with --api-key, set SGLANG_DIFFUSION_API_KEY env var

Generate a video

bash
python3 {baseDir}/scripts/generate_video.py --prompt "a curious raccoon exploring a garden"

Useful flags

bash
python3 {baseDir}/scripts/generate_video.py --prompt "ocean waves at sunset" --size 1280x720
python3 {baseDir}/scripts/generate_video.py --prompt "city timelapse" --negative-prompt "blurry, low quality"
python3 {baseDir}/scripts/generate_video.py --prompt "dancing robot" --steps 50 --guidance-scale 7.5 --seed 42
python3 {baseDir}/scripts/generate_video.py --prompt "flying through clouds" --seconds 8 --fps 24 --out ./my-video.mp4
python3 {baseDir}/scripts/generate_video.py --prompt "flying through clouds" --server http://192.168.1.100:30000 --out ./my-video.mp4
python3 {baseDir}/scripts/generate_video.py --prompt "cat playing" --poll-interval 15 --timeout 1800
python3 {baseDir}/scripts/generate_video.py --prompt "animate this scene" --input-image /tmp/scene.png

API key (optional)

Only needed if the SGLang-Diffusion server was started with --api-key. Set SGLANG_DIFFUSION_API_KEY, or pass --api-key directly:

bash
python3 {baseDir}/scripts/generate_video.py --prompt "hello" --api-key sk-my-key

Or configure in ~/.openclaw/openclaw.json:

json5
{
 skills: {
   "sglang-diffusion-video": {
     env: { SGLANG_DIFFUSION_API_KEY: "sk-my-key" },
   },
 },
}

Notes

  • The script prints a MEDIA: line for OpenClaw to auto-attach on supported chat providers.
  • Output defaults to timestamped MP4 in /tmp/.
  • Video generation typically takes 5-15 minutes depending on GPU and model size.
  • Do not read the video back; report the saved path only.

© LeoYeAI, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 3 other files (scripts) in skills/sglang-diffusion-video of LeoYeAI/openclaw-master-skills.

  • SKILL.md
  • .clawhub/origin.json
  • _meta.json
  • scripts/generate_video.py

Open the folder on GitHubat commit e5199b5

Compare with similar skills

Sglang Diffusion Video next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Sglang Diffusion Video compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Sglang Diffusion Video this skillLeoYeAI/openclaw-master-skills2.2k—~676Automated safety check: PassMIT
SN Motion HTMLOpenSenseNova/SenseNova-Skills5.7k—~2.2kAutomated safety check: NotesMIT
Wedding Video Guided Wizardaaronyi97/wedding-video-guided-wizard310—~1kAutomated safety check: PassMIT
WorkrallyTencent/workrally166—~3.7kAutomated safety check: PassMIT-0
Gc Still Image Motion DirectorLiamGvchi/gc-still-image-motion-director152—~1.4kAutomated safety check: PassMIT
RunninghubHM-RunningHub/OpenClaw_RH_Skills142—~1.6kAutomated safety check: PassApache-2.0

Similar skills

  • SN Motion HTML

    OpenSenseNova/SenseNova-Skills

    Builds HTML stories where one continuous camera journey advances with page progress, using researched structure, AI stills, Seedance video clips and browser QA.

    5.7k GitHub stars~2.2k tokensUpdated yesterday
    Media & CreativeAuto-check: notes
  • Wedding Video Guided Wizard

    aaronyi97/wedding-video-guided-wizard

    Guide a creator through a real couple's custom wedding video, from a shareable story intake card and Kimi writing pack through narration, external GPT image prompts, image-to-video packs, music and…

    310 GitHub stars~1k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Workrally

    Tencent/workrally

    WorkRally CLI (workrally) — 面向 AI Agent 的 AIGC 漫剧视频创作全流程工具集。支持 AI 生图、AI 生视频、视频提示词优化、画布生音频/音乐、混元 3D 模型生成、AI 生音频、项目/剧集/场次/分镜的完整 CRUD、资产库、媒资管理、无限画布、文件上传下载等。Use when user asks to generate images…

    166 GitHub stars~3.7k tokensUpdated 10 days ago
    Media & CreativeAuto-check passed
  • Gc Still Image Motion Director

    LiamGvchi/gc-still-image-motion-director

    Analyze still images and design restrained, image-specific motion for image-to-video generation.

    152 GitHub stars~1.4k tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed
  • Runninghub

    HM-RunningHub/OpenClaw_RH_Skills

    Generate images, videos, audio, and 3D models via RunningHub API (420 endpoints) and run any RunningHub AI Application (custom ComfyUI workflow) by webappId.

    142 GitHub stars~1.6k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Media Gen

    clacky-ai/openclacky

    Generate or edit images, videos, or audio in the current task.

    1.2k GitHub stars~7.3k tokensUpdated today
    Media & CreativeAuto-check passed

More from LeoYeAI/openclaw-master-skills

All 1,235 skills in this repo
  • DevOps Pipeline Management

    LeoYeAI/openclaw-master-skills

    Manages pipelines on a DevOps quality and efficiency platform through its OpenAPI: list workspaces and templates, create, update, run and cancel pipelines, and read run records.

    2.2k GitHub stars~4.2k tokensUpdated 2 mo ago
    Auto-check: notes
  • Feishu Document Collaboration

    LeoYeAI/openclaw-master-skills

    Patches OpenClaw's Feishu extension so an edited document triggers an isolated agent session that reads the doc and replies inline, turning it into a live chat space.

    2.2k GitHub stars~2k tokensUpdated 2 mo ago
    Auto-check passed
  • Files Memory System

    LeoYeAI/openclaw-master-skills

    Multi-context memory management system for OpenClaw agents with group-isolated storage, global shared memory, workspace organization, and group-specific skills isolation.

    2.2k GitHub stars~3.8k tokensUpdated 2 mo ago
    Auto-check passed
  • GEO-Claw AI Visibility Agent

    LeoYeAI/openclaw-master-skills

    Runs a brand's AI-search visibility work end to end: diagnosing how AI platforms represent it, repositioning it, producing AI-optimized content and monitoring ongoing mentions.

    2.2k GitHub stars~4.7k tokensUpdated 2 mo ago
    Auto-check passed
  • Google Workspace CLI

    LeoYeAI/openclaw-master-skills

    Installs and authenticates the gws CLI, then automates Gmail, Drive, Sheets, Calendar, Docs, Chat and Tasks with ready-made recipes, persona bundles and security audits.

    2.2k GitHub stars~2.6k tokensUpdated 2 mo ago
    Auto-check: notes
  • HealthFit Health Advisors

    LeoYeAI/openclaw-master-skills

    Runs four advisor roles, a fitness coach, nutritionist, data analyst and TCM practitioner, to build a health profile and track workouts, diet and wellness over time.

    2.2k GitHub stars~4.4k tokensUpdated 2 mo ago
    Auto-check passed

Works with

Questions about Sglang Diffusion Video

What does Sglang Diffusion Video do?

Generate videos using a local SGLang-Diffusion server (Wan2.2, Hunyuan, FastWan, etc.). Sglang Diffusion Video is an agent skill from LeoYeAI/openclaw-master-skills.).

When should I use Sglang Diffusion Video?

Sglang Diffusion Video fits situations like: : user asks to generate; render a video with a locally running SGLang-Diffusion instance.

How do I install Sglang Diffusion Video in Claude Code?

Run `npx skills add LeoYeAI/openclaw-master-skills --skill sglang-diffusion-video -a claude-code`. Or copy the skill folder (skills/sglang-diffusion-video in LeoYeAI/openclaw-master-skills) into .claude/skills/sglang-diffusion-video in your project. Claude Code loads it when a task matches its description.

How do I install Sglang Diffusion Video in Codex?

Run `npx skills add LeoYeAI/openclaw-master-skills --skill sglang-diffusion-video -a codex`. Or copy the skill folder (skills/sglang-diffusion-video in LeoYeAI/openclaw-master-skills) into .agents/skills/sglang-diffusion-video in your project. Codex loads it when a task matches its description.

Can I use Sglang Diffusion Video in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add LeoYeAI/openclaw-master-skills --skill sglang-diffusion-video -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/sglang-diffusion-video, .gemini/skills/sglang-diffusion-video, .github/skills/sglang-diffusion-video and .opencode/skills/sglang-diffusion-video in your project.

What does Sglang Diffusion Video need to run?

Going by SKILL.md and its folder, Sglang Diffusion Video needs Python for the scripts in its folder, the command-line tools its instructions call (python3) and credentials named SGLANG_DIFFUSION_API_KEY. Our summary lists: Python 3; A credential in SGLANG_DIFFUSION_API_KEY.

Does Sglang Diffusion Video access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Sglang Diffusion Video safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Sglang Diffusion Video use?

Sglang Diffusion Video is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Sglang Diffusion Video use?

About 676 tokens (SKILL.md is roughly 2.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Sglang Diffusion Video?

Skills that share tags, products or a category with Sglang Diffusion Video: SN Motion HTML (OpenSenseNova/SenseNova-Skills, 5.7k stars), Wedding Video Guided Wizard (aaronyi97/wedding-video-guided-wizard, 310 stars), Workrally (Tencent/workrally, 166 stars) and Gc Still Image Motion Director (LiamGvchi/gc-still-image-motion-director, 152 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Sglang Diffusion Video?

LeoYeAI (a GitHub user) maintains it in LeoYeAI/openclaw-master-skills, which has 2,161 GitHub stars. The repository holds 1,235 skills in this directory. The repository was last updated on July 20, 2026.

Source: LeoYeAI/openclaw-master-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.