Agent skill

Video

by guaardvark in guaardvark/guaardvark

Generate video clips on the user's own GPU through Guaardvark: text-to-video, image-to-video, first+last frame animation, clips with their own soundtrack and dialogue (MiniMax H3), short looping…

MITAuto-check passedMedia & Creative

Install Video

skills CLI
$ npx skills add guaardvark/guaardvark --skill video -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install guaardvark/guaardvark video --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/guaardvark/guaardvark.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/video .claude/skills/video && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
video
GitHub stars
257
Token cost
~1.1k tokens
SKILL.md length
449 words
Files
1
Skills in repo
15
Repo updated
First seen
Licence
MIT

At a glance

Generate video clips on the user's own GPU through Guaardvark: text-to-video, image-to-video, first+last frame animation, clips with their own soundtrack and dialogue (MiniMax H3), short looping…

  • The user asks to make
  • SKILL.md covers Pick the model, One clip: MCP generate_video, Looping animation: MCP… and Batch: REST, plus 1 more section
  • Calls curl
  • Batch-generate video locally

What it does

Video is an agent skill from guaardvark/guaardvark. Generate video clips on the user's own GPU through Guaardvark: text-to-video, image-to-video, first+last frame animation, clips with their own soundtrack and dialogue (MiniMax H3), short looping animations, and batch runs. Use when the user asks to make, render, animate, or batch-generate video locally.

Its SKILL.md is about 1.1k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering AI video generation and Music and audio generation. It works with MiniMax, Model Context Protocol and ComfyUI. The repository describes itself as: Self-hosted AI studio on your own GPU: chat with your files (RAG), screen and browser agents, coding swarms, LoRA training, an MCP server for Claude Code and Cursor, and local… The licence is MIT.

When your agent uses it

  • The user asks to make
  • Batch-generate video locally

Example prompts

  • “/video”

What it can do on your machine

Read from SKILL.md and the folder at commit 2a33110. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • curl

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use curl, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Video loads about 1.1k tokens when it runs. Until then it costs about 78 tokens; SKILL.md has 449 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~78
When it runs · the whole SKILL.md, loaded when a task matches
~1.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from guaardvark/guaardvark at commit 2a33110, republished under its MIT licence (© guaardvark). 449 words, ~1,124 tokens.

Download SKILL.mdSave it as .claude/skills/video/SKILL.md (or your agent's skills folder).
name
video
description
Generate video clips on the user's own GPU through Guaardvark: text-to-video, image-to-video, first+last frame animation, clips with their own soundtrack and dialogue (MiniMax H3), short looping animations, and batch runs. Use when the user asks to make, render, animate, or batch-generate video locally.

Video with Guaardvark

Read setup first if the backend or the comfyui plugin state is unknown. Video needs a 16 GB-class NVIDIA card. Clips take minutes, so every route is queued.

Pick the model

GET ${GUAARDVARK_URL:-http://localhost:5000}/api/batch-video/models lists the registry with is_downloaded / is_ready (and missing_files), capabilities (modes t2v/i2v, max_frames, native_fps, aspect_ratios, min_steps, speed_profiles, audio_out), vram_mb, size_gb, license. Installed on a typical box:

idwhat it isnotes
wan22-5bWan 2.2 TI2V-5B, t2v + i2v, 24 fps, up to 121 framesthe everyday default
wan22-14b / wan22-14b-i2vWan 2.2 14B MoE, 16 fps, 81 framesbest quality; speed_profile: lightx2v-4 for 4-step Lightning
ltx23-distilled-fp8LTX-2.3 distilled, 8 steps, 161 framesfastest long clips
minimax-h3-int8MiniMax H3, text / first / last / first+last frame, generates its own stereo soundtrack and spoken linespass audio: true; ~6.5 min for 5 s on a 16 GB card
cogvideox-5b / -i2vCogVideoX, 8 fps, 49 frameslegacy

The active default is GET /api/settings/active_video_model (resolved.t2v, resolved.i2v, resolved.scene). Never pass a step count below the model's min_steps; the server raises it and the result would be smeared anyway.

One clip: MCP generate_video

  • prompt: scene, subject, motion, style, and any spoken lines (H3 speaks them).
  • model, duration_s (clamped to the model), aspect_ratio (one the model declares), style (cinematic, realistic, anime, 3d_animation, ...), num_inference_steps, speed_profile.
  • audio: true forces a soundtrack-capable model (H3) and fails on a silent family.
  • first_image / last_image: document id or path; last frame needs a first+last mode model.
  • reference_images / reference_audio: lock a person, look or voice (reference build only).
  • wait_for_result default false: the tool returns a batch id and a Studio deep link at once. Give the user the link; poll get_generation_status(batch_id=...) (MCP) or GET /api/batch-video/status/<batch_id> if they ask you to wait. wait_for_result: true blocks for the clip, up to 30 minutes.
Show full SKILL.md (161 more words)Show less

Looping animation: MCP generate_animation

Frame-morph GIF/MP4 via img2img: prompt, motion, frames 2-24, strength 0.1-0.5, format gif|mp4|both. Use for short loops and stickers, not for cinema clips. Over MCP it answers at once with a job id (tooljob_...); poll get_generation_status for the GIF and MP4 URLs, or pass wait_for_result: true to wait up to 60 s for them.

Batch: REST

bash
B=${GUAARDVARK_URL:-http://localhost:5000}
# text to video, one clip per prompt
curl -s -X POST $B/api/batch-video/generate/text -H 'Content-Type: application/json' -d '{
  "prompts": ["a red kite over a grey sea", "the same kite at dusk"],
  "model": "wan22-5b", "prompt_style": "cinematic", "enhance_prompt": true, "seed": 42
}'
# image to video
curl -s -X POST $B/api/batch-video/generate/image -H 'Content-Type: application/json' -d '{
  "image_paths": ["/abs/path/frame.png"], "prompt": "slow push in, wind in the grass", "model": "wan22-5b"
}'

Optional keys the server honours: negative_prompt, guidance_scale, motion_strength, interpolation_multiplier (RIFE frame interpolation), combine_frames, lora_name + lora_strength, adapters, guides (per-prompt audio/image anchors on models that declare them), last_frame_paths (image route), storyboard_concept / storyboard_shots. Poll GET $B/api/batch-video/status/<batch_id>; files via GET $B/api/batch-video/video/<batch_id>/<video_name>; cancel POST $B/api/batch-video/batch/<batch_id>/cancel; retry POST $B/api/batch-video/retry/<batch_id>.

Rules

  • Quote the model and the queued batch id back to the user. Never claim a clip is done until the status route says so.
  • The GPU is exclusive: while video renders, chat models are evicted. Warn before queuing a long batch.
  • The Studio page (Video Gen) shows the same queue; the user may prefer to watch it there.

© guaardvark, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .agents/skills/video of guaardvark/guaardvark.

Open the folder on GitHubat commit 2a33110

Compare with similar skills

Video next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Video compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Video this skillguaardvark/guaardvark257—~1.1kAutomated safety check: PassMIT
ComfyUI Local DriverSlavaSexton/ComfyUI-Agent-Kit105—~12kAutomated safety check: PassApache-2.0
VRGDG H3 Short Film Pipelinevrgamegirl19/comfyui-vrgamedevgirl763—~4.2kAutomated safety check: PassCustom licence
H3 Videoagent-next/video-agent120—~2.9kAutomated safety check: PassApache-2.0
BlockrunBlockRunAI/blockrun-mcp391—~2.7kAutomated safety check: PassMIT
Open Videoagent-next/video-agent120—~3.2kAutomated safety check: PassApache-2.0

Similar skills

  • ComfyUI Local Driver

    SlavaSexton/ComfyUI-Agent-Kit

    Drives a local ComfyUI install over its HTTP API to generate and edit images, video and audio, with per-model prompt recipes and workflow guidance.

    105 GitHub stars~12k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • VRGDG H3 Short Film Pipeline

    vrgamegirl19/comfyui-vrgamedevgirl

    Builds an AI short film in a local ComfyUI with the VRGDG Video Builder, MiniMax H3 scenes, reference images, a music score, QA and a final edit.

    763 GitHub stars~4.2k tokensUpdated today
    Media & CreativeAuto-check passed
  • H3 Video

    agent-next/video-agent

    OpenVideo skill (v0.1.0): generate high-quality local video with the OpenVideo product (MiniMax H3 backend).

    120 GitHub stars~2.9k tokensUpdated 4 days ago
    Media & CreativeAuto-check passed
  • Blockrun

    BlockRunAI/blockrun-mcp

    Pay-per-call access to AI models, real-time data, media generation and multi-chain RPC over x402 micropayments (USDC on Base or Solana), or a BlockRun account API key.

    391 GitHub stars~2.7k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Open Video

    agent-next/video-agent

    Generate, edit, or direct videos via open-source models (MiniMax H3 baseline; Wan2.2 / LTX future).

    120 GitHub stars~3.2k tokensUpdated 4 days ago
    Media & CreativeAuto-check passed
  • Miora Video Studio

    kangarooking/director-skills

    miora 视频生成的可靠性与规格层——补内置技能未覆盖的部分:四必填参数闸门(模式/分辨率/时长/画幅缺一不可提交)、时长区间闸门、完成信号的判定与轮询等待、官方规格口径与本通道实测差异(时长/分辨率/参考图上限)、文生/首尾帧/参考生三种模式的参数事实、提示词的自动修正边界、生成状态向用户的同步。当用户要生成视频、动效、短视频,或提到参考生视频/首尾帧/图生视频、miora…

    175 GitHub stars~4.6k tokensUpdated 2 days ago
    Media & CreativeAuto-check passed

More from guaardvark/guaardvark

All 15 skills in this repo
  • Cast

    guaardvark/guaardvark

    Build consistent characters, environments and props in Guaardvark's Cast Library and train LoRAs for them locally (reference photos → vision bible → sample plan → approved samples → training).

    257 GitHub stars~673 tokensUpdated today
    Auto-check passed
  • Image

    guaardvark/guaardvark

    Generate or edit images on the user's own GPU through Guaardvark: single images, instruction edits, background cut-outs, inpaint and outpaint, consistent characters from the Cast Library, and batch…

    257 GitHub stars~1.8k tokensUpdated today
    Auto-check passed
  • Music

    guaardvark/guaardvark

    Generate full songs with vocals or instrumentals (ACE-Step) and sound effects or ambience (Stable Audio Open) on the user's GPU through Guaardvark's Audio Foundry.

    257 GitHub stars~710 tokensUpdated today
    Auto-check passed
  • Ops

    guaardvark/guaardvark

    Operate a running Guaardvark: GPU and VRAM state, plugin start/stop, logs, Celery tasks, the Interconnector sync to other machines, overnight RAG autoresearch, and infographics.

    257 GitHub stars~779 tokensUpdated today
    Auto-check passed
  • Setup

    guaardvark/guaardvark

    Connect this agent to a running Guaardvark (self-hosted AI studio) and check what it can do right now.

    257 GitHub stars~1.2k tokensUpdated today
    Auto-check passed
  • Swarm

    guaardvark/guaardvark

    Launch and watch Guaardvark's Swarm Orchestrator: parallel coding agents, each in its own git worktree, working a markdown plan and merging back deterministically.

    257 GitHub stars~746 tokensUpdated today
    Auto-check passed

Questions about Video

What does Video do?

Generate video clips on the user's own GPU through Guaardvark: text-to-video, image-to-video, first+last frame animation, clips with their own soundtrack and dialogue (MiniMax H3), short looping…. Video is an agent skill from guaardvark/guaardvark. Generate video clips on the user's own GPU through Guaardvark: text-to-video, image-to-video, first+last frame animation, clips with their own soundtrack and dialogue (MiniMax H3), short looping animations, and batch runs.

When should I use Video?

Video fits situations like: the user asks to make; batch-generate video locally.

How do I install Video in Claude Code?

Run `npx skills add guaardvark/guaardvark --skill video -a claude-code`. Or copy the skill folder (.agents/skills/video in guaardvark/guaardvark) into .claude/skills/video in your project. Claude Code loads it when a task matches its description.

How do I install Video in Codex?

Run `npx skills add guaardvark/guaardvark --skill video -a codex`. Or copy the skill folder (.agents/skills/video in guaardvark/guaardvark) into .agents/skills/video in your project. Codex loads it when a task matches its description.

Can I use Video in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add guaardvark/guaardvark --skill video -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/video, .gemini/skills/video, .github/skills/video and .opencode/skills/video in your project.

What does Video need to run?

Going by SKILL.md and its folder, Video needs the command-line tools its instructions call (curl).

Does Video access the network?

SKILL.md contains no URLs. Its commands use curl, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Video safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Video use?

Video is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Video use?

About 1.1k tokens (SKILL.md is roughly 4.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Video?

Skills that share tags, products or a category with Video: ComfyUI Local Driver (SlavaSexton/ComfyUI-Agent-Kit, 105 stars), VRGDG H3 Short Film Pipeline (vrgamegirl19/comfyui-vrgamedevgirl, 763 stars), H3 Video (agent-next/video-agent, 120 stars) and Blockrun (BlockRunAI/blockrun-mcp, 391 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Video?

guaardvark (a GitHub user) maintains it in guaardvark/guaardvark, which has 257 GitHub stars. The repository holds 15 skills in this directory. The repository was last updated on October 9, 2026.

Source: guaardvark/guaardvark on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.