Agent skill

Volcengine Video Understanding

by freestylefly in freestylefly/canghe-skills

火山视频理解 - 使用火山方舟视频理解 API 分析视频内容。通过 Files API 上传视频(推荐),支持大文件(最大512MB),可用于视频内容分析、物体识别、动作理解等。当用户需要分析视频、理解视频内容、提取视频信息时激活此技能。

No licenceAuto-check: notesAI & LLM Engineering

Install Volcengine Video Understanding

skills CLI
$ npx skills add freestylefly/canghe-skills --skill volcengine-video-understanding -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install freestylefly/canghe-skills volcengine-video-understanding --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/freestylefly/canghe-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/canghe-volcengine-video-understanding .claude/skills/volcengine-video-understanding && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
volcengine-video-understanding
GitHub stars
461
Token cost
~1k tokens
SKILL.md length
179 words
Files
2 (incl. scripts)
Skills in repo
19
Repo updated
First seen
Licence
None found

At a glance

火山视频理解 - 使用火山方舟视频理解 API 分析视频内容。通过 Files API 上传视频(推荐),支持大文件(最大512MB),可用于视频内容分析、物体识别、动作理解等。当用户需要分析视频、理解视频内容、提取视频信息时激活此技能。

  • Works in 5 steps: 基础视频分析(Files API 方式 - 推荐) → 视频问答 → 情感分析 → …
  • Tasks that involve Computer vision
  • SKILL.md covers 功能, 前置要求, 使用方法 and 参数说明, plus 7 more sections
  • Runs Python scripts from its folder; calls python3 and curl; reaches ark.cn-beijing.volces.com; needs ARK_API_KEY

What it does

Volcengine Video Understanding is an agent skill from freestylefly/canghe-skills. 火山视频理解 - 使用火山方舟视频理解 API 分析视频内容。通过 Files API 上传视频(推荐),支持大文件(最大512MB),可用于视频内容分析、物体识别、动作理解等。当用户需要分析视频、理解视频内容、提取视频信息时激活此技能。

Its SKILL.md is about 1k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/video_understand.py`).

It sits in AI & LLM Engineering, covering Computer vision. The repository describes itself as: 苍何的技能skills仓库,搜集好用的 skills,辅助提效.

When your agent uses it

  • Tasks that involve Computer vision

Example prompts

  • “/volcengine-video-understanding”

Requirements

  • Python 3
  • A credential in ARK_API_KEY

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. 基础视频分析(Files API 方式 - 推荐)
  2. 视频问答
  3. 情感分析
  4. 指定模型和帧率
  5. 保存结果到文件

What it can do on your machine

Read from SKILL.md and the folder at commit dd0bf35. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python3
    • curl

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • ark.cn-beijing.volces.com

    Also links to:

    • volcengine.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • ARK_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Volcengine Video Understanding loads about 1k tokens when it runs. Until then it costs about 38 tokens; SKILL.md has 179 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~38
When it runs · the whole SKILL.md, loaded when a task matches
~1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NoteMentions a .env fileSKILL.md:30
    anghe-skills/.env.example .canghe-skills/.env
  • NoteMentions a .env fileSKILL.md:33
    2. 编辑 `.canghe-skills/.env` 文件,填写你的 API Key:
  • NoteMentions a .env fileSKILL.md:47
    2. 当前目录 `.canghe-skills/.env`
  • NoteMentions a .env fileSKILL.md:48
    3. 用户主目录 `~/.canghe-skills/.env`

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 179 words (~1,035 tokens).

name
volcengine-video-understanding

Read the full SKILL.md on GitHub

Files

SKILL.md and 1 other file (scripts) in skills/canghe-volcengine-video-understanding of freestylefly/canghe-skills.

  • SKILL.md
  • scripts/video_understand.py

Open the folder on GitHubat commit dd0bf35

Compare with similar skills

Volcengine Video Understanding next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Volcengine Video Understanding compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Volcengine Video Understanding this skillfreestylefly/canghe-skills461—~1kAutomated safety check: NotesNone
Segment Anything Model GuideOrchestra-Research/AI-Research-SKILLs13k8 repos~3.3kAutomated safety check: PassMIT
CLIP Image-Text MatchingOrchestra-Research/AI-Research-SKILLs13k7 repos~1.7kAutomated safety check: PassMIT
Yolo Master AgentTencent/YOLO-Master745—~755Automated safety check: PassAGPL-3.0
Video Understandjjyaoao/HelloAgents3.2k1 repos~6.2kAutomated safety check: PassMIT
LLaVA Vision-Language ModelOrchestra-Research/AI-Research-SKILLs13k6 repos~2kAutomated safety check: PassMIT

Similar skills

  • Segment Anything Model Guide

    Orchestra-Research/AI-Research-SKILLs

    Guide to using Meta's Segment Anything Model for zero-shot image segmentation with point, box or mask prompts, or automatic mask generation.

    13k GitHub starsUsed in 8 repos~3.3k tokens
    AI & LLM EngineeringAuto-check passed
  • CLIP Image-Text Matching

    Orchestra-Research/AI-Research-SKILLs

    Explains OpenAI's CLIP model for zero-shot image classification, image-text similarity, semantic image search and content moderation, with install steps and code patterns.

    13k GitHub starsUsed in 7 repos~1.7k tokens
    AI & LLM EngineeringAuto-check passed
  • Yolo Master Agent

    Tencent/YOLO-Master

    A skill your agent uses when the user wants to run a YOLO-Master task (train/val/predict/track/export/benchmark) or use the Agent Skill dispatcher.

    745 GitHub stars~755 tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Video Understand

    jjyaoao/HelloAgents

    Implement specialized video understanding capabilities using the z-ai-web-dev-sdk.

    3.2k GitHub starsUsed in 1 repo~6.2k tokens
    AI & LLM EngineeringAuto-check passed
  • LLaVA Vision-Language Model

    Orchestra-Research/AI-Research-SKILLs

    Guide to LLaVA for image chat, visual question answering and captioning, with model sizes, CLI and Gradio usage and multi-turn conversation code.

    13k GitHub starsUsed in 6 repos~2k tokens
    AI & LLM EngineeringAuto-check passed
  • Motioneyes Visual Analysis

    edwardsanchez/MotionEyes

    Pixel-based motion and UI change analysis from frame sequences or screenshots using computer vision and visual comparison.

    229 GitHub stars~2k tokensUpdated 6 mo ago
    AI & LLM EngineeringAuto-check passed

More from freestylefly/canghe-skills

All 19 skills in this repo
  • Canghe Comic

    freestylefly/canghe-skills

    Knowledge comic creator supporting multiple art styles and tones.

    461 GitHub starsUsed in 8 repos~3.2k tokens
    Auto-check passed
  • Canghe Post To Wechat

    freestylefly/canghe-skills

    Posts content to WeChat Official Account (微信公众号) via API or Chrome CDP.

    461 GitHub starsUsed in 3 repos~3.4k tokens
    Auto-check: notes
  • Canghe Post To X

    freestylefly/canghe-skills

    Posts content and articles to X (Twitter). An agent skill from freestylefly/canghe-skills.

    461 GitHub starsUsed in 4 repos~1.7k tokens
    Auto-check: warnings
  • Canghe URL To Markdown

    freestylefly/canghe-skills

    Fetch any URL and convert to markdown using Chrome CDP. An agent skill from freestylefly/canghe-skills.

    461 GitHub starsUsed in 4 repos~1.1k tokens
    Auto-check passed
  • Canghe Xhs Images

    freestylefly/canghe-skills

    Generates Xiaohongshu (Little Red Book) infographic series with 10 visual styles and 8 layouts.

    461 GitHub starsUsed in 4 repos~4.9k tokens
    Auto-check passed
  • Canghe Markdown To HTML

    freestylefly/canghe-skills

    Converts Markdown to styled HTML with WeChat-compatible themes.

    461 GitHub starsUsed in 2 repos~1.6k tokens
    Auto-check passed

Questions about Volcengine Video Understanding

What does Volcengine Video Understanding do?

火山视频理解 - 使用火山方舟视频理解 API 分析视频内容。通过 Files API 上传视频(推荐),支持大文件(最大512MB),可用于视频内容分析、物体识别、动作理解等。当用户需要分析视频、理解视频内容、提取视频信息时激活此技能。. Volcengine Video Understanding is an agent skill from freestylefly/canghe-skills.

When should I use Volcengine Video Understanding?

Volcengine Video Understanding fits situations like: tasks that involve Computer vision.

How do I install Volcengine Video Understanding in Claude Code?

Run `npx skills add freestylefly/canghe-skills --skill volcengine-video-understanding -a claude-code`. Or copy the skill folder (skills/canghe-volcengine-video-understanding in freestylefly/canghe-skills) into .claude/skills/volcengine-video-understanding in your project. Claude Code loads it when a task matches its description.

How do I install Volcengine Video Understanding in Codex?

Run `npx skills add freestylefly/canghe-skills --skill volcengine-video-understanding -a codex`. Or copy the skill folder (skills/canghe-volcengine-video-understanding in freestylefly/canghe-skills) into .agents/skills/volcengine-video-understanding in your project. Codex loads it when a task matches its description.

Can I use Volcengine Video Understanding in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add freestylefly/canghe-skills --skill volcengine-video-understanding -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/volcengine-video-understanding, .gemini/skills/volcengine-video-understanding, .github/skills/volcengine-video-understanding and .opencode/skills/volcengine-video-understanding in your project.

What does Volcengine Video Understanding need to run?

Going by SKILL.md and its folder, Volcengine Video Understanding needs Python for the scripts in its folder, the command-line tools its instructions call (python3 and curl) and credentials named ARK_API_KEY. Our summary lists: Python 3; A credential in ARK_API_KEY.

Does Volcengine Video Understanding access the network?

SKILL.md names 2 domains. In commands or code: ark.cn-beijing.volces.com; the agent is likely to contact it when it follows the instructions. As links in the text: volcengine.com. This is read from the text; nothing was executed.

Is Volcengine Video Understanding safe to install?

Our automated static check of SKILL.md found notes only (mentions a .env file), nothing it rates as a warning. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Volcengine Video Understanding use?

No licence was found for Volcengine Video Understanding or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Volcengine Video Understanding use?

About 1k tokens (SKILL.md is roughly 4.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Volcengine Video Understanding?

Skills that share tags, products or a category with Volcengine Video Understanding: Segment Anything Model Guide (Orchestra-Research/AI-Research-SKILLs, 13k stars), CLIP Image-Text Matching (Orchestra-Research/AI-Research-SKILLs, 13k stars), Yolo Master Agent (Tencent/YOLO-Master, 745 stars) and Video Understand (jjyaoao/HelloAgents, 3.2k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Volcengine Video Understanding?

freestylefly (a GitHub user) maintains it in freestylefly/canghe-skills, which has 461 GitHub stars. The repository holds 19 skills in this directory. The repository was last updated on June 8, 2026.

Source: freestylefly/canghe-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.