Agent skill

Image To Video

by notque in notque/vexjoy-agent

FFmpeg-based video creation from image and audio. An agent skill from notque/vexjoy-agent.

MITAuto-check: notesMedia & Creative

Install Image To Video

skills CLI
$ npx skills add notque/vexjoy-agent --skill image-to-video -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install notque/vexjoy-agent image-to-video --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/notque/vexjoy-agent.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/content/image-to-video .claude/skills/image-to-video && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
image-to-video
GitHub stars
439
Token cost
~817 tokens
SKILL.md length
268 words
Files
7 (incl. scripts, references)
Skills in repo
61
Repo updated
First seen
Licence
MIT

At a glance

FFmpeg-based video creation from image and audio. An agent skill from notque/vexjoy-agent.

  • Works in 4 steps: VALIDATE → PREPARE → ENCODE → …
  • Tasks that involve AI video generation
  • SKILL.md covers Deep References, Phase 1: VALIDATE, Phase 2: PREPARE and Phase 3: ENCODE, plus 2 more sections
  • Runs Python scripts from its folder; calls python3, ffmpeg and ffprobe

What it does

Image To Video is an agent skill from notque/vexjoy-agent. FFmpeg-based video creation from image and audio.

Its SKILL.md is about 820 tokens, which your agent loads only when the skill is triggered. The skill folder holds 11 other files, including scripts and reference files (for example `references/ffmpeg-filters.md`, `scripts/image_to_video.py` and `workspace/README.md`).

It sits in Media & Creative, covering AI video generation and Video production. It works with FFmpeg. The repository describes itself as: VexJoy AI Agent with Jev Intelligent Routing - /do routes plain-English requests to the right specialist agent and gates the work with reviews, tests, and a learning loop. The licence is MIT.

When your agent uses it

  • Tasks that involve AI video generation
  • Tasks that involve Video production

Example prompts

  • “/image-to-video”

Requirements

  • Python 3
  • Pre-approved tools (allowed-tools): Read, Write, Bash, Grep, Glob, Edit

Workflow steps

4 steps, taken from the step headings in SKILL.md.

  1. VALIDATE
  2. PREPARE
  3. ENCODE
  4. VERIFY

What it can do on your machine

Read from SKILL.md and the folder at commit 5218674. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Read
    • Write
    • Bash
    • Grep
    • Glob
    • Edit

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python3
    • ffmpeg
    • ffprobe
    • apt
    • brew

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Image To Video loads about 817 tokens when it runs, and up to ~2.1k if it reads all its reference files. Until then it costs about 16 tokens; SKILL.md has 268 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~16
When it runs · the whole SKILL.md, loaded when a task matches
~817
With references · SKILL.md plus every file in references/, read only if the agent opens them
~2.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NotePre-approves every shell command (allowed-tools: Bash)SKILL.md
    allowed-tools: Read, Write, Bash, Grep, Glob, Edit

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from notque/vexjoy-agent at commit 5218674, republished under its MIT licence (© notque). 268 words, ~817 tokens.

Download SKILL.mdSave it as .claude/skills/image-to-video/SKILL.md (or your agent's skills folder). This skill also uses 6 other files; get the full folder from GitHub.
name
image-to-video
description
FFmpeg-based video creation from image and audio.
allowed-tools
Read, Write, Bash, Grep, Glob, Edit
promoted_to
video-editing
user-invocable
false
routing.triggers
image to video, audio visualization, static video, mp4 from image, music video, podcast video, video from image, combine image audio, album art video, cover…
routing.pairs_with
workflow
routing.complexity
simple
routing.category
video-creation

Image to Video Skill

Combine a static image with an audio file to produce an MP4 via FFmpeg. Supports resolution presets, audio visualization overlays, and batch processing. For image generation, use image-gen.

Deep References

SignalLoadWhy
FFmpeg filter graphs for visualization modesreferences/ffmpeg-filters.mdScale/pad, showwaves, showspectrum, overlay filters

Phase 1: VALIDATE

  1. Check FFmpeg: ffmpeg -version. If missing, stop with install instructions.
  2. Verify both input files exist with absolute paths and non-zero size. Supported: PNG/JPG/JPEG/GIF/WEBP/BMP (image), MP3/WAV/M4A/OGG/FLAC (audio).
  3. Determine parameters from the user's request -- do not default to static when the user requested a visualization.
PresetDimensionsPlatform
1080p1920x1080YouTube HD (default)
720p1280x720Standard HD
square1080x1080Instagram, social
vertical1080x1920Stories, Reels, TikTok

Visualization modes (off unless requested): waveform, spectrum, cqt, bars.

Gate: FFmpeg installed, both files exist, parameters resolved.

Phase 2: PREPARE

Use the user's output path or derive from audio filename (/same/dir/filename.mp4). Verify directory is writable.

Phase 3: ENCODE

Defaults: libx264 preset medium, CRF 23, yuv420p, 192k AAC.

bash
python3 $HOME/vexjoy-agent/skills/content/image-to-video/scripts/image_to_video.py \
  --image /path/to/image.png --audio /path/to/audio.mp3 \
  --output /path/to/output.mp4 --resolution 1080p --visualization static

Batch mode (matched pairs in workspace/input/):

bash
python3 $HOME/vexjoy-agent/skills/content/image-to-video/scripts/image_to_video.py \
  --process-workspace --visualization waveform

Gate: Script exits 0.

Phase 4: VERIFY

FFmpeg can exit 0 but produce a corrupt file. Always probe:

bash
ffprobe -v error -show_entries format=duration,size -show_entries stream=codec_name,width,height \
  -of default=noprint_wrappers=1 /path/to/output.mp4

Confirm video duration matches audio (within 1s). Report: file path, size, duration, resolution, visualization mode.

Error Handling

ErrorCauseSolution
FFmpeg not foundNot installedapt install ffmpeg or brew install ffmpeg
Image/audio not foundWrong or relative pathUse absolute paths; check with ls -la
FFmpeg filter errorsMinimal build lacks showwaves/showcqtInstall full FFmpeg; fall back to --visualization static
Cannot determine audio durationCorrupted audio fileTest with ffprobe; convert: ffmpeg -i input -acodec pcm_s16le output.wav

© notque, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 6 other files (scripts, references) in skills/content/image-to-video of notque/vexjoy-agent.

  • SKILL.md
  • references/ffmpeg-filters.md
  • scripts/image_to_video.py
  • workspace/.gitignore
  • workspace/README.md
  • workspace/completed/.gitkeep
  • workspace/output/.gitkeep

Open the folder on GitHubat commit 5218674

Compare with similar skills

Image To Video next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Image To Video compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Image To Video this skillnotque/vexjoy-agent439—~817Automated safety check: NotesMIT
HyperFrames Video Entry Pointheygen-com/hyperframes59k3 repos~5.2kAutomated safety check: PassApache-2.0
Video Shotseternityspring/reelbench-skills8721 repos~1.8kAutomated safety check: NotesApache-2.0
Video Scrubeternityspring/reelbench-skills872—~1.3kAutomated safety check: NotesApache-2.0
Stage EditOrkas-AI/Orkas-VideoStudio498—~2.4kAutomated safety check: PassMIT
Re Walkthrough Procharlesdove977/re-walkthrough-pro164—~1.1kAutomated safety check: NotesMIT

Similar skills

  • HyperFrames Video Entry Point

    heygen-com/hyperframes

    Entry point for making, editing and rendering videos from HTML compositions with HyperFrames, routing each request to the right workflow.

    59k GitHub starsUsed in 3 repos~5.2k tokens
    Media & CreativeAuto-check passed
  • Video Shots

    eternityspring/reelbench-skills

    拉片:把一条成片拆成逐镜头的分析表——每个镜头的时长、景别、类别、运镜、画面. An agent skill from eternityspring/reelbench-skills.

    872 GitHub starsUsed in 1 repo~1.8k tokens
    Media & CreativeAuto-check: notes
  • Video Scrub

    eternityspring/reelbench-skills

    把一条视频重建成「只有画面和声音」的干净文件——源片的元数据一概不搬: GPS、设备型号、账号 ID、创建时间、章节、GoPro 的遥测轨,全部留在原地。

    872 GitHub stars~1.3k tokensUpdated 18 days ago
    Media & CreativeAuto-check: notes
  • Stage Edit

    Orkas-AI/Orkas-VideoStudio

    Intelligent editing of real user-supplied footage—understand it with transcript/inspected-frame/scene/silence/quality evidence, then choose deterministic timeline operations or a constrained…

    498 GitHub stars~2.4k tokensUpdated 17 days ago
    Media & CreativeAuto-check passed
  • Re Walkthrough Pro

    charlesdove977/re-walkthrough-pro

    Turn a Zillow listing into a cinematic room-by-room walkthrough video (Apify scrape → Higgsfield image-to-video → ffmpeg stitch) to sell to real estate agents

    164 GitHub stars~1.1k tokensUpdated 2 mo ago
    Media & CreativeAuto-check: notes
  • Video Production

    ffroliva/gflow-cli

    A skill your agent uses when the user wants a finished video out of gflow rather than a single clip — a scripted scene, a talking-head or dialogue piece, an explainer, a product montage, a story…

    266 GitHub stars~7k tokensUpdated yesterday
    Media & CreativeAuto-check passed

More from notque/vexjoy-agent

All 61 skills in this repo
  • Game Asset Generator

    notque/vexjoy-agent

    Deterministic palette/matrix pixel art (not AI). An agent skill from notque/vexjoy-agent.

    439 GitHub stars~2.3k tokensUpdated 6 days ago
    Auto-check: notes
  • PR Workflow

    notque/vexjoy-agent

    Pull request lifecycle: commit, codex review, sync, review, fix, status, cleanup, and PR mining.

    439 GitHub stars~2.8k tokensUpdated 6 days ago
    Auto-check: notes
  • Architecture Deepening

    notque/vexjoy-agent

    Improve architecture across modules by deepening interfaces.

    439 GitHub stars~3.3k tokensUpdated 6 days ago
    Auto-check: notes
  • Code Quality

    notque/vexjoy-agent

    Code quality: cleanup, linting, formatting, quality gates. An agent skill from notque/vexjoy-agent.

    439 GitHub stars~1.5k tokensUpdated 6 days ago
    Auto-check: notes
  • Codebase Analyzer

    notque/vexjoy-agent

    Statistical rule discovery from Go codebase patterns. An agent skill from notque/vexjoy-agent.

    439 GitHub stars~2k tokensUpdated 6 days ago
    Auto-check: notes
  • Comment Quality

    notque/vexjoy-agent

    Review and fix temporal references in code comments. An agent skill from notque/vexjoy-agent.

    439 GitHub stars~2k tokensUpdated 6 days ago
    Auto-check: notes

Works with

Questions about Image To Video

What does Image To Video do?

FFmpeg-based video creation from image and audio. An agent skill from notque/vexjoy-agent. Image To Video is an agent skill from notque/vexjoy-agent. FFmpeg-based video creation from image and audio.

When should I use Image To Video?

Image To Video fits situations like: tasks that involve AI video generation; tasks that involve Video production.

How do I install Image To Video in Claude Code?

Run `npx skills add notque/vexjoy-agent --skill image-to-video -a claude-code`. Or copy the skill folder (skills/content/image-to-video in notque/vexjoy-agent) into .claude/skills/image-to-video in your project. Claude Code loads it when a task matches its description.

How do I install Image To Video in Codex?

Run `npx skills add notque/vexjoy-agent --skill image-to-video -a codex`. Or copy the skill folder (skills/content/image-to-video in notque/vexjoy-agent) into .agents/skills/image-to-video in your project. Codex loads it when a task matches its description.

Can I use Image To Video in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add notque/vexjoy-agent --skill image-to-video -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/image-to-video, .gemini/skills/image-to-video, .github/skills/image-to-video and .opencode/skills/image-to-video in your project.

What does Image To Video need to run?

Going by SKILL.md and its folder, Image To Video needs Python for the scripts in its folder and the command-line tools its instructions call (python3, ffmpeg, ffprobe, apt and brew). Our summary lists: Python 3. Its frontmatter pre-approves these tools: Read, Write, Bash, Grep, Glob, Edit.

Does Image To Video access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Image To Video safe to install?

Our automated static check of SKILL.md found notes only (pre-approves every shell command (allowed-tools: bash)), nothing it rates as a warning. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Image To Video use?

Image To Video is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Image To Video use?

About 817 tokens (SKILL.md is roughly 3.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 1.3k tokens, read only when the agent opens those files.

What are the alternatives to Image To Video?

Skills that share tags, products or a category with Image To Video: HyperFrames Video Entry Point (heygen-com/hyperframes, 59k stars), Video Shots (eternityspring/reelbench-skills, 872 stars), Video Scrub (eternityspring/reelbench-skills, 872 stars) and Stage Edit (Orkas-AI/Orkas-VideoStudio, 498 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Image To Video?

notque (a GitHub user) maintains it in notque/vexjoy-agent, which has 439 GitHub stars. The repository holds 61 skills in this directory. The repository was last updated on October 3, 2026.

Source: notque/vexjoy-agent on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.