Agent skill

Comfyui

by calesthio in calesthio/OpenMontage

A skill your agent uses when working with ComfyUI workflows in OpenMontage, including comfyuiimage/comfyuivideo/comfyuimusic, custom workflowjson/workflowpath inputs, outputnode selection, missing…

AGPL-3.0Auto-check passedAI & LLM Engineering

Install Comfyui

skills CLI
$ npx skills add calesthio/OpenMontage --skill comfyui -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install calesthio/OpenMontage comfyui --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/calesthio/OpenMontage.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/comfyui .claude/skills/comfyui && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
comfyui
GitHub stars
66k
Token cost
~2k tokens
SKILL.md length
1,043 words
Files
1
Skills in repo
41
Repo updated
First seen
Licence
AGPL-3.0

At a glance

A skill your agent uses when working with ComfyUI workflows in OpenMontage, including comfyuiimage/comfyuivideo/comfyuimusic, custom workflowjson/workflowpath inputs, outputnode selection, missing…

  • Working with ComfyUI workflows in OpenMontage
  • SKILL.md covers Server Contract, Choosing a Workflow, Output Node Contract and Templated vs Fixed Nodes, plus 4 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • Including comfyuiimage/comfyuivideo/comfyuimusic

What it does

Comfyui is an agent skill from calesthio/OpenMontage. Use when working with ComfyUI workflows in OpenMontage, including comfyuiimage/comfyuivideo/comfyuimusic, custom workflowjson/workflowpath inputs, outputnode selection, missing model setup, LoRAs, low-VRAM workflow choices, and community workflow imports.

Its SKILL.md is about 2k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in AI & LLM Engineering, covering Diffusion and image models and Fine-tuning. It works with ComfyUI and MiniMax. The repository describes itself as: World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent skill and production-knowledge files. Turn your AI coding assistant… The licence is AGPL-3.0.

When your agent uses it

  • Working with ComfyUI workflows in OpenMontage
  • Including comfyuiimage/comfyuivideo/comfyuimusic
  • Custom workflowjson/workflowpath inputs
  • Outputnode selection

Example prompts

  • “/comfyui”

What it can do on your machine

Read from SKILL.md and the folder at commit 9327439. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Comfyui loads about 2k tokens when it runs. Until then it costs about 67 tokens; SKILL.md has 1,043 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~67
When it runs · the whole SKILL.md, loaded when a task matches
~2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from calesthio/OpenMontage at commit 9327439, republished under its AGPL-3.0 licence (© calesthio). 1,043 words, ~2,022 tokens.

Download SKILL.mdSave it as .claude/skills/comfyui/SKILL.md (or your agent's skills folder).
name
comfyui
description
Use when working with ComfyUI workflows in OpenMontage, including comfyui_image/comfyui_video/comfyui_music, custom workflow_json/workflow_path inputs, output_node selection, missing model setup, LoRAs, low-VRAM workflow choices, and community workflow imports.

ComfyUI Workflows in OpenMontage

Use this skill before calling comfyui_image, comfyui_video, or comfyui_music, and when converting a community ComfyUI workflow into an OpenMontage tool call.

Server Contract

  • ComfyUI must be running before the tool can generate. The default server is http://localhost:8188; override it with COMFYUI_SERVER_URL.
  • Running separate ComfyUI instances per capability (different GPU, different model set)? COMFYUI_IMAGE_SERVER_URL / COMFYUI_VIDEO_SERVER_URL / COMFYUI_MUSIC_SERVER_URL each override COMFYUI_SERVER_URL for that one tool only. Optional -- a single-server setup needs none of these.
  • Health and hardware status come from GET /system_stats.
  • Jobs are submitted to POST /prompt, completed outputs are read from GET /history/{prompt_id}, and artifact bytes are downloaded with GET /view.
  • Long waits (video, music) prefer ComfyUI's websocket feed for immediate completion/error detection and transparently fall back to REST polling if websocket-client isn't installed. Either way, a timeout is recoverable: pass the error's prompt_id back in as resume_prompt_id to resume waiting on the same job instead of resubmitting it.
  • Export workflows with ComfyUI's API-format JSON, not the UI layout format. If a downloaded workflow will not submit, re-export it from ComfyUI with API format enabled.
Partner Nodes are hosted
  • gemini_omni_flash, seedance_2.5, and minimax_h3_api in comfyui_video are official ComfyUI Partner Nodes. They call hosted APIs and require network access, a logged-in Comfy account, and prepaid credits.
  • Do not describe Partner Nodes as local, offline, or free merely because the graph runs in a local ComfyUI process.
  • minimax_h3_local is a separate open-weight path. It requires the official MiniMax H3 workflow exported in API format, its output_node, and the model stack reported by the tool.

Choosing a Workflow

  • Use bundled workflows when the requested operation matches and the local machine has the required models and VRAM.
  • Use a custom workflow_json or workflow_path when the user needs a community recipe, a lower-VRAM model, a different style family, or custom nodes.
  • For 8GB-12GB GPUs, prefer lower-footprint workflows such as Wan 2.1 1.3B, LTXV FP8 or quantized workflows, or Wan 2.2 GGUF/quantized community workflows. The bundled Wan 2.2 14B FP8 video workflows are a 16GB-class path, not a provider-wide floor.
  • Do not promise that arbitrary custom workflows will fit a machine. The workflow, quantization, resolution, frame count, and offload settings determine the real resource envelope.

Output Node Contract

  • Custom workflows must pass output_node.
  • Pick the node that writes the artifact, usually SaveImage, SaveVideo, VHS_VideoCombine, or another terminal saver node.
  • Pass the node ID as a string, for example "108". Do not pass the class name.
  • If a workflow has multiple savers, choose the final deliverable node, not previews or intermediates.

Templated vs Fixed Nodes

  • Identify templated nodes before execution: prompt text, seed, dimensions, frame count, source image, sampler settings, and output filename prefix.
  • Fixed nodes are model loaders, VAEs, text encoders, LoRA loaders, schedulers, and graph wiring. Do not mutate those unless the workflow author intended that customization.
  • For community workflows, inspect each loader node and note every required model or custom node before running. Missing models should be handled through the tool's structured missing_models payload when available.
Video workflows need a temporal latent
  • A video workflow's empty-latent node must be a video latent -- EmptyHunyuanLatentVideo for Wan 2.2 t2v, Wan22ImageToVideoLatent or WanImageToVideo for the image-conditioned variants. The frame count goes in length; batch_size is how many separate clips to generate and stays at 1.
  • EmptyLatentImage with batch_size set to the frame count is a trap: it asks for N unrelated images, and the resulting MP4 has the right frame count, duration and codec and passes ffprobe cleanly -- it just strobes. Check the latent node before running any unfamiliar video workflow.
Show full SKILL.md (456 more words)Show less

Model and LoRA Setup

  • Use ComfyUI Manager or the workflow author's model links when available, and respect model licenses.
  • Place models in the folders expected by the loader nodes: diffusion models under ComfyUI/models/diffusion_models/, text encoders under ComfyUI/models/text_encoders/, VAEs under ComfyUI/models/vae/, and LoRAs under ComfyUI/models/loras/.
  • For LoRA stacks, use LoraLoader or LoraLoaderModelOnly chains in the workflow. Record each LoRA name plus strength_model and strength_clip when applicable.
  • The current ComfyUI tools do not inject LoRAs into arbitrary graphs. To use LoRAs, provide a workflow that already contains the LoRA loader chain and pass model-stack provenance.

Provenance

  • For custom workflows, provide workflow_name and workflow_model when known.
  • Provide workflow_model_stack for reproducibility when the workflow is not bundled. Include base checkpoint or diffusion model, quantization, text encoder, VAE, LoRAs and strengths, sampler or scheduler, steps, and guidance if the workflow exposes them.
  • The tools record the final workflow hash. Treat that hash plus the model stack, seed, dimensions, and prompt as the reproducibility contract.

Failure Handling

  • If the server is unavailable, surface the structured setup offer. Starting ComfyUI or setting COMFYUI_SERVER_URL is the first fix.
  • If models are missing, read data.missing_models[]; each item should include the file name, role, destination hint, and download URL when OpenMontage knows it.
  • If custom nodes are missing, ask the user to install them through ComfyUI Manager or the workflow author's documented install path, then restart ComfyUI.
  • If a long render times out locally, check ComfyUI history before retrying from scratch; the server may still have completed the prompt -- or just call again with resume_prompt_id set to the prompt_id from the timeout error.

Music (comfyui_music)

  • Bundled default is ACE-Step v1 (3.5B) text-to-audio, built from ComfyUI's native TextEncodeAceStepAudio/EmptyAceStepLatentAudio nodes (core, not a third-party pack) -- unlike ACE-Step 1.5 or other custom node packs, v1's interface is standardized enough to bundle safely.
  • prompt maps to the bundled workflow's tags field (style/genre/mood, e.g. "upbeat electronic pop, female vocals"), matching the same "prompt = music description" convention suno_music uses. lyrics is a separate optional field -- leave empty for instrumental, or use [verse]/[chorus]/[bridge] structure tags and [zh]/[ja]/[ko]-style language-code prefixes for non-English lines.
  • duration_seconds, steps, cfg, lyrics_strength, and seed are patchable on the bundled workflow. Missing ace_step_v1_3.5b.safetensors surfaces through the same data.missing_models[] contract as image/video.
  • Need ACE-Step 1.5, a different node pack, or a non-ACE-Step audio model? Fall back to workflow_json/workflow_path + output_node, exactly like a custom image/video workflow -- in that mode prompt becomes provenance/logging only again and must already be baked into the graph.
  • output_node (bundled or custom) should be the node that writes the final audio -- the bundled workflow's is SaveAudioMP3. The client reads artifacts from that node's "audio" output key (parallel to "images" for image/video savers).
  • For custom workflows, provide workflow_name/workflow_model/workflow_model_stack for provenance exactly as you would for a custom image/video workflow.

© calesthio, AGPL-3.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .agents/skills/comfyui of calesthio/OpenMontage.

Open the folder on GitHubat commit 9327439

Compare with similar skills

Comfyui next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Comfyui compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Comfyui this skillcalesthio/OpenMontage66k—~2kAutomated safety check: PassAGPL-3.0
Minimax H3 Videoartokun/comfyui-mcp800—~3.4kAutomated safety check: PassMIT
Continuity Renderroadmaus/ComfyUI-Continuity133—~1.7kAutomated safety check: PassMIT
AI Toolkit Trainerartokun/comfyui-mcp800—~2.7kAutomated safety check: PassMIT
Civitaiartokun/comfyui-mcp800—~1.1kAutomated safety check: PassMIT
H3 Prompt Directordagthomas/comfyui_dagthomas290—~1.9kAutomated safety check: PassMIT

Similar skills

  • Minimax H3 Video

    artokun/comfyui-mcp

    Build MiniMax H3 (Hailuo) local video workflows with native T2V/I2V/R2V nodes, Comfy-Org INT8 weights, turbo LoRAs for 8GB VRAM, 15-second stereo-audio clips, and the official MiniMax prompting…

    800 GitHub stars~3.4k tokensUpdated 4 days ago
    AI & LLM EngineeringAuto-check passed
  • Continuity Render

    roadmaus/ComfyUI-Continuity

    Render videos and pictures on a ComfyUI server that has the Continuity node pack (MiniMax H3, LTX 2.5, Krea 2, Ideogram 4, Qwen Image, Flux 2 Klein) with one command, the render.py bundled in this…

    133 GitHub stars~1.7k tokensUpdated yesterday
    AI & LLM EngineeringAuto-check passed
  • AI Toolkit Trainer

    artokun/comfyui-mcp

    Train custom LoRAs with ostris AI-Toolkit. An agent skill from artokun/comfyui-mcp.

    800 GitHub stars~2.7k tokensUpdated 4 days ago
    AI & LLM EngineeringAuto-check passed
  • Civitai

    artokun/comfyui-mcp

    Discover Civitai models with the BUILT-IN downloadmodel action:"searchcivitai" and install/generate them locally.

    800 GitHub stars~1.1k tokensUpdated 4 days ago
    AI & LLM EngineeringAuto-check passed
  • H3 Prompt Director

    dagthomas/comfyui_dagthomas

    Core craft for writing MiniMax H3 video prompts as the engine behind the APNext H3 nodes in ComfyUI - how to obey the node's directives (task type, duration, shot plan, camera, dialogue, wildness)…

    290 GitHub stars~1.9k tokensUpdated 1 mo ago
    AI & LLM EngineeringAuto-check passed
  • Setup

    guaardvark/guaardvark

    Connect this agent to a running Guaardvark (self-hosted AI studio) and check what it can do right now.

    257 GitHub stars~1.2k tokensUpdated today
    AI & LLM EngineeringAuto-check passed

More from calesthio/OpenMontage

All 41 skills in this repo
  • Video Understand

    calesthio/OpenMontage

    Understand video content locally using ffmpeg frame extraction and Whisper transcription.

    66k GitHub stars~841 tokensUpdated 6 days ago
    Auto-check passed
  • Avatar Video

    calesthio/OpenMontage

    Create AI avatar videos with precise control over avatars, voices, scripts, scenes, and backgrounds using HeyGen's v2 API.

    66k GitHub stars~1.6k tokensUpdated 6 days ago
    Auto-check passed
  • D3 Viz

    calesthio/OpenMontage

    Creating interactive data visualisations using d3.js. An agent skill from calesthio/OpenMontage.

    66k GitHub starsUsed in 3 repos~5.4k tokens
    Auto-check passed
  • Create Video

    calesthio/OpenMontage

    Create videos from a text prompt using HeyGen's Video Agent.

    66k GitHub stars~1.3k tokensUpdated 6 days ago
    Auto-check passed
  • Threejs World Generation

    calesthio/OpenMontage

    Build deterministic, editable, free-viewpoint Three.js worlds from text or structured briefs.

    66k GitHub stars~2k tokensUpdated 6 days ago
    Auto-check passed
  • Video Edit

    calesthio/OpenMontage

    Edit videos locally using ffmpeg. An agent skill from calesthio/OpenMontage.

    66k GitHub stars~855 tokensUpdated 6 days ago
    Auto-check: notes

Works with

Questions about Comfyui

What does Comfyui do?

A skill your agent uses when working with ComfyUI workflows in OpenMontage, including comfyuiimage/comfyuivideo/comfyuimusic, custom workflowjson/workflowpath inputs, outputnode selection, missing…. Comfyui is an agent skill from calesthio/OpenMontage. Use when working with ComfyUI workflows in OpenMontage, including comfyuiimage/comfyuivideo/comfyuimusic, custom workflowjson/workflowpath inputs, outputnode selection, missing model setup, LoRAs, low-VRAM workflow choices, and community workflow imports.

When should I use Comfyui?

Comfyui fits situations like: working with ComfyUI workflows in OpenMontage; including comfyuiimage/comfyuivideo/comfyuimusic; custom workflowjson/workflowpath inputs; outputnode selection.

How do I install Comfyui in Claude Code?

Run `npx skills add calesthio/OpenMontage --skill comfyui -a claude-code`. Or copy the skill folder (.agents/skills/comfyui in calesthio/OpenMontage) into .claude/skills/comfyui in your project. Claude Code loads it when a task matches its description.

How do I install Comfyui in Codex?

Run `npx skills add calesthio/OpenMontage --skill comfyui -a codex`. Or copy the skill folder (.agents/skills/comfyui in calesthio/OpenMontage) into .agents/skills/comfyui in your project. Codex loads it when a task matches its description.

Can I use Comfyui in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add calesthio/OpenMontage --skill comfyui -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/comfyui, .gemini/skills/comfyui, .github/skills/comfyui and .opencode/skills/comfyui in your project.

What does Comfyui need to run?

SKILL.md names no scripts, command-line tools or credentials: Comfyui is instructions for the agent only.

Does Comfyui access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Comfyui safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Comfyui use?

Comfyui is published under the AGPL-3.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Comfyui use?

About 2k tokens (SKILL.md is roughly 8.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Comfyui?

Skills that share tags, products or a category with Comfyui: Minimax H3 Video (artokun/comfyui-mcp, 800 stars), Continuity Render (roadmaus/ComfyUI-Continuity, 133 stars), AI Toolkit Trainer (artokun/comfyui-mcp, 800 stars) and Civitai (artokun/comfyui-mcp, 800 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Comfyui?

calesthio (a GitHub user) maintains it in calesthio/OpenMontage, which has 65,614 GitHub stars. The repository holds 41 skills in this directory. The repository was last updated on October 3, 2026.

Source: calesthio/OpenMontage on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.