Agent skill

Triton Sageattention

by artokun in artokun/comfyui-mcp

Install Triton + SageAttention to accelerate ComfyUI (the sageattn attentionmode and inductor torch.compile used by WanVideoWrapper / many video graphs).

MITAuto-check passedAI & LLM Engineering

Install Triton Sageattention

skills CLI
$ npx skills add artokun/comfyui-mcp --skill triton-sageattention -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install artokun/comfyui-mcp triton-sageattention --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/artokun/comfyui-mcp.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugin/skills/triton-sageattention .claude/skills/triton-sageattention && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
triton-sageattention
GitHub stars
803
Token cost
~5k tokens
SKILL.md length
2,174 words
Files
1
Skills in repo
42
Repo updated
First seen
Licence
MIT

At a glance

Install Triton + SageAttention to accelerate ComfyUI (the sageattn attentionmode and inductor torch.compile used by WanVideoWrapper / many video graphs).

  • Works in 5 steps: find the RIGHT python (NOT system python) → read the installed torch + CUDA + python → install triton-windows (matched to torch) → …
  • A loader crashes with No module named sageattention
  • SKILL.md covers Prefer kitchen INT8 attention…, Overview, Decide first: do you even need… and The safe sdpa / no-compile…, plus 7 more sections
  • Calls pip and python; reaches github.com

What it does

Triton Sageattention is an agent skill from artokun/comfyui-mcp. Install Triton + SageAttention to accelerate ComfyUI (the sageattn attentionmode and inductor torch.compile used by WanVideoWrapper / many video graphs). Windows-first (triton-windows + woct0rdho prebuilt SageAttention wheels matched to torch/CUDA/python into the RIGHT python), plus Linux (official triton + build) and Mac (N/A → sdpa/MPS). Also covers the SAFE sdpa / no-compile fallback so an example that assumes sageattn + torch.compile still runs when these aren't installed (video-extend TRAP 5). Use when a…

Its SKILL.md is about 5k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in AI & LLM Engineering, covering Diffusion and image models. It works with Python, ComfyUI, CUDA and Linux. The repository describes itself as: Local-first, agent-native control plane for ComfyUI — MCP server + sidebar agent that generates images, video & audio, authors and runs workflows, and edits your live graph in… The licence is MIT.

When your agent uses it

  • A loader crashes with No module named sageattention
  • Reports triton unavailable
  • Asked to speed up Wan/video workflows
  • Deciding whether to install acceleration vs

Example prompts

  • “t installed (video-extend TRAP 5). Use when a loader crashes with”
  • “sageattention”
  • “/triton-sageattention”

Requirements

  • Python 3

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. find the RIGHT python (NOT system python)
  2. read the installed torch + CUDA + python
  3. install triton-windows (matched to torch)
  4. install SageAttention (prebuilt wheel, matched to torch+CUDA)
  5. verify (Windows)

What it can do on your machine

Read from SKILL.md and the folder at commit 6ad6fc0. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • pip
    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • github.com

    Also links to:

    • download.pytorch.org

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Triton Sageattention loads about 5k tokens when it runs. Until then it costs about 182 tokens; SKILL.md has 2,174 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~182
When it runs · the whole SKILL.md, loaded when a task matches
~5k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from artokun/comfyui-mcp at commit 6ad6fc0, republished under its MIT licence (© artokun). 2,174 words, ~4,952 tokens.

Download SKILL.mdSave it as .claude/skills/triton-sageattention/SKILL.md (or your agent's skills folder).
name
triton-sageattention
description
Install Triton + SageAttention to accelerate ComfyUI (the sageattn attention_mode and inductor torch.compile used by WanVideoWrapper / many video graphs). Windows-first (triton-windows + woct0rdho prebuilt SageAttention wheels matched to torch/CUDA/python into the RIGHT python), plus Linux (official triton + build) and Mac (N/A → sdpa/MPS). Also covers the SAFE sdpa / no-compile fallback so an example that assumes sageattn + torch.compile still runs when these aren't installed (video-extend TRAP 5). Use when a loader crashes with "No module named 'sageattention'" or reports triton unavailable, when asked to speed up Wan/video workflows, or when deciding whether to install acceleration vs. fall back.
globs
**/*.json, **/packs/**

Triton + SageAttention (ComfyUI acceleration)

See also comfyui-launch-flags for the full attention / VRAM / cache flag matrix. Note the Z-Image exception: Z-Image is broken under --use-sage-attention, so launch it with --use-pytorch-cross-attention instead.

Prefer kitchen INT8 attention when it is available

If kitchen action:"status" (or panel_kitchen) reports kitchen present and int8_attention_is_available on this GPU, launch with --use-ck-attention and skip the sageattention wheel dance. Kitchen INT8 attention is a ComfyUI flag; it does not need a version-matched sageattention wheel. Restart required, consent-gated like every restart.

Only fall through to the Triton + SageAttention install below when kitchen INT8 is unknown or not available. A failed kitchen probe is unknown, not a no.

Overview

Two optional accelerators that many modern video graphs (especially kijai's ComfyUI-WanVideoWrapper) reference by default:

  • SageAttention (import sageattention), a quantized attention kernel. Selected via a node's attention_mode = sageattn (WanVideoWrapper) or ComfyUI's --use-sage-attention startup flag. ~20 to 40% faster sampling on supported NVIDIA GPUs.
  • Triton, the GPU kernel compiler that inductor torch.compile needs. WanVideoWrapper's WanVideoTorchCompileSettings (and any torch.compile/ inductor node) compiles the model through Triton for another speedup.

The risk. Both are version-locked to your exact torch + CUDA + python. A wrong wheel does worse than fail to install. It can break the torch install (mismatched CUDA DLLs, ImportError, or silent NaNs). And the failure mode of not having them is a hard crash before any sampling: ValueError: Can't import SageAttention: No module named 'sageattention', or compile errors / triton: unavailable in the startup log. This is exactly the video-extend TRAP 5.

Therefore the default is to get a working render FIRST with the sdpa / no-compile fallback, then OFFER to install acceleration for speed. Never run a torch-breaking install unannounced to "fix" a workflow. Fall back, render, then ask.

Verification note (June 2026). Wheel sources, the triton↔torch table, and the live attention_mode enum below were verified against woct0rdho/triton-windows, woct0rdho/SageAttention releases, and WanVideoWrapper's nodes (see Sources). Versions move fast, so always re-read the live torch/CUDA/python first (commands below) and pick the wheel that matches. Flag anything you can't confirm rather than guessing.


Decide first: do you even need them?

Workflow crashes "No module named 'sageattention'"  ──┐
  or "triton: unavailable" / torch.compile error     ─┤
                                                       ▼
              1. APPLY THE SDPA / NO-COMPILE FALLBACK  → render works now
                                                       ▼
              2. OFFER acceleration, in this order:
                 a. If kitchen INT8 attention is available:
                    "Want --use-ck-attention? No sageattention wheel."
                 b. Else:
                    "Want me to install Triton + SageAttention for ~20–40%
                     faster sampling? It's a version-matched install that
                     touches your torch env — I'll verify torch/CUDA/python
                     first and can roll back."
                                                       ▼
              3. Only on YES → install per-OS below → verify → re-enable
                 sageattn + torch.compile in the workflow.

Mac (no CUDA): skip the install entirely. The answer is always sdpa/MPS.


The safe sdpa / no-compile fallback (DO THIS FIRST)

When Triton/SageAttention aren't installed, make the workflow run unaccelerated but correct by switching attention to sdpa (PyTorch's built-in scaled dot-product attention, always available, no extra deps) and removing the torch.compile/inductor wiring.

WanVideoWrapper (the common case):

  1. On every WanVideoModelLoader set attention_mode to sdpa.
    • Confirmed enum values: sdpa, flash_attn_2, flash_attn_3, sageattn, sparse_sage_attention. The examples ship with sageattn; sdpa is the universal safe one.
  2. Disconnect WanVideoTorchCompileSettings from each loader's compile_args input (or delete/bypass the node). No compile = no Triton needed.
  3. (If present) bypass any WanVideoSetRadialAttention / sparse_sage_attention node. Those also route through SageAttention.

Generic ComfyUI: don't launch with --use-sage-attention; bypass any TorchCompileModel / inductor node.

This costs you speed, not quality. Use create_workflow (action:"modify") / the panel's strip-and-re-point flow to flip the widget and drop the link, then enqueue. Once it renders, offer the install.

Cross-ref: video-extend documents this exact fix as TRAP 5 for the Pusa extension graph (both WanVideoModelLoaders → attention_mode=sdpa, disconnect WanVideoTorchCompileSettings).


Windows install (the priority)

Windows has no official Triton or SageAttention build. You use community prebuilt wheels, and they must match torch + CUDA + python exactly. The panel agent has a shell (Bash for Claude / exec for Codex). Use it to run these in the correct python, never the system python.

Step 1 — find the RIGHT python (NOT system python)

ComfyUI on Windows comes in three flavors; each has its own python whose pip you must target:

VariantWhere its python livesHow to invoke pip
Desktop (standalone)a standalone-env\ (or venv) beside the install, e.g. C:\Users\<you>\ComfyUI-Installs\ComfyUI\standalone-env\python.exe"<install>\standalone-env\python.exe" -m pip ...
PortableComfyUI_windows_portable\python_embeded\python.exe"<...>\python_embeded\python.exe" -m pip ...
Manual venvthe venv you created (venv\Scripts\python.exe)activate it, then python -m pip ...

Detect it from the live server, the surest way to hit the same python ComfyUI runs on:

  • install_comfyui (action:"environment") / get_system_stats report embedded_python (true → Portable), the python version and the pytorch_version (e.g. 2.10.0+cu130).
  • Inspect the running process's argv (from get_system_stats). The path to main.py reveals the install root; its sibling standalone-env / python_embeded holds the python.
  • Last resort, ask the user for their ComfyUI folder.

Installing into the wrong python (e.g. a global pip install) is the #1 Windows mistake. The package lands somewhere ComfyUI never imports from, so the loader still crashes "No module named 'sageattention'". Always use "<that python>" -m pip.

Step 2 — read the installed torch + CUDA + python

Run with the python you found:

bash
"<python>" -c "import sys, torch; print(sys.version.split()[0], torch.__version__, torch.version.cuda)"

Example live output on this machine: 3.13.12 2.10.0+cu130 13.0, meaning python 3.13, torch 2.10, CUDA line cu130. You'll pick wheels for that triple.

Step 3 — install triton-windows (matched to torch)

Source: woct0rdho/triton-windows (the canonical Windows Triton fork; also on PyPI as triton-windows). The pin is an upper bound. pip resolves the right build for your torch:

bash
"<python>" -m pip install -U "triton-windows<3.7"

Why <3.7: each torch minor pins a Triton minor. Verified table:

PyTorchtriton-windowsconstraint to use
2.73.3"triton-windows<3.4"
2.83.4"triton-windows<3.5"
2.93.5"triton-windows<3.6"
2.103.6"triton-windows<3.7"

(torch 2.6 or older → triton 3.2 or earlier.) Pick the row for your torch.

  • CUDA toolkit: since triton-windows 3.2.0.post11 a minimal CUDA toolchain is bundled in the wheel, so you do NOT need a separate CUDA Toolkit install for Triton itself. (Triton 3.3 through 3.6 bundle the CUDA 12.8 line; works against cu12x/cu13x torch.)
  • MSVC / vcredist: Triton compiles C++ at runtime, so it needs the MSVC toolchain and "Visual C++ Redistributable 2015-2022" present. A TinyCC is bundled (since 3.2.0.post13) which covers many cases, but installing the Visual Studio Build Tools (C++ workload) plus the latest vcredist is the reliable fix if you hit compiler errors (see Traps).
  • Embedded/Portable python only: the embedded distro ships without C headers, so Triton can't compile. Download the matching python_<ver>_include_libs.zip from the triton-windows releases and copy its include and libs (note: libs, not lib) folders into python_embeded\. The Desktop standalone-env usually already has these.
Step 4 — install SageAttention (prebuilt wheel, matched to torch+CUDA)

Prefer the prebuilt wheel. Building from source needs the full CUDA Toolkit (nvcc) plus MSVC and often fails on Windows. Source: woct0rdho/SageAttention releases (Windows wheels; v2 = SageAttention 2.x).

Latest verified tag: v2.2.0-windows.post5, with these four wheels (all cp310-abi3, so they work on python 3.10 through 3.13+ via the stable ABI; one wheel covers all those pythons):

Wheel filenameFor
sageattention-2.2.0+cu128torch2.9.1.post5-cp310-abi3-win_amd64.whlCUDA 12.8 line, torch 2.9.x
sageattention-2.2.0+cu128torch2.10.0andhigher.post5-cp310-abi3-win_amd64.whlCUDA 12.8 line, torch ≥2.10
sageattention-2.2.0+cu130torch2.9.1.post5-cp310-abi3-win_amd64.whlCUDA 13.0 line, torch 2.9.x
sageattention-2.2.0+cu130torch2.10.0andhigher.post5-cp310-abi3-win_amd64.whlCUDA 13.0 line, torch ≥2.10

Pick by your CUDA line (cu128 vs cu130, from torch.version.cuda: 12.8 → cu128, 13.0 → cu130) and torch minor. For the live machine above (torch 2.10.0+cu130, py3.13) that is the last wheel. Install by full URL:

bash
"<python>" -m pip install "https://github.com/woct0rdho/SageAttention/releases/download/v2.2.0-windows.post5/sageattention-2.2.0+cu130torch2.10.0andhigher.post5-cp310-abi3-win_amd64.whl"
  • The cpXXX-abi3 tag means one wheel works across python ≥ its base (3.10+), so py3.13 is covered even though there's no cp313-specific wheel. This is expected, not a mismatch.
  • Always check the releases page for a newer tag than .post5 and newer torch variants. The filename pattern is stable (+cu<line>torch<minor>...abi3).
  • Don't build from source unless no wheel matches your torch/CUDA at all (then you need CUDA Toolkit plus MSVC; flag the cost to the user first).
Step 5 — verify (Windows)
bash
"<python>" -c "import triton; print('triton', triton.__version__)"
"<python>" -c "import sageattention; print('sageattention OK')"
"<python>" -c "import torch; print('torch still ok', torch.__version__, torch.cuda.is_available())"

All three must succeed and torch must still import with CUDA. If the third line now fails, the install clobbered torch (see Traps, roll back). Then restart ComfyUI and confirm the startup log no longer prints Could not load sageattention / triton: unavailable. Finally re-enable in the workflow: WanVideoModelLoader.attention_mode = sageattn and reconnect WanVideoTorchCompileSettings, enqueue, and confirm it samples (a torch.compile node will spend extra time on the first run compiling, which is normal).


Show full SKILL.md (929 more words)Show less

Linux install

Official builds exist here, so this is much simpler:

bash
# Triton: official, pip-installable; torch usually already pulls a matching triton.
pip install -U triton          # or let torch's pinned triton stand; match torch minor

# SageAttention: pip, or build from source for your GPU arch
pip install sageattention      # if a matching wheel exists for your torch/CUDA
  • Use the python that runs ComfyUI (its venv/conda env), the same rule as Windows.
  • Version matching still applies. torch pins a triton minor (e.g. torch 2.9.x ↔ triton 3.5.x, torch 2.10 ↔ 3.6); patch versions within a minor are interchangeable. Don't pip install triton blindly if it would upgrade past what your torch pins.
  • Build deps (if building SageAttention from source): the CUDA Toolkit with nvcc (matching your torch CUDA line), gcc/g++, and the torch headers. If CUDA is in a nonstandard path, export PATH=/usr/local/cuda-<ver>/bin:$PATH so the right nvcc is found. Building is GPU-arch specific and slow, so prefer a matching prebuilt wheel when one exists.
  • Verify exactly as in Windows Step 5 (import triton, import sageattention, torch still imports with CUDA).

Mac

Triton and SageAttention are N/A on Mac. There is no CUDA. Do not attempt to install them. Use PyTorch sdpa attention (the fallback above is the permanent answer), which on Apple Silicon runs on the MPS backend. Set any attention_mode to sdpa, never load torch.compile/inductor (Triton) nodes, and run unaccelerated. If a workflow hard-requires sageattn, edit it to sdpa rather than trying to satisfy the dependency.


Verification checklist (any OS)

  1. import triton succeeds and prints a version matching your torch (table above).
  2. import sageattention succeeds.
  3. torch STILL imports and torch.cuda.is_available() is True (the install didn't break the env).
  4. ComfyUI startup log: no Could not load sageattention, no triton: unavailable.
  5. In the graph: attention_mode = sageattn loads without the No module named 'sageattention' ValueError; a torch.compile/WanVideoTorchCompileSettings node completes its (slow) first-run compile and then samples.
  6. A real render completes and looks correct (SageAttention can rarely introduce NaN/noise on some GPUs; if output degrades vs. sdpa, fall back to sdpa).

Traps

  • Wrong python / global pip. Installing into system python (or the wrong venv) means ComfyUI never imports it, so the loader still crashes. Always "<that exact python>" -m pip; for Portable that's python_embeded\python.exe, for Desktop the standalone-env\python.exe. Verify with pip show sageattention run by that python.
  • torch / CUDA / python wheel mismatch breaks torch. Installing a cu128 wheel on a cu130 torch (or a torch2.9 wheel on torch2.10) can drag in mismatched CUDA DLLs and break import torch itself, or show up as a runtime DLL error. Match cu128↔12.x / cu130↔13.0 and the torch minor exactly. Pin and verify: before installing, record pip freeze | grep -i torch; after, confirm torch still imports with CUDA. If broken, roll back (pip install torch==<old>+cu<line> --index-url https://download.pytorch.org/whl/cu<line>, or uninstall the bad wheel) and re-apply the sdpa fallback.
  • Stale Triton cache after a torch/GPU/driver change. Triton caches compiled kernels in ~/.triton (%USERPROFILE%\.triton on Windows). After upgrading torch, swapping GPUs, a driver update, or a failed compile, that cache can go stale and cause torch.compile/SageAttention runs to fail even though the install is correct. Symptoms are recurring compile errors, RuntimeError in a Triton kernel, or a hang on the first sample. Fix: clear the cache and re-run (Triton recompiles fresh):
    # Windows
    rmdir /s /q "%USERPROFILE%\.triton"
    # macOS / Linux
    rm -rf ~/.triton
    Safe to delete; it's a pure cache. Do this BEFORE assuming the wheel is wrong (it's a much cheaper fix than a reinstall or roll-back). If it recurs every run, the install is mismatched (see the wheel-mismatch trap above).
  • MSVC missing (Windows Triton). torch.compile/Triton errors like "Microsoft Visual C++ ... required", cl.exe not found, or PY_SSIZE_T_CLEAN/DLL load failures usually mean no MSVC toolchain. Install Visual Studio Build Tools (C++ workload) plus the latest "Visual C++ Redistributable 2015-2022"; copying msvcp140.dll/vcruntime140*.dll into the python folder is the documented last-resort fix.
  • Embedded python has no headers. Portable's python_embeded lacks include/libs, so Triton can't compile and torch.compile fails. Copy the matching python_<ver>_include_libs.zip include and libs (not lib) folders from the triton-windows releases into python_embeded\.
  • py3.13 "no wheel" panic. SageAttention's Windows wheels are cp310-abi3, so one wheel covers py3.10 through 3.13+. The absence of a cp313 filename is normal; do not conclude "no wheel for 3.13." (Source builds, by contrast, can lag on the newest python, another reason to use the abi3 wheel.) Triton-windows does ship py3.13-specific builds.
  • CUDA line confusion. torch.version.cuda is the source of truth: 12.8 → pick cu128 wheels, 13.0 → cu130. Don't read the system CUDA driver version. Match what torch was built against.
  • "Install can break torch." Treat every acceleration install as risky to the env. Get a working sdpa render first, capture the torch version, install, re-verify torch, and be ready to roll back. Never leave the user with a broken torch and no render.
  • SageAttention numerical artifacts. On some GPUs (reported on H100/Hopper) sageattn produces noise that sdpa doesn't. If a render looks worse than the sdpa version, switch that workflow back to sdpa. Correctness over speed.
  • First torch.compile run is slow. Inductor compiles on the first sample (tens of seconds to minutes); that's expected, not a hang. Subsequent runs are fast. Don't "fix" it by ripping out compile unless it actually errors.

See also

  • video-extend. TRAP 5 is the canonical example. The Pusa graph ships with attention_mode=sageattn and WanVideoTorchCompileSettings; this skill is how you either satisfy or safely fall back from that. Read its TRAP 5 for the exact node-by-node sdpa fix.
  • troubleshooting. "Torch / CUDA Version Errors" and "Missing Nodes" sections for diagnosing a torch env that an install broke.
  • installer-packs. Packs note SageAttention/ Triton requirements in pack.yaml notes/post_install; acceleration is an opt-in post-install step, never baked into a model download.

Sources

© artokun, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in plugin/skills/triton-sageattention of artokun/comfyui-mcp.

Open the folder on GitHubat commit 6ad6fc0

Compare with similar skills

Triton Sageattention next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Triton Sageattention compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Triton Sageattention this skillartokun/comfyui-mcp803—~5kAutomated safety check: PassMIT
ComfyUI Custom Node BuilderConstantineB6/comfy-pilot230—~897Automated safety check: PassMIT
Add Comfyui NodeMooshieblob1/MooshieUI207—~936Automated safety check: PassAGPL-3.0
Migrate Workflow Ec2 To Osdcpytorch/test-infra113—~2kAutomated safety check: PassCustom licence
Edit Comfy Workflowpeteromallet/VibeComfy150—~2.2kAutomated safety check: PassMIT
ComfyUI Custom Node Basicsjtydhr88/comfyui-custom-node-skills296—~1.6kAutomated safety check: PassMIT

Similar skills

  • ComfyUI Custom Node Builder

    ConstantineB6/comfy-pilot

    Helps an agent write ComfyUI custom nodes in Python, including wrapping an existing script, mapping data types and handling image batches.

    230 GitHub stars~897 tokensUpdated 7 mo ago
    AI & LLM EngineeringAuto-check passed
  • Add Comfyui Node

    Mooshieblob1/MooshieUI

    Adds a custom ComfyUI Python node to MooshieUI — Python class in mooshienodes.py, Rust required-class registration, and optional workflow template chain hookup.

    207 GitHub stars~936 tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Step-by-step playbook for migrating a pytorch/pytorch .github/workflows/.yml from EC2 to OSDC (ARC) runners — covers both dial-up and 100% opt-in patterns, with the inputs that must be plumbed…

    113 GitHub stars~2k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Edit Comfy Workflow

    peteromallet/VibeComfy

    Edit an existing VibeComfy or ComfyUI workflow, ready template, recipe, scratchpad, or target graph.

    150 GitHub stars~2.2k tokensUpdated 10 days ago
    AI & LLM EngineeringAuto-check passed
  • ComfyUI Custom Node Basics

    jtydhr88/comfyui-custom-node-skills

    Explains the V3 API for ComfyUI custom nodes: node classes, schema, inputs and outputs, registration and how it differs from the legacy V1 style.

    296 GitHub stars~1.6k tokensUpdated 2 mo ago
    AI & LLM EngineeringAuto-check passed
  • Setup

    guaardvark/guaardvark

    Connect this agent to a running Guaardvark (self-hosted AI studio) and check what it can do right now.

    258 GitHub stars~1.2k tokensUpdated today
    AI & LLM EngineeringAuto-check passed

More from artokun/comfyui-mcp

All 42 skills in this repo
  • AI Toolkit Trainer

    artokun/comfyui-mcp

    Train custom LoRAs with ostris AI-Toolkit. An agent skill from artokun/comfyui-mcp.

    803 GitHub stars~2.7k tokensUpdated 6 days ago
    Auto-check passed
  • Anima Base

    artokun/comfyui-mcp

    Anime/illustration text-to-image (ANIMA 1.0, ~2B Cosmos DiT).

    803 GitHub stars~4k tokensUpdated 6 days ago
    Auto-check passed
  • Civitai

    artokun/comfyui-mcp

    Discover Civitai models with the BUILT-IN downloadmodel action:"searchcivitai" and install/generate them locally.

    803 GitHub stars~1.1k tokensUpdated 6 days ago
    Auto-check passed
  • Color Correction

    artokun/comfyui-mcp

    Diagnose and fix video/image color OBJECTIVELY with the getimage (action:"analyzecolor") tool (scopes/stats such as black/white points, contrast, saturation, clipping, cast) instead of eyeballing a…

    803 GitHub stars~2.4k tokensUpdated 6 days ago
    Auto-check passed
  • Comfyui Frontend Extensions

    artokun/comfyui-mcp

    Authoring ComfyUI v2 frontend extensions with @comfyorg/extension-api, covering defineNode/defineExtension/defineWidget, shell UI (sidebar tabs, commands, hotkeys), typed events, and handles.

    803 GitHub stars~5.4k tokensUpdated 6 days ago
    Auto-check passed
  • Comfyui Launch Flags

    artokun/comfyui-mcp

    Pick the right ComfyUI startup flags for VRAM, attention, caching, and speed.

    803 GitHub stars~3.1k tokensUpdated 6 days ago
    Auto-check passed

Questions about Triton Sageattention

What does Triton Sageattention do?

Install Triton + SageAttention to accelerate ComfyUI (the sageattn attentionmode and inductor torch.compile used by WanVideoWrapper / many video graphs). Triton Sageattention is an agent skill from artokun/comfyui-mcp.compile used by WanVideoWrapper / many video graphs).

When should I use Triton Sageattention?

Triton Sageattention fits situations like: A loader crashes with No module named sageattention; reports triton unavailable; asked to speed up Wan/video workflows; deciding whether to install acceleration vs.

How do I install Triton Sageattention in Claude Code?

Run `npx skills add artokun/comfyui-mcp --skill triton-sageattention -a claude-code`. Or copy the skill folder (plugin/skills/triton-sageattention in artokun/comfyui-mcp) into .claude/skills/triton-sageattention in your project. Claude Code loads it when a task matches its description.

How do I install Triton Sageattention in Codex?

Run `npx skills add artokun/comfyui-mcp --skill triton-sageattention -a codex`. Or copy the skill folder (plugin/skills/triton-sageattention in artokun/comfyui-mcp) into .agents/skills/triton-sageattention in your project. Codex loads it when a task matches its description.

Can I use Triton Sageattention in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add artokun/comfyui-mcp --skill triton-sageattention -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/triton-sageattention, .gemini/skills/triton-sageattention, .github/skills/triton-sageattention and .opencode/skills/triton-sageattention in your project.

What does Triton Sageattention need to run?

Going by SKILL.md and its folder, Triton Sageattention needs the command-line tools its instructions call (pip and python). Our summary lists: Python 3.

Does Triton Sageattention access the network?

SKILL.md names 2 domains. In commands or code: github.com; the agent is likely to contact it when it follows the instructions. As links in the text: download.pytorch.org. This is read from the text; nothing was executed.

Is Triton Sageattention safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Triton Sageattention use?

Triton Sageattention is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Triton Sageattention use?

About 5k tokens (SKILL.md is roughly 20k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Triton Sageattention?

Skills that share tags, products or a category with Triton Sageattention: ComfyUI Custom Node Builder (ConstantineB6/comfy-pilot, 230 stars), Add Comfyui Node (Mooshieblob1/MooshieUI, 207 stars), Migrate Workflow Ec2 To Osdc (pytorch/test-infra, 113 stars) and Edit Comfy Workflow (peteromallet/VibeComfy, 150 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Triton Sageattention?

artokun (a GitHub user) maintains it in artokun/comfyui-mcp, which has 803 GitHub stars. The repository holds 42 skills in this directory. The repository was last updated on October 5, 2026.

Source: artokun/comfyui-mcp on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.