Agent skill

Add Target Atom Op

by ROCm in ROCm/FlyDSL

Add a new target-specific Mma / Copy Op type to a FlyDSL backend dialect (lib/Dialect/Fly<TARGET/<SUBTARGET/ + include/flydsl/Dialect/Fly<TARGET/IR/).

Custom licenceAuto-check: notesAI & LLM Engineering

Install Add Target Atom Op

skills CLI
$ npx skills add ROCm/FlyDSL --skill add-target-atom-op -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install ROCm/FlyDSL add-target-atom-op --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/ROCm/FlyDSL.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/add-target-atom-op .claude/skills/add-target-atom-op && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
add-target-atom-op
GitHub stars
290
Token cost
~6.9k tokens
SKILL.md length
2,556 words
Files
1
Skills in repo
19
Repo updated
First seen
Licence
Custom licence

At a glance

Add a new target-specific Mma / Copy Op type to a FlyDSL backend dialect (lib/Dialect/Fly<TARGET/<SUBTARGET/ + include/flydsl/Dialect/Fly<TARGET/IR/).

  • Works in 5 steps: Inherent Design: How FlyDSL Atoms Work → The Files You Will Touch → Recipe: Add a Stateless MmaOp → …
  • Adding a new tensor-core / matrix instruction (MFMA
  • SKILL.md covers 1. Inherent Design: How FlyDSL…, 2. The Files You Will Touch, 3. Recipe: Add a Stateless MmaOp and 4. Recipe: Add a Stateful CopyOp, plus 1 more section
  • Calls bash

What it does

Add Target Atom Op is an agent skill from ROCm/FlyDSL. Add a new target-specific Mma / Copy Op type to a FlyDSL backend dialect (lib/Dialect/Fly<TARGET/<SUBTARGET/ + include/flydsl/Dialect/Fly<TARGET/IR/). Covers the MmaOp/CopyOp type design, the stateful-vs-stateless variants, and the emitAtomCall / emitAtomCallSSA lowering contract to the backend dialect (LLVM/ROCDL/NVVM/SPIR-V/...). Use when adding a new tensor-core / matrix instruction (MFMA, WMMA, HMMA, WGMMA, ...), a new buffer / shared-memory / global copy atom, a new stateful copy (per-atom offset or…

Its SKILL.md is about 6.9k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in AI & LLM Engineering, covering GPU and accelerator computing. The repository describes itself as: FlyDSL is the Python front‑end of the project: a Flexible Layout Python DSL for expressing tiling, partitioning, data movement, and kernel structure at a high level.

When your agent uses it

  • Adding a new tensor-core / matrix instruction (MFMA
  • A new buffer / shared-memory / global copy atom
  • A new stateful copy (per-atom offset
  • Bringing up a new backend dialect (FlyPTX

Example prompts

  • “/add-target-atom-op”

Requirements

  • Pre-approved tools (allowed-tools): Read, Edit, Bash, Grep, Glob, Agent

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. Inherent Design: How FlyDSL Atoms Work
  2. The Files You Will Touch
  3. Recipe: Add a Stateless MmaOp
  4. Recipe: Add a Stateful CopyOp
  5. Adding a New AtomStateField (rare)

What it can do on your machine

Read from SKILL.md and the folder at commit 3ad47c1. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Read
    • Edit
    • Bash
    • Grep
    • Glob
    • Agent

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • bash

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Add Target Atom Op loads about 6.9k tokens when it runs. Until then it costs about 172 tokens; SKILL.md has 2,556 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~172
When it runs · the whole SKILL.md, loaded when a task matches
~6.9k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NotePre-approves every shell command (allowed-tools: Bash)SKILL.md
    allowed-tools: Read, Edit, Bash, Grep, Glob, Agent

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 2,556 words (~6,876 tokens).

“Step-by-step recipe for authoring a new MmaOp*Type or CopyOp*Type in a backend dialect (fly_rocdl, or a future fly_ptx / ...), plus the inherent design contract every Op author must understand before writing a single line of code.”

— opening of SKILL.md by ROCm, Custom licence
name
add-target-atom-op
allowed-tools
Read, Edit, Bash, Grep, Glob, Agent

Read the full SKILL.md on GitHub

Files

Just SKILL.md in .claude/skills/add-target-atom-op of ROCm/FlyDSL.

Open the folder on GitHubat commit 3ad47c1

Compare with similar skills

Add Target Atom Op next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Add Target Atom Op compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Add Target Atom Op this skillROCm/FlyDSL290—~6.9kAutomated safety check: NotesCustom licence
Hugging Face Local Model Evalshuggingface/skills11k2 repos~1.6kAutomated safety check: PassApache-2.0
Fla Triton To Gluonfla-org/flash-linear-attention5.8k—~4.2kAutomated safety check: PassMIT
Liger Kernel Perflinkedin/Liger-Kernel6.7k—~1.5kAutomated safety check: PassBSD-2-Clause
Hugging Face LLM Trainerhuggingface/skills11k1 repos~7.2kAutomated safety check: PassApache-2.0
MUSA GPU Training Optimizeropen-infra-skills/infra-skills141—~1.7kAutomated safety check: PassApache-2.0

Similar skills

  • Official

    Runs evaluations of Hugging Face Hub models on local hardware with inspect-ai or lighteval, and helps choose between vLLM, Transformers and accelerate backends.

    11k GitHub starsUsed in 2 repos~1.6k tokens
    AI & LLM EngineeringAuto-check passed
  • Fla Triton To Gluon

    fla-org/flash-linear-attention

    Workflow for porting an existing Triton kernel in fla/ops/ to Gluon (triton.experimental.gluon) to gain explicit control over tensor layouts, shared memory, async data movement (cp.async / TMA), MMA…

    5.8k GitHub stars~4.2k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Liger Kernel Perf

    linkedin/Liger-Kernel

    Optimizes the performance of existing Liger Kernel Triton kernels.

    6.7k GitHub stars~1.5k tokensUpdated yesterday
    AI & LLM EngineeringAuto-check passed
  • Hugging Face LLM Trainer

    huggingface/skills

    Official

    Trains or fine-tunes language and vision models with TRL or Unsloth on Hugging Face Jobs cloud GPUs, then converts the results to GGUF.

    11k GitHub starsUsed in 1 repo~7.2k tokens
    AI & LLM EngineeringAuto-check passed
  • MUSA GPU Training Optimizer

    open-infra-skills/infra-skills

    Profiles, benchmarks and tunes AI training workloads on Moore Threads MUSA GPUs with a measurement-first process that keeps model behavior unchanged.

    141 GitHub stars~1.7k tokensUpdated 3 mo ago
    AI & LLM EngineeringAuto-check passed
  • Cuda Kernel Optimizer

    KernelFlow-ops/cuda-optimized-skill

    Iteratively optimize a CUDA/CUTLASS/Triton kernel only when strict on-device compilation, correctness, timing, and NCU evidence gates pass.

    213 GitHub stars~4.3k tokensUpdated 1 mo ago
    AI & LLM EngineeringAuto-check passed

More from ROCm/FlyDSL

All 19 skills in this repo
  • Llvm

    ROCm/FlyDSL

    Tune and analyse a FlyDSL kernel at the LLVM level: pick a compile hint, function attribute, or AMDGPU backend flag, then PROVE it reached codegen.

    290 GitHub stars~4.9k tokensUpdated today
    Auto-check: notes
  • Detect per-kernel GPU resource regressions (VGPR, SGPR, register spills, scratch, static LDS) by diffing the final ISA before and after a change, using its isaresourcetable.py helper.

    290 GitHub stars~2.7k tokensUpdated today
    Auto-check: notes
  • API Stability

    ROCm/FlyDSL

    Review a FlyDSL PR, commit, branch, kernel, or consuming module for API-stability compliance.

    290 GitHub stars~3.2k tokensUpdated today
    Auto-check: notes
  • Build Rocm Image

    ROCm/FlyDSL

    Connect to a remote host via SSH and build a Docker image with rocprofv3, aiter, and FlyDSL.

    290 GitHub stars~1.2k tokensUpdated today
    Auto-check: notes
  • Debug FlyDSL GPU kernels that produce NaN, inf, wrong results, or crash.

    290 GitHub stars~2.6k tokensUpdated today
    Auto-check: notes
  • Guided step-by-step wizard for producing a new FlyDSL GPU kernel from a requirement: classify the kernel type, pick a skeleton, fill in compute, add control flow / sync / LDS, then test on GPU.

    290 GitHub stars~4.6k tokensUpdated today
    Auto-check: notes

Questions about Add Target Atom Op

What does Add Target Atom Op do?

Add a new target-specific Mma / Copy Op type to a FlyDSL backend dialect (lib/Dialect/Fly<TARGET/<SUBTARGET/ + include/flydsl/Dialect/Fly<TARGET/IR/). Add Target Atom Op is an agent skill from ROCm/FlyDSL. Add a new target-specific Mma / Copy Op type to a FlyDSL backend dialect (lib/Dialect/Fly<TARGET/<SUBTARGET/ + include/flydsl/Dialect/Fly<TARGET/IR/).

When should I use Add Target Atom Op?

Add Target Atom Op fits situations like: adding a new tensor-core / matrix instruction (MFMA; A new buffer / shared-memory / global copy atom; A new stateful copy (per-atom offset; bringing up a new backend dialect (FlyPTX.

How do I install Add Target Atom Op in Claude Code?

Run `npx skills add ROCm/FlyDSL --skill add-target-atom-op -a claude-code`. Or copy the skill folder (.claude/skills/add-target-atom-op in ROCm/FlyDSL) into .claude/skills/add-target-atom-op in your project. Claude Code loads it when a task matches its description.

How do I install Add Target Atom Op in Codex?

Run `npx skills add ROCm/FlyDSL --skill add-target-atom-op -a codex`. Or copy the skill folder (.claude/skills/add-target-atom-op in ROCm/FlyDSL) into .agents/skills/add-target-atom-op in your project. Codex loads it when a task matches its description.

Can I use Add Target Atom Op in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add ROCm/FlyDSL --skill add-target-atom-op -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/add-target-atom-op, .gemini/skills/add-target-atom-op, .github/skills/add-target-atom-op and .opencode/skills/add-target-atom-op in your project.

What does Add Target Atom Op need to run?

Going by SKILL.md and its folder, Add Target Atom Op needs the command-line tools its instructions call (bash). Its frontmatter pre-approves these tools: Read, Edit, Bash, Grep, Glob, Agent.

Does Add Target Atom Op access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Add Target Atom Op safe to install?

Our automated static check of SKILL.md found notes only (pre-approves every shell command (allowed-tools: bash)), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.

What licence does Add Target Atom Op use?

Add Target Atom Op has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Add Target Atom Op use?

About 6.9k tokens (SKILL.md is roughly 28k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Add Target Atom Op?

Skills that share tags, products or a category with Add Target Atom Op: Hugging Face Local Model Evals (huggingface/skills, 11k stars), Fla Triton To Gluon (fla-org/flash-linear-attention, 5.8k stars), Liger Kernel Perf (linkedin/Liger-Kernel, 6.7k stars) and Hugging Face LLM Trainer (huggingface/skills, 11k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Add Target Atom Op?

ROCm (a GitHub organization) maintains it in ROCm/FlyDSL, which has 290 GitHub stars. The repository holds 19 skills in this directory. The repository was last updated on October 9, 2026.

Source: ROCm/FlyDSL on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.