Agent skill

Fla Dispatch Backends

by fla-org in fla-org/flash-linear-attention

Workflow for FLA backend dispatch decorators and backend implementations.

MITAuto-check passed

Install Fla Dispatch Backends

skills CLI
$ npx skills add fla-org/flash-linear-attention --skill fla-dispatch-backends -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install fla-org/flash-linear-attention fla-dispatch-backends --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/fla-org/flash-linear-attention.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/fla-dispatch-backends .claude/skills/fla-dispatch-backends && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
fla-dispatch-backends
GitHub stars
5.8k
Token cost
~1.1k tokens
SKILL.md length
533 words
Files
1
Skills in repo
9
Repo updated
First seen
Licence
MIT

At a glance

Workflow for FLA backend dispatch decorators and backend implementations.

  • Works in 6 steps: Add a BaseBackend subclass under the… → Set backend_type, package_name, env_var,… → Implement _verifier(...) with the same… → …
  • Touching fla.ops.backends
  • SKILL.md covers Core model, Backend implementation checklist, Verifier rules and Decorator placement, plus 2 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Fla Dispatch Backends is an agent skill from fla-org/flash-linear-attention. Workflow for FLA backend dispatch decorators and backend implementations. Use when touching fla.ops.backends, @dispatch-decorated functions, BaseBackend subclasses, backend verifier methods, backend env vars, or backend tests.

Its SKILL.md is about 1.1k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

The repository describes itself as: 🚀 Efficient implementations for emerging model architectures. The licence is MIT.

When your agent uses it

  • Touching fla.ops.backends
  • @dispatch-decorated functions
  • BaseBackend subclasses
  • Backend verifier methods

Example prompts

  • “/fla-dispatch-backends”

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. Add a BaseBackend subclass under the operation's backends/ package.
  2. Set backend_type, package_name, env_var, default_enable, and priority.
  3. Implement _verifier(...) with the same public call
  4. Implement (...) and keep return values identical to
  5. Register the backend in the operation's backends/init.py.
  6. Add tests that cover accepted dispatch, verifier rejection, and fallback.

What it can do on your machine

Read from SKILL.md and the folder at commit b8ff848. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Fla Dispatch Backends loads about 1.1k tokens when it runs. Until then it costs about 62 tokens; SKILL.md has 533 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~62
When it runs · the whole SKILL.md, loaded when a task matches
~1.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from fla-org/flash-linear-attention at commit b8ff848, republished under its MIT licence (© fla-org). 533 words, ~1,136 tokens.

Download SKILL.mdSave it as .claude/skills/fla-dispatch-backends/SKILL.md (or your agent's skills folder).
name
fla-dispatch-backends
description
Workflow for FLA backend dispatch decorators and backend implementations. Use when touching fla.ops.backends, @dispatch-decorated functions, BaseBackend subclasses, backend verifier methods, backend env vars, or backend tests.

FLA Dispatch Backends Skill

Use this skill for the runtime backend dispatch system implemented in fla/ops/backends/__init__.py.

Core model

  • Public functions opt in with @dispatch('<operation>').
  • First call lazily imports fla.ops.<operation>.backends, unless the operation has a custom module in _OPERATION_BACKEND_MODULES (for example modules).
  • Backend modules create BackendRegistry('<operation>') and register BaseBackend subclasses.
  • Dispatch tries registered backends sorted by priority where lower means higher priority.
  • A backend is considered only when is_available() and is_enabled() are both true.
  • Runtime dispatch checks is_available() and is_enabled() directly; do not rely on the cached can_use() path inside code that must be torch.compile friendly.
  • If <func_name>_verifier exists, it must return (True, None) or (False, reason). Rejected calls fall back to the next backend.
  • If no backend handles the call, dispatch runs the original implementation.
  • FLA_DISABLE_BACKEND_DISPATCH=1 bypasses the decorator entirely.
  • The dispatch wrapper is marked with torch.compiler.disable, so keep backend selection logic outside compiled graphs and keep compiled work inside the selected backend implementation.

Backend implementation checklist

For a new backend:

  1. Add a BaseBackend subclass under the operation's backends/ package.
  2. Set backend_type, package_name, env_var, default_enable, and priority.
  3. Implement <public_function_name>_verifier(...) with the same public call surface as the decorated function.
  4. Implement <public_function_name>(...) and keep return values identical to the default implementation.
  5. Register the backend in the operation's backends/__init__.py.
  6. Add tests that cover accepted dispatch, verifier rejection, and fallback.

Verifier rules

  • Verifiers must be cheap, deterministic, and side-effect free.
  • Return a specific rejection reason; it is logged once and is useful in CI logs.
  • Check dtype, shape, layout, inference/training mode, external package requirements, env flags, and unsupported options before calling backend code.
  • Do not silently copy or normalize inputs in a verifier; do that in the backend implementation only when it is part of the backend contract.
  • If a backend supports only inference, check torch.is_grad_enabled() or torch.is_inference_mode_enabled() as appropriate.
  • Do not mutate global backend registries, environment variables, tensors, RNG state, or caches from a verifier.
Show full SKILL.md (214 more words)Show less

Decorator placement

  • Decorate public operation entry points, not private helpers that are only used inside one backend.
  • Keep the decorated function as the semantic fallback implementation. A user should be able to set FLA_DISABLE_BACKEND_DISPATCH=1 and still get the same API behavior.
  • Use the operation name that maps to the backend package. For normal ops, @dispatch('kda') maps to fla.ops.kda.backends; special cases belong in _OPERATION_BACKEND_MODULES.
  • Do not add import-time side effects in backend packages beyond registering backends.

Testing guidance

  • Test the public decorated function, not only the backend helper.
  • Force dispatch off with FLA_DISABLE_BACKEND_DISPATCH=1 when comparing against the Triton/default path.
  • Force or disable backend-specific env vars (FLA_FLASH_KDA, FLA_TILELANG, FLA_INTRACARD_CP) when testing route behavior.
  • Include at least one rejection test for each verifier branch added or changed.
  • For backend changes under fla/ops/<op>/backends/, ensure dependent op tests still run; scripts/find_dependent_tests.py maps backend changes back to the decorated op files.

Style constraints

  • Use platform helpers from fla.utils for hardware/platform decisions instead of adding new direct torch.cuda checks in public code or tests. If no helper covers the condition, add a small helper in fla.utils first.
  • Keep backend imports lazy inside backend implementations when importing an optional package would otherwise break environments without that package.
  • Keep error/rejection messages precise and user-facing; they appear in logs and tests may assert them.

© fla-org, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .agents/skills/fla-dispatch-backends of fla-org/flash-linear-attention.

Open the folder on GitHubat commit b8ff848

Compare with similar skills

Fla Dispatch Backends next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Fla Dispatch Backends compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Fla Dispatch Backends this skillfla-org/flash-linear-attention5.8k—~1.1kAutomated safety check: PassMIT
Implementsickn33/agentic-awesome-skills47k5 repos~306Automated safety check: PassMIT
Backend Patternsaffaan-m/ECC274k1 repos~3.5kAutomated safety check: PassMIT
Backend Patternsaffaan-m/ECC274k—~3.4kAutomated safety check: PassMIT
Backendredis/RedisInsight8.9k—~1.6kAutomated safety check: PassCustom licence
Dispatchsickn33/agentic-awesome-skills47k1 repos~2.1kAutomated safety check: PassMIT

Similar skills

  • Implement

    sickn33/agentic-awesome-skills

    Implement a piece of work based on a PRD or set of issues. An agent skill from sickn33/agentic-awesome-skills.

    47k GitHub starsUsed in 5 repos~306 tokens
    Product & Project ManagementAuto-check passed
  • Backend Patterns

    affaan-m/ECC

    Node.js, Express ve Next.js API routes için backend mimari kalıpları, API tasarımı, veritabanı optimizasyonu ve sunucu tarafı en iyi uygulamalar.

    274k GitHub starsUsed in 1 repo~3.5k tokens
    Backend & APIsAuto-check passed
  • Backend Patterns

    affaan-m/ECC

    Patrones de arquitectura backend, diseño de API, optimización de base de datos y buenas prácticas del lado del servidor para Node.js, Express y rutas API de Next.js.

    274k GitHub stars~3.4k tokensUpdated 2 days ago
    Backend & APIsAuto-check passed
  • Backend

    redis/RedisInsight

    Official

    NestJS backend development patterns for the RedisInsight API: module structure, services, controllers, DTOs, dependency injection, and error handling.

    8.9k GitHub stars~1.6k tokensUpdated 3 days ago
    Backend & APIsAuto-check passed
  • Dispatch

    sickn33/agentic-awesome-skills

    Delegate tasks to OpenAI Codex CLI and Google Antigravity CLI from Claude Code with topic-aware sessions

    47k GitHub starsUsed in 1 repo~2.1k tokens
    Auto-check passed
  • Senior Backend

    alirezarezvani/claude-skills

    Designs and implements backend systems including REST APIs, microservices, database architectures, authentication flows, and security hardening.

    28k GitHub starsUsed in 1 repo~3.8k tokens
    Backend & APIsAuto-check passed

More from fla-org/flash-linear-attention

All 9 skills in this repo
  • Fla Ascend Performance

    fla-org/flash-linear-attention

    Guidelines for Ascend NPU kernel / Triton-Ascend backend performance work in the FLA repo.

    5.8k GitHub stars~5.6k tokensUpdated yesterday
    Auto-check passed
  • Fla Optimization Loop

    fla-org/flash-linear-attention

    Disciplined, reproducible loop for making an FLA kernel faster (Triton, Gluon, TileLang, CuTe) without ever breaking or gaming correctness.

    5.8k GitHub stars~2.6k tokensUpdated yesterday
    Auto-check passed
  • Fla Triton To Gluon

    fla-org/flash-linear-attention

    Workflow for porting an existing Triton kernel in fla/ops/ to Gluon (triton.experimental.gluon) to gain explicit control over tensor layouts, shared memory, async data movement (cp.async / TMA), MMA…

    5.8k GitHub stars~4.2k tokensUpdated yesterday
    Auto-check passed
  • Fla Correctness Coverage

    fla-org/flash-linear-attention

    Guidelines for kernel correctness testing and coverage in fla/ops/ and related modules, including common Triton grid/addressing pitfalls.

    5.8k GitHub stars~1.2k tokensUpdated yesterday
    Auto-check passed
  • Fla Design Coverage

    fla-org/flash-linear-attention

    Contract-first design and coverage discipline for FLA kernel and numerical changes.

    5.8k GitHub stars~3.3k tokensUpdated yesterday
    Auto-check passed
  • Fla Kda

    fla-org/flash-linear-attention

    FLA KDA kernel workflow and public technical notes. An agent skill from fla-org/flash-linear-attention.

    5.8k GitHub stars~1.3k tokensUpdated yesterday
    Auto-check passed

Questions about Fla Dispatch Backends

What does Fla Dispatch Backends do?

Workflow for FLA backend dispatch decorators and backend implementations. Fla Dispatch Backends is an agent skill from fla-org/flash-linear-attention. Workflow for FLA backend dispatch decorators and backend implementations.

When should I use Fla Dispatch Backends?

Fla Dispatch Backends fits situations like: touching fla.ops.backends; @dispatch-decorated functions; baseBackend subclasses; backend verifier methods.

How do I install Fla Dispatch Backends in Claude Code?

Run `npx skills add fla-org/flash-linear-attention --skill fla-dispatch-backends -a claude-code`. Or copy the skill folder (.agents/skills/fla-dispatch-backends in fla-org/flash-linear-attention) into .claude/skills/fla-dispatch-backends in your project. Claude Code loads it when a task matches its description.

How do I install Fla Dispatch Backends in Codex?

Run `npx skills add fla-org/flash-linear-attention --skill fla-dispatch-backends -a codex`. Or copy the skill folder (.agents/skills/fla-dispatch-backends in fla-org/flash-linear-attention) into .agents/skills/fla-dispatch-backends in your project. Codex loads it when a task matches its description.

Can I use Fla Dispatch Backends in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add fla-org/flash-linear-attention --skill fla-dispatch-backends -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/fla-dispatch-backends, .gemini/skills/fla-dispatch-backends, .github/skills/fla-dispatch-backends and .opencode/skills/fla-dispatch-backends in your project.

What does Fla Dispatch Backends need to run?

SKILL.md names no scripts, command-line tools or credentials: Fla Dispatch Backends is instructions for the agent only.

Does Fla Dispatch Backends access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Fla Dispatch Backends safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Fla Dispatch Backends use?

Fla Dispatch Backends is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Fla Dispatch Backends use?

About 1.1k tokens (SKILL.md is roughly 4.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Fla Dispatch Backends?

Skills that share tags, products or a category with Fla Dispatch Backends: Implement (sickn33/agentic-awesome-skills, 47k stars), Backend Patterns (affaan-m/ECC, 274k stars), Backend Patterns (affaan-m/ECC, 274k stars) and Backend (redis/RedisInsight, 8.9k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Fla Dispatch Backends?

fla-org (a GitHub organization) maintains it in fla-org/flash-linear-attention, which has 5,828 GitHub stars. The repository holds 9 skills in this directory. The repository was last updated on October 6, 2026.

Source: fla-org/flash-linear-attention on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.