O2 Review Loop
openobserve/openobserve
Splits a change into planner, coder and independent reviewer roles: you confirm a spec, a subagent implements it, and a separate reviewer checks each round's local WIP commit.
Runtime-V2 kernel-writing workflow. An agent skill from mirage-project/mirage.
$ npx skills add mirage-project/mirage --skill v2-kernel-writing -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install mirage-project/mirage v2-kernel-writing --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/mirage-project/mirage.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/v2-kernel-writing .claude/skills/v2-kernel-writing && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "v2-kernel-writing" agent skill from https://github.com/mirage-project/mirage/tree/mpk/.claude/skills/v2-kernel-writing into .claude/skills/v2-kernel-writing/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "v2-kernel-writing", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/mirage-project/mirage/tree/mpk/.claude/skills/v2-kernel-writingType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add mirage-project/mirage --skill v2-kernel-writing -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install mirage-project/mirage v2-kernel-writing --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/mirage-project/mirage.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.claude/skills/v2-kernel-writing .agents/skills/v2-kernel-writing && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "v2-kernel-writing" agent skill from https://github.com/mirage-project/mirage/tree/mpk/.claude/skills/v2-kernel-writing into .agents/skills/v2-kernel-writing/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "v2-kernel-writing", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add mirage-project/mirage --skill v2-kernel-writing -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install mirage-project/mirage v2-kernel-writing --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/mirage-project/mirage.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.claude/skills/v2-kernel-writing .cursor/skills/v2-kernel-writing && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "v2-kernel-writing" agent skill from https://github.com/mirage-project/mirage/tree/mpk/.claude/skills/v2-kernel-writing into .cursor/skills/v2-kernel-writing/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "v2-kernel-writing", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/mirage-project/mirage.git --path .claude/skills/v2-kernel-writing--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add mirage-project/mirage --skill v2-kernel-writing -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install mirage-project/mirage v2-kernel-writing --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/mirage-project/mirage.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.claude/skills/v2-kernel-writing .gemini/skills/v2-kernel-writing && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "v2-kernel-writing" agent skill from https://github.com/mirage-project/mirage/tree/mpk/.claude/skills/v2-kernel-writing into .gemini/skills/v2-kernel-writing/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "v2-kernel-writing", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install mirage-project/mirage v2-kernel-writingInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add mirage-project/mirage --skill v2-kernel-writing -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/mirage-project/mirage.git skills-src && mkdir -p .github/skills && cp -r skills-src/.claude/skills/v2-kernel-writing .github/skills/v2-kernel-writing && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "v2-kernel-writing" agent skill from https://github.com/mirage-project/mirage/tree/mpk/.claude/skills/v2-kernel-writing into .github/skills/v2-kernel-writing/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "v2-kernel-writing", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add mirage-project/mirage --skill v2-kernel-writing -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install mirage-project/mirage v2-kernel-writing --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/mirage-project/mirage.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.claude/skills/v2-kernel-writing .opencode/skills/v2-kernel-writing && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "v2-kernel-writing" agent skill from https://github.com/mirage-project/mirage/tree/mpk/.claude/skills/v2-kernel-writing into .opencode/skills/v2-kernel-writing/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "v2-kernel-writing", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
v2-kernel-writingRuntime-V2 kernel-writing workflow. An agent skill from mirage-project/mirage.
V2 Kernel Writing is an agent skill from mirage-project/mirage. Runtime-V2 kernel-writing workflow. Use when writing, porting, or rewriting ANY Runtime-V2 task kernel (tasks/blackwellv2/.cuh + registration) — a new op, a v1→v2 port, or a rewrite toward the reference linearsm100v2 warp-role pipeline idiom. Drives the staged loop SPEC→IMPLEMENT→WIRE→VALIDATE→PERF→REVIEW with per-stage subagents and the b200- sub-skills, and enforces the M=1 anti-loop evidence + the v2 protocol invariants (§1.1 dep-prefix, stale-arrival re-init, skipafterstep0, taskoffset wiring).
Its SKILL.md is about 3.7k tokens, which your agent loads only when the skill is triggered. The skill folder holds 11 other files, including reference files (for example `applications/attn-ffn-reference-rewrite-plan.md`, `applications/ferret_dispatch_w13w2.md` and `applications/ffn_item1_spec.md`).
It sits in Agent Workflows, covering Subagents. It works with Git. The repository describes itself as: Mirage Persistent Kernel: Compiling LLMs into a MegaKernel. The licence is Apache-2.0.
7 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit f9eb70c. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
gitFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use git, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
V2 Kernel Writing loads about 3.7k tokens when it runs, and up to ~23k if it reads all its reference files. Until then it costs about 132 tokens; SKILL.md has 1,600 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from mirage-project/mirage at commit f9eb70c, republished under its Apache-2.0 licence (© mirage-project). 1,600 words, ~3,742 tokens.
.claude/skills/v2-kernel-writing/SKILL.md (or your agent's skills folder). This skill also uses 9 other files; get the full folder from GitHub.You are writing a Runtime-V2 task kernel for MPK (dsv3-decode-clean). The house style is
the reference linear_sm100_v2.cuh warp-role pipeline (loader W4 TMA → launcher W5
tcgen05/TMEM → consumers W0-3 epilogue → storer W6 page release), with consumer-only as the
sanctioned idiom for non-GEMM-shaped ops. Quality bar and every protocol invariant live in
this skill's references — the loop below tells you when to read what and what to dispatch.
For MODEL-level bring-up (whole compute graph → demo) use v2-model-support; this skill is
the per-KERNEL inner loop that pipeline dispatches into.
Everything the loop REQUIRES travels with the repo: this skill's references/ +
applications/ docs, the reference kernels (include/mirage/persistent_kernel/tasks/blackwell_v2/),
the harness (tests/runtime_python/blackwell_v2/), and the wiring surface. The rest is
machine-local and DEGRADES GRACEFULLY:
b200-* sub-skills (the Skill-tool names in the index below) are USER-global
(~/.claude/skills/b200-*), NOT in this repo — a fresh clone still has them only under
the same user account. If missing: proceed anyway; references/house-style.md +
references/upstream-kernel-catalog.md carry the distilled protocol/layout contracts.git show mirage-project/runtime_refactor:<path>) need the
remote: git remote add mirage-project https://github.com/mirage-project/mirage.git && git fetch mirage-project runtime_refactor. Optional — the catalog doc is self-contained.~/ferret/ (TEMPLATE_v2.yaml, docs/v2_runtime_notes.md, workspace1..8) — required ONLY
for the ferret-v2 engine; ~/kda-workspaces/ only for kda; ~/kernel_tools/
(ncu_profile.sh) only for the NCU bound-check. Absent ⇒ route Stage 2/5 work to
v2-kernel-engineer (the default anyway) and use b200-kernel-roofline-triage/manual NCU.~/.claude/projects/-home-muhengl-mirage/memory/) — optional context
only (same-user machines). references/m1-decode-evidence.md is the self-contained
distillation of the evidence rows; treat the memory dir as its (optional) source citations.v2-model-support/references/box-orchestration.md) — without it,
deliver with local gates + an explicit "pending TIER-1".mcp__codex__codex) for Stage-6 double-checks — if unconfigured, the
ablation-logic-reviewer pass still runs; note the missing second engine in the verdict.| Doc | What it is | Read at |
|---|---|---|
references/house-style.md | Reference methodology spec (roles, SEM tables, SMEM regions, TMA/tcgen05 patterns, quality bar) | Stage 0, and by every subagent |
references/upstream-kernel-catalog.md | Per-kernel/family pattern catalog of upstream runtime_refactor@0eadb3fd (which family to copy, gotchas, sync-with-upstream list) | Stage 0 (find your op's family) + Stage 1 |
references/m1-decode-evidence.md | ANTI-LOOP map: DEAD / WIN / UNTESTED at M=1 decode | Stage 0 + Stage 1 |
references/wiring-recipe.md | The v2 8-file registration checklist + footguns | Stage 3 |
references/validation-debug.md | Gates, hang/crash tooling, TIER measurement, profiler | Stage 4 + 5 |
references/ferret-v2-dispatch.md | Stage-2/5 engine routing (engineer|ferret-v2|kda) + the ferret-v2 flow/contract | Stage 2 + 5 |
applications/attn-ffn-reference-rewrite-plan.md | The staged attn/FFN rewrite (user directive) | when working that campaign |
applications/ffn_item1_spec.md | Worked Stage-1 SPEC exemplar (W13/W2 per-tile pipeline) — the deliverable shape Stage 1 must produce | Stage 1 (as template) |
applications/ferret_dispatch_w13w2.md | Worked ferret-v2 dispatch brief exemplar (targets/gate/protocol_frozen/budget) | Stage 2/5 ferret dispatches (as template) |
The MAIN THREAD (or one lead subagent for a multi-kernel campaign) is the orchestrator: it
runs Stages 0/3/6 itself and dispatches ONE subagent per heavy stage — a designer
(Stage 1), an implementer (Stage 2 — use the v2-kernel-engineer agent if defined in
.claude/agents/, else general-purpose with that discipline pasted in), a validator
(Stage 4). Subagents do not dispatch subagents. Each dispatch prompt MUST name: the stage's
reference docs (absolute paths), the sub-skills to load via the Skill tool, the op contract,
and the exact deliverable. Stages are sequential; iterate 2↔4 on failures, 5→1 on a perf
verdict that changes the design.
Hard rules for every stage: default build byte-identical (new task types additive; levers env-gated default-OFF); no repo-wide refactors; GPU safety (test-mode first, never crash-loop the megakernel); every non-trivial conclusion → Stage 6 review before acting on it.
Read references/house-style.md + references/m1-decode-evidence.md IN FULL before any
design. Then classify the op: shape (M, N, K / attention / elementwise), dtype, per-token
work, TP/EP sharding, where it sits in the layer DAG (producer/consumer events), and which
evidence rows (D*/W*/U*) touch it. Match the op to an upstream family (A pipeline / B
consumer+regions / C consumer+monolith / D consumer+no-SMEM / E sub-op helper) in
references/upstream-kernel-catalog.md — the family names the file to crib from. If the op
matches a DEAD row and no new mechanism is on offer — STOP and say so; that is a successful
outcome of this stage.
Deliverable: a spec.h-style design doc (markdown). Location convention: campaign items
that should travel with the repo go to
.claude/skills/v2-kernel-writing/applications/<item>_spec.md (exemplar:
applications/ffn_item1_spec.md); throwaway/exploratory specs go to scratch/
(git-ignored, machine-local). The spec contains:
b200-tma-pipeline-designer (stage ring,
swizzle, load-vs-store completion), b200-tcgen05-mma-contract-builder (tile/dtype/
cta_group, SMEM operand layout, I-desc), b200-tmem-lifecycle-planner (TMEM columns,
alloc/dealloc, ld/wait). Anchor every choice to house-style §2/§5.ceil(items/(ntasks*nwarps)).attention_sm100.cuh (house-style §0);
for FA-style rewrites additionally load b200-flash-attention4-planner.b200-scope-layout-dispatch first.Engine choice first (routing rule + flow in references/ferret-v2-dispatch.md):
v2-kernel-engineer = protocol-heavy house-style port / first bring-up (default);
ferret-v2 (ferret-kernel-agent V2 MODE) = beat-a-numeric-TIER-2-target optimization
loop once a faithful FROZEN gate exists (op wired + harness case + live anchor) — also the
"ferret writes from spec" path over a wired stub; kda = verdict-grade honest transfer
when the number decides a campaign verdict and over-claim is costly. Ferret-v2 rearranges
the stages (S3-stub wiring precedes the run — the gate substrate is the in-tree harness);
its friction escape falls back to engineer-shell + ferret-math-only-body.
Write <op>_v2.cuh + <op>_v2_spec.h to the spec. House-style code conventions:
namespace kernel { namespace <op>_v2 {; spec.h constants + static_asserts pinning every
mirrored constant; role-split __device__ __noinline__ functions (one per role).fence.mbarrier_init.release.cluster after inits;
tcgen05.fence::after_thread_sync at MMA↔wait boundaries. Before finalizing, invoke
b200-mbarrier-protocol-auditor on the barrier ledger (every mbar: init count, arrivers,
tx-bytes, waiters, phase evolution, re-init ownership).extern __shared__ __align__(1024); SMEM only via task_desc->smem_region_offset(REGION_*).__syncthreads() in role loops — named barriers (bar.sync <free-id>, 128) or
tag-flags only; elect_sync() for single-thread issue; no blockIdx for identity.b200-layout-contract-auditor. Build-flag doubts (sm_100a, -rdc=true) →
blackwell-build-compatibility-auditor.Follow references/wiring-recipe.md top to bottom — enum, task_header include, register fn
(§1.1 dep-prefix is the first line of the consumer body — MANDATORY), graph.cc tuple,
runtime.cc task_type_to_name + task_offset=bid.x list, py wrapper (num_tasks==num_workers
gate if grid-wide), builder use_v2 branch, skip_after_step0 on any monotonic-barrier scratch.
Tick the §10 ship checklist explicitly.
In order, no skipping: (1) test-mode numeric vs torch in tests/runtime_python/blackwell_v2/
(cos ≥ 0.999, rel_max ≤ 3e-2, no NaN; v1-counterpart compare); (2) §1.1/protocol static audit;
(3) in-MPK --layers 0-3 probe; (4) MULTI-STEP run, iter ≥ 3 — iter-0-fine/iter-1-hang =
persistent-state re-init (skip_after_step0), NOT a missing event; (5) on any hang: watchdog
(-DMPK_V2_BREADCRUMB + MPK_V2_HANG_WATCHDOG_S); on any crash: compute-sanitizer memcheck
= ground truth (breadcrumb counts are base-rate-biased). Deadlock/wrong-result debugging →
b200-warp-specialized-debugger (roles/storage/handoff/lifetime worksheet, one handoff at a
time). Math-changing on TP8 → poison-fill gate, not token-identity.
TIER hierarchy is the law: TIER 1 in-MPK %globaltimer slowCTA @ production grid = the only
verdict-grade number; harness slowCTA corroborates; cudaEvent-wall / standalone-warm are
diagnostic-only. Compare against the reference/v1 body anchor from the spec. Bottleneck
classification → b200-kernel-roofline-triage (achievable-floor rules from
m1-decode-evidence §4 apply — same-grid xor-consumer floor, never theoretical peak). For a
pipeline kernel that is correct-but-slow, climb b200-gemm-optimization-ladder one rung at a
time. Profiler: buffer = 120000*128 entries; export via scripts/v2_perfetto_export.py.
For a sustained beat-a-numeric-target optimization loop on one kernel, dispatch ferret-v2
(references/ferret-v2-dispatch.md): it iterates the pair against the frozen harness gate
(TIER-2 body_span) autonomously; TIER-1 in-MPK slowCTA stays the final verdict here.
For a whole perf-optimization CAMPAIGN around this kernel (measure→plan→implement→re-measure
→land, agent roster + history contract) use the sibling skill v2-perf-iteration.
ablation-logic-reviewer subagent + Codex MCP double-check (default params)
BEFORE acting on or reporting it.mpk-correctness-gate for anything math-adjacent, then mpk-commit-reviewer
before git commit (staged-path + byte-identity + message gates). Verdicts →
mpk-memory-keeper (experiment_history INDEX + memory; update m1-decode-evidence sources).| Sub-skill | Use at | For |
|---|---|---|
b200-scope-layout-dispatch | S1 | op→kernel mapping: scope/layout/dispatch/handoff contract |
b200-tma-pipeline-designer | S1/S2 | TMA descriptors, stage ring, swizzle, completion protocol |
b200-tcgen05-mma-contract-builder | S1/S2 | MMA tile/dtype/descriptor contract |
b200-tmem-lifecycle-planner | S1/S2 | TMEM columns, alloc/ld/wait/dealloc lifecycle |
b200-flash-attention4-planner | S1 (attn rewrites) | QKᵀ/PV + online-softmax tile & barrier graph |
b200-mbarrier-protocol-auditor | S2 gate | per-barrier ledger audit before finalizing |
b200-layout-contract-auditor | S2/S4 | shape-stride/swizzle/operand-contract bugs |
blackwell-build-compatibility-auditor | S2/S3 | sm_100a flags, PTX/cubin, JIT |
b200-warp-specialized-debugger | S4 | deadlock / IMA / wrong-result / correct-but-slow |
b200-kernel-roofline-triage | S5 | bound classification + minimal falsifying experiment |
b200-gemm-optimization-ladder | S5 | staged GEMM perf climb with gates |
© mirage-project, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 9 other files (references) in .claude/skills/v2-kernel-writing of mirage-project/mirage.
Open the folder on GitHubat commit f9eb70c
V2 Kernel Writing next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| V2 Kernel Writing this skillmirage-project/mirage | 2.5k | — | ~3.7k | Automated safety check: Pass | Apache-2.0 | |
| O2 Review Loopopenobserve/openobserve | 22k | — | ~3.7k | Automated safety check: Pass | AGPL-3.0 | |
| ClawTeam Multi-Agent Swarmwin4r/ClawTeam-OpenClaw | 1.5k | 1 repos | ~2.9k | Automated safety check: Pass | MIT | |
| Agent Deckasheshgoplani/agent-deck | 1k | — | ~1.7k | Automated safety check: Pass | MIT | |
| Clawteamwin4r/ClawTeam-OpenClaw | 1.5k | — | ~3.1k | Automated safety check: Pass | MIT | |
| Puppetmaster Agent Orchestrationprofessorpalmer/Puppetmaster | 467 | — | ~3.2k | Automated safety check: Pass | MIT |
openobserve/openobserve
Splits a change into planner, coder and independent reviewer roles: you confirm a spec, a subagent implements it, and a separate reviewer checks each round's local WIP commit.
win4r/ClawTeam-OpenClaw
Launches a swarm of specialist Hermes agents in git-worktree-isolated tmux windows with a kanban board and file-based inboxes, using built-in templates like hedge-fund and code-review.
asheshgoplani/agent-deck
agent-deck, the terminal session manager for AI coding agents.
win4r/ClawTeam-OpenClaw
Multi-agent swarm orchestration. An agent skill from win4r/ClawTeam-OpenClaw.
professorpalmer/Puppetmaster
Operates and supervises Puppetmaster, a multi-agent orchestrator, through its MCP tools or CLI, picking the right verb for edits, reviews, audits and long-running jobs.
leeguooooo/claude-code-usage-bar
Manage cs (claude-statusbar) — switch theme/style/density, override severity colors, preview combinations, run doctor, reset config, install, upgrade (cs upgrade — the only supported upgrade path)…
mirage-project/mirage
Runtime-V2 performance-iteration workflow. An agent skill from mirage-project/mirage.
mirage-project/mirage
Step-by-step guide for adding a new task implementation to Mirage Persistent Kernel (MPK).
mirage-project/mirage
A skill your agent uses when the user wants to design or extend a FlashAttention-style forward kernel on B200/Blackwell, involving the two MMAs QKᵀ and PV, online softmax, S/P/O in TMEM, warp roles…
mirage-project/mirage
Build or run a FAITHFUL in-MPK per-task latency gate (slowCTA at the production grid + cos) for a DeepSeek-V3 MPK decode kernel or shape.
mirage-project/mirage
A skill your agent uses when a batch of env-gated (ifdef MPKDSV3 / os.environ-controlled, default-OFF) MPK optimization levers needs to be consolidated into a single clean code path for a PR…
mirage-project/mirage
Guide for using MPK test mode to unit-test individual layers or multi-layer pipelines through the full compilation pipeline.
Works with
Categories
Runtime-V2 kernel-writing workflow. An agent skill from mirage-project/mirage. V2 Kernel Writing is an agent skill from mirage-project/mirage. Runtime-V2 kernel-writing workflow.
V2 Kernel Writing fits situations like: rewriting ANY Runtime-V2 task kernel (tasks/blackwellv2/.cuh + registration) — a new op; A rewrite toward the reference linearsm100v2 warp-role pipeline idiom.
Run `npx skills add mirage-project/mirage --skill v2-kernel-writing -a claude-code`. Or copy the skill folder (.claude/skills/v2-kernel-writing in mirage-project/mirage) into .claude/skills/v2-kernel-writing in your project. Claude Code loads it when a task matches its description.
Run `npx skills add mirage-project/mirage --skill v2-kernel-writing -a codex`. Or copy the skill folder (.claude/skills/v2-kernel-writing in mirage-project/mirage) into .agents/skills/v2-kernel-writing in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add mirage-project/mirage --skill v2-kernel-writing -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/v2-kernel-writing, .gemini/skills/v2-kernel-writing, .github/skills/v2-kernel-writing and .opencode/skills/v2-kernel-writing in your project.
Going by SKILL.md and its folder, V2 Kernel Writing needs the command-line tools its instructions call (git).
SKILL.md contains no URLs. Its commands use git, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
V2 Kernel Writing is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 3.7k tokens (SKILL.md is roughly 15k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 19k tokens, read only when the agent opens those files.
Skills that share tags, products or a category with V2 Kernel Writing: O2 Review Loop (openobserve/openobserve, 22k stars), ClawTeam Multi-Agent Swarm (win4r/ClawTeam-OpenClaw, 1.5k stars), Agent Deck (asheshgoplani/agent-deck, 1k stars) and Clawteam (win4r/ClawTeam-OpenClaw, 1.5k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
mirage-project (a GitHub organization) maintains it in mirage-project/mirage, which has 2,541 GitHub stars. The repository holds 24 skills in this directory. The repository was last updated on October 7, 2026.
Source: mirage-project/mirage on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.