File Organization And Structure
kitchen-engineer42/pdf2skills
Organize routines within files using blank line separation, consider alphabetical ordering when appropriate, and follow C++ standard file structure.
Run the sched2tlx perf/correctness harness over the modulo-scheduling example corpus (case1-9: GEMM, persistent GEMM, FA fwd/bwd, addmm+bias, LayerNorm, wgrad+bias, multiphase GEMM, scaledmm).
$ npx skills add facebookexperimental/triton --skill sched2tlx-perf-testing -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install facebookexperimental/triton sched2tlx-perf-testing --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/facebookexperimental/triton.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/sched2tlx-perf-testing .claude/skills/sched2tlx-perf-testing && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "sched2tlx-perf-testing" agent skill from https://github.com/facebookexperimental/triton/tree/main/.claude/skills/sched2tlx-perf-testing into .claude/skills/sched2tlx-perf-testing/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "sched2tlx-perf-testing", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/facebookexperimental/triton/tree/main/.claude/skills/sched2tlx-perf-testingType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add facebookexperimental/triton --skill sched2tlx-perf-testing -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install facebookexperimental/triton sched2tlx-perf-testing --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/facebookexperimental/triton.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.claude/skills/sched2tlx-perf-testing .agents/skills/sched2tlx-perf-testing && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "sched2tlx-perf-testing" agent skill from https://github.com/facebookexperimental/triton/tree/main/.claude/skills/sched2tlx-perf-testing into .agents/skills/sched2tlx-perf-testing/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "sched2tlx-perf-testing", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add facebookexperimental/triton --skill sched2tlx-perf-testing -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install facebookexperimental/triton sched2tlx-perf-testing --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/facebookexperimental/triton.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.claude/skills/sched2tlx-perf-testing .cursor/skills/sched2tlx-perf-testing && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "sched2tlx-perf-testing" agent skill from https://github.com/facebookexperimental/triton/tree/main/.claude/skills/sched2tlx-perf-testing into .cursor/skills/sched2tlx-perf-testing/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "sched2tlx-perf-testing", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/facebookexperimental/triton.git --path .claude/skills/sched2tlx-perf-testing--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add facebookexperimental/triton --skill sched2tlx-perf-testing -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install facebookexperimental/triton sched2tlx-perf-testing --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/facebookexperimental/triton.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.claude/skills/sched2tlx-perf-testing .gemini/skills/sched2tlx-perf-testing && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "sched2tlx-perf-testing" agent skill from https://github.com/facebookexperimental/triton/tree/main/.claude/skills/sched2tlx-perf-testing into .gemini/skills/sched2tlx-perf-testing/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "sched2tlx-perf-testing", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install facebookexperimental/triton sched2tlx-perf-testingInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add facebookexperimental/triton --skill sched2tlx-perf-testing -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/facebookexperimental/triton.git skills-src && mkdir -p .github/skills && cp -r skills-src/.claude/skills/sched2tlx-perf-testing .github/skills/sched2tlx-perf-testing && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "sched2tlx-perf-testing" agent skill from https://github.com/facebookexperimental/triton/tree/main/.claude/skills/sched2tlx-perf-testing into .github/skills/sched2tlx-perf-testing/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "sched2tlx-perf-testing", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add facebookexperimental/triton --skill sched2tlx-perf-testing -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install facebookexperimental/triton sched2tlx-perf-testing --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/facebookexperimental/triton.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.claude/skills/sched2tlx-perf-testing .opencode/skills/sched2tlx-perf-testing && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "sched2tlx-perf-testing" agent skill from https://github.com/facebookexperimental/triton/tree/main/.claude/skills/sched2tlx-perf-testing into .opencode/skills/sched2tlx-perf-testing/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "sched2tlx-perf-testing", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
sched2tlx-perf-testingRun the sched2tlx perf/correctness harness over the modulo-scheduling example corpus (case1-9: GEMM, persistent GEMM, FA fwd/bwd, addmm+bias, LayerNorm, wgrad+bias, multiphase GEMM, scaledmm).
Sched2tlx Perf Testing is an agent skill from facebookexperimental/triton, published by the product's own GitHub organization. Run the sched2tlx perf/correctness harness over the modulo-scheduling example corpus (case1-9: GEMM, persistent GEMM, FA fwd/bwd, addmm+bias, LayerNorm, wgrad+bias, multiphase GEMM, scaledmm). Use when the user asks to benchmark generated-vs-handwritten kernels, check corpus correctness, compare emitter revisions, or regenerate schedulegraph.json fixtures. Never run perf unless explicitly asked.
Its SKILL.md is about 1.9k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Productivity & Automation. It works with C++. The repository describes itself as: Github mirror of trition-lang/triton repo. The licence is MIT.
2 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 6f3dd70. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
makeuvpythonFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use uv, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Sched2tlx Perf Testing loads about 1.9k tokens when it runs. Until then it costs about 106 tokens; SKILL.md has 778 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from facebookexperimental/triton at commit 6f3dd70, republished under its MIT licence (© facebookexperimental). 778 words, ~1,917 tokens.
.claude/skills/sched2tlx-perf-testing/SKILL.md (or your agent's skills folder).Never run performance tests unless the user explicitly asks.
Harness: third_party/tlx/tools/sched2tlx/examples/testing/perf_regression/perf_harness.py
Corpus: third_party/tlx/tools/sched2tlx/examples/case*/
Before running any performance test for a C++ scheduler change, rebuild
Triton, then regenerate both schedule_graph.json and generated.py for
every case being tested. Do not benchmark stale fixtures.
Use one build path consistently for compilation, fixture regeneration, correctness, and timing:
buck2 is available, use buck2 run to compile and
run every performance test, including compare and per-case benchmarks.
Load the running-with-buck skill for the required working directory,
@mode/opt, beta-Triton modifier, GPU architecture, and CUDA flags. Do not
silently fall back to the repo venv when Buck is available. If the requested
benchmark has no runnable Buck target, report the missing target instead.buck2 is unavailable, build and run all
performance tests with $REPO/.venv. The login shell's lmod modules break
both build and runtime, so prefix every Python and triton-opt invocation
with env -u LD_LIBRARY_PATH. Ignore the lua/posix noise every command
prints. Use one of the repo-venv rebuild methods below.buck2 build rebuilds the selected target
and its changed C++ dependencies incrementally; buck2 run does the same
before running it. Follow the running-with-buck skill and never guess a
target name. A Buck rebuild does not update source-tree schedule_graph.json
or generated.py unless the selected target explicitly regenerates them.env -u LD_LIBRARY_PATH \
PATH="$REPO/.venv/bin:$HOME/.local/bin:/usr/local/bin:/usr/bin:/bin" \
VIRTUAL_ENV="$REPO/.venv" PYTHON="$REPO/.venv/bin/python" makeenv -u LD_LIBRARY_PATH \
PATH="$REPO/.venv/bin:$HOME/.local/bin:/usr/local/bin:/usr/bin:/bin" \
VIRTUAL_ENV="$REPO/.venv" CC=/usr/bin/gcc CXX=/usr/bin/g++ MAX_JOBS=14 \
uv pip install -e . --no-build-isolationmake dev-install-triton is the Makefile wrapper for the same editable
installation flow when PYTHON points to $REPO/.venv/bin/python.Regardless of the rebuild path, regenerate schedule_graph.json and
generated.py for every selected case before performance testing a scheduler
change.
When Buck is available, run the applicable performance-runner target from
fbsource/fbcode, following the running-with-buck skill, and pass
compare, --rev, and --cases as program arguments after --.
Only when Buck is unavailable, run from
examples/testing/perf_regression/:
env -u LD_LIBRARY_PATH $REPO/.venv/bin/python perf_harness.py compare \
[--rev origin/main] [--cases case7_wgrad_bias,case9_scaled_mm/blockwise]One row per case, four columns:
| column | meaning |
|---|---|
case | case dir relative to examples/ (nested variants like case9_scaled_mm/blockwise included) |
main (gen/hw) | per-shape gen/handwritten throughput ratios for --rev's committed generated.py (default origin/main) |
branch (gen/hw) | the same for the working tree's generated.py |
improvement | per-shape % change of the branch's GENERATED-kernel throughput vs --rev's (positive = branch faster) |
Semantics:
bench_spec.py files are discovered RECURSIVELY under examples/; top-level
case*/ dirs without any spec are listed as (no bench_spec), never
silently dropped. All of case1–case9 currently have specs.generated.py, not schedule_graph.json:
JSON op ids are pointer-derived and unstable across regenerations, while
byte-identical generated source means the kernels are identical. When the
revision's and working tree's generated.py are byte-identical, benchmark
the revision's gen/hw result once for the left column, skip a duplicate
working-tree benchmark, show unchanged in the branch column, and show -
for improvement.handwritten.py) show raw generated TFLOPS instead of a gen/hw ratio; the
improvement column still works. case9_scaled_mm/blockwise wires hw_call
to handwritten.blackwell_scaled_mm_ws, so it reports gen/hw ratios.(error: ...) cell instead of crashing the
table.Deep-dive per-case scripts (outside the harness): case4
perf_generated.py (gen vs no-WS vs handwritten WS) and run_generated.py
(all three gradients); case8 bench_general.py (all three outputs +
pool-vs-sum A/B); any case's run_*.py runner for correctness-only
(case8's is run_triple_gemm_nows.py).
The corpus fixtures (schedule_graph.json and the committed generated.py)
are produced by the Modulo Scheduling pass.
When Buck is available, regenerate fixtures with the Buck-built beta
triton-opt and the applicable sched2tlx Buck runner, following the same
build-path rule above. The commands below are only for the no-Buck venv
fallback:
TRITON_MODULO_DUMP_SCHEDULE=<case>/schedule_graph.json \
build/cmake.*/bin/triton-opt -allow-unregistered-dialect \
--nvgpu-modulo-schedule <case>/<kernel>_pre_modulo.ttgir -o /dev/null
env -u LD_LIBRARY_PATH PYTHONPATH=third_party/tlx/tools/sched2tlx \
$REPO/.venv/bin/python -m sched2tlx <case>/schedule_graph.json -o <case>/generated.pyJSON op ids are pointer-derived and never byte-stable — regen always churns
schedule_graph.json; the meaningful diff and benchmark-identity signal is
generated.py.
Known: case3 may need TRITON_MODULO_SELECT_VARIANT=2; case2 fixtures are
ancient (fresh dumps differ, pre-existing); case8's committed generated.py
predates the emitter's multiphase support landing (regen produces a
single-phase kernel — don't "refresh" it casually).
triton.testing.do_bench for every timing run. Configure a nonzero
warmup before measurement; measured iterations must clear L2 before each
invocation. Do not add ad-hoc CUDA-event timing loops.nvidia-smi first; if a run hangs for minutes, run
third_party/tlx/killgpu.sh.© facebookexperimental, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in .claude/skills/sched2tlx-perf-testing of facebookexperimental/triton.
Open the folder on GitHubat commit 6f3dd70
Sched2tlx Perf Testing next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Sched2tlx Perf Testing this skillfacebookexperimental/triton | 201 | — | ~1.9k | Automated safety check: Pass | MIT | |
| File Organization And Structurekitchen-engineer42/pdf2skills | 134 | — | ~500 | Automated safety check: Pass | None | |
| Update Milvus SDK Docsmilvus-io/web-content | 138 | — | ~13k | Automated safety check: Pass | Apache-2.0 | |
| Dependency Watchtelegramdesktop/tdesktop | 33k | 1 repos | ~2.2k | Automated safety check: Pass | GPL-3.0 | |
| Process Inboxtelegramdesktop/tdesktop | 33k | 1 repos | ~5.4k | Automated safety check: Pass | GPL-3.0 | |
| Feishu Docopenclaw/openclaw | 392k | — | ~516 | Automated safety check: Pass | MIT |
kitchen-engineer42/pdf2skills
Organize routines within files using blank line separation, consider alphabetical ordering when appropriate, and follow C++ standard file structure.
milvus-io/web-content
Update the Milvus SDK API reference documentation under APIReference/ in the web-content repository so it reflects a new SDK release, using the SDK repository's git tags as ground truth.
telegramdesktop/tdesktop
Audit Telegram Desktop dependencies on freshly fetched origin/dev for releases and security fixes, including upstream lag and backport candidates in patched forks.
telegramdesktop/tdesktop
Process the local ignored ai-tdesktop inbox into durable, independently testable Telegram Desktop task records while task execution worktrees remain active.
openclaw/openclaw
Feishu document read/write workflows. An agent skill from openclaw/openclaw.
PaddlePaddle/Paddle
A skill your agent uses when needing to compile, rebuild, or install Paddle from source after code changes.
facebookexperimental/triton
Collect, validate, package, and inspect rocprofv3 Advanced Thread Trace bundles for AMD GPU kernels.
facebookexperimental/triton
Design and run Triton TTGIR debugging ablations using iroverride.
facebookexperimental/triton
Execute the TLX Kernel Optimization Agent CLI on a Triton or TLX kernel.
facebookexperimental/triton
Run NVIDIA compute-sanitizer (memcheck, racecheck, initcheck, synccheck) against a Triton/TLX kernel to find runtime memory and synchronization bugs.
facebookexperimental/triton
Recover from GPU-busy / GPU-unavailable failures. An agent skill from facebookexperimental/triton.
facebookexperimental/triton
Debug Triton compilation by dumping IR at each stage (TTIR, TTGIR, LLVM, PTX).
Works with
Categories
Run the sched2tlx perf/correctness harness over the modulo-scheduling example corpus (case1-9: GEMM, persistent GEMM, FA fwd/bwd, addmm+bias, LayerNorm, wgrad+bias, multiphase GEMM, scaledmm). Sched2tlx Perf Testing is an agent skill from facebookexperimental/triton, published by the product's own GitHub organization. Run the sched2tlx perf/correctness harness over the modulo-scheduling example corpus (case1-9: GEMM, persistent GEMM, FA fwd/bwd, addmm+bias, LayerNorm, wgrad+bias, multiphase GEMM, scaledmm).
Sched2tlx Perf Testing fits situations like: the user asks to benchmark generated-vs-handwritten kernels; check corpus correctness; compare emitter revisions; regenerate schedulegraph.json fixtures.
Run `npx skills add facebookexperimental/triton --skill sched2tlx-perf-testing -a claude-code`. Or copy the skill folder (.claude/skills/sched2tlx-perf-testing in facebookexperimental/triton) into .claude/skills/sched2tlx-perf-testing in your project. Claude Code loads it when a task matches its description.
Run `npx skills add facebookexperimental/triton --skill sched2tlx-perf-testing -a codex`. Or copy the skill folder (.claude/skills/sched2tlx-perf-testing in facebookexperimental/triton) into .agents/skills/sched2tlx-perf-testing in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add facebookexperimental/triton --skill sched2tlx-perf-testing -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/sched2tlx-perf-testing, .gemini/skills/sched2tlx-perf-testing, .github/skills/sched2tlx-perf-testing and .opencode/skills/sched2tlx-perf-testing in your project.
Going by SKILL.md and its folder, Sched2tlx Perf Testing needs the command-line tools its instructions call (make, uv and python). Our summary lists: Python 3.
SKILL.md contains no URLs. Its commands use uv, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Sched2tlx Perf Testing is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.9k tokens (SKILL.md is roughly 7.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Sched2tlx Perf Testing: File Organization And Structure (kitchen-engineer42/pdf2skills, 134 stars), Update Milvus SDK Docs (milvus-io/web-content, 138 stars), Dependency Watch (telegramdesktop/tdesktop, 33k stars) and Process Inbox (telegramdesktop/tdesktop, 33k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
facebookexperimental (a GitHub organization, an official publisher) maintains it in facebookexperimental/triton, which has 201 GitHub stars. The repository holds 18 skills in this directory. The repository was last updated on October 10, 2026.
Source: facebookexperimental/triton on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.