Agent skill

Slime User

by yzlnew in yzlnew/infra-skills

Guide for using SLIME (LLM post-training framework for RL Scaling).

No licenceAuto-check passedAI & LLM Engineering

Install Slime User

skills CLI
$ npx skills add yzlnew/infra-skills --skill slime-user -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install yzlnew/infra-skills slime-user --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/yzlnew/infra-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/slime-user .claude/skills/slime-user && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
slime-user
GitHub stars
149
Token cost
~3.2k tokens
SKILL.md length
740 words
Files
4 (incl. references)
Skills in repo
8
Repo updated
First seen
Licence
None found

At a glance

Guide for using SLIME (LLM post-training framework for RL Scaling).

  • Works in 5 steps: Standard Single-Turn Training → Multi-Turn Tool Calling → Dynamic Sampling (DAPO-style) → …
  • Working with SLIME for reinforcement learning training of language models
  • SKILL.md covers Quick Start Workflow, Documentation Navigation, Core Concepts and Parameter Quick Reference, plus 6 more sections
  • Calls hf, python and bash

What it does

Slime User is an agent skill from yzlnew/infra-skills. Guide for using SLIME (LLM post-training framework for RL Scaling). Use when working with SLIME for reinforcement learning training of language models, including setup, configuration, training execution, multi-turn interactions, custom reward models, tool calling scenarios, or troubleshooting SLIME workflows. Covers GRPO, GSPO, PPO, Reinforce++, multi-agent RL, VLM training, FSDP/Megatron backends, SGLang integration, dynamic sampling, and custom generation functions.

Its SKILL.md is about 3.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 4 other files, including reference files (for example `references/doc_navigation.md`, `references/examples_reference.md` and `references/source_code_reference.md`).

It sits in AI & LLM Engineering, covering Reinforcement learning, Deep learning and Structured output and tool calling. It works with NVIDIA AI Platform and SGLang. The repository describes itself as: A collection of specialized agent skills for AI infrastructure development, enabling Claude Code to write, optimize, and debug high-performance systems.

When your agent uses it

  • Working with SLIME for reinforcement learning training of language models
  • Including setup
  • Training execution
  • Multi-turn interactions

Example prompts

  • “/slime-user”

Requirements

  • Python 3
  • Docker

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. Standard Single-Turn Training
  2. Multi-Turn Tool Calling
  3. Dynamic Sampling (DAPO-style)
  4. FSDP Backend (No Weight Conversion)
  5. Multi-Node Training

What it can do on your machine

Read from SKILL.md and the folder at commit f3a8d7d. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • hf
    • python
    • bash
    • docker

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • github.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Slime User loads about 3.2k tokens when it runs, and up to ~9.2k if it reads all its reference files. Until then it costs about 121 tokens; SKILL.md has 740 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~121
When it runs · the whole SKILL.md, loaded when a task matches
~3.2k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~9.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 740 words (~3,216 tokens).

“SLIME is an LLM post-training framework for RL Scaling developed by THUDM. It supports various RL algorithms (GRPO, GSPO, PPO, Reinforce++), multiple training backends (Megatron, FSDP), and advanced features like multi-turn interactions, tool calling, and dynamic sampling.”

— opening of SKILL.md by yzlnew
name
slime-user

Read the full SKILL.md on GitHub

Files

SKILL.md and 3 other files (references) in slime-user of yzlnew/infra-skills.

  • SKILL.md
  • references/doc_navigation.md
  • references/examples_reference.md
  • references/source_code_reference.md

Open the folder on GitHubat commit f3a8d7d

Compare with similar skills

Slime User next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Slime User compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Slime User this skillyzlnew/infra-skills149—~3.2kAutomated safety check: PassNone
Graphsignalgraphsignal/graphsignal257—~6.2kAutomated safety check: PassApache-2.0
SGLang Structured ServingOrchestra-Research/AI-Research-SKILLs13k2 repos~2.9kAutomated safety check: PassMIT
Spark Environment Setupwshobson/agents40k—~2kAutomated safety check: PassMIT
Serving Systemsuw-syfi/vibesys103—~2.9kAutomated safety check: PassMIT
Verl Quickstartascend-ai-coding/awesome-ascend-skills174—~592Automated safety check: PassNone

Similar skills

  • Graphsignal

    graphsignal/graphsignal

    Profile AI inference workloads (vLLM, SGLang, TensorRT-LLM, PyTorch, any GPU application) with the Graphsignal profiler and read the results from its local /signals JSON endpoint.

    257 GitHub stars~6.2k tokensUpdated 11 days ago
    AI & LLM EngineeringAuto-check passed
  • SGLang Structured Serving

    Orchestra-Research/AI-Research-SKILLs

    Covers serving LLMs with SGLang, whose RadixAttention reuses cached prefixes, and constraining output to JSON, regex or grammar for agent and tool-calling workloads.

    13k GitHub starsUsed in 2 repos~2.9k tokens
    AI & LLM EngineeringAuto-check passed
  • Set up a working ML training/inference environment on NVIDIA DGX Spark (GB10, aarch64, CUDA 13).

    40k GitHub stars~2k tokensUpdated 4 days ago
    AI & LLM EngineeringAuto-check passed
  • Serving Systems

    uw-syfi/vibesys

    LLM and multimodal serving systems. An agent skill from uw-syfi/vibesys.

    103 GitHub stars~2.9k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Verl Quickstart

    ascend-ai-coding/awesome-ascend-skills

    Generates an executable, end-to-end VERL reinforcement learning quickstart runbook for Ascend/NPU (docker image, dataset preprocessing, model setup, mainppo training, and examples/run.sh flow).

    174 GitHub stars~592 tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • slime RL Post-Training

    Orchestra-Research/AI-Research-SKILLs

    Guides reinforcement-learning post-training of LLMs with slime, which pairs Megatron-LM training with SGLang rollouts, including GRPO runs on GLM, Qwen3 and Llama 3 models.

    13k GitHub starsUsed in 4 repos~2.8k tokens
    AI & LLM EngineeringAuto-check passed

More from yzlnew/infra-skills

All 8 skills in this repo
  • Hf Architecture Tikz

    yzlnew/infra-skills

    Draw Sebastian-Raschka-gallery-style TikZ architecture diagrams for any HuggingFace decoder-only LLM, with per-block parameter formulas and concrete numbers.

    149 GitHub stars~2k tokensUpdated 3 mo ago
    Auto-check passed
  • Megatron Memory Estimator

    yzlnew/infra-skills

    Estimate GPU memory usage for Megatron-based MoE (Mixture of Experts) and dense models.

    149 GitHub stars~2.2k tokensUpdated 3 mo ago
    Auto-check passed
  • HTML Flowchart Anthropic

    yzlnew/infra-skills

    Create and revise pure HTML/CSS flowcharts using an Anthropic-inspired design language.

    149 GitHub stars~1.5k tokensUpdated 3 mo ago
    Auto-check passed
  • Openai Dotcom Viz

    yzlnew/infra-skills

    Build figures in OpenAI's blog / research / system-card "dotcom" visual style — both (a) bar charts (monochrome bars with a darker same-hue stroke, rounded corners, a black y-axis with outward ticks…

    149 GitHub stars~1.3k tokensUpdated 3 mo ago
    Auto-check passed
  • Tilelang Developer

    yzlnew/infra-skills

    Write, optimize, and debug high-performance AI compute kernels using TileLang (a Python DSL for GPU programming).

    149 GitHub stars~2.4k tokensUpdated 3 mo ago
    Auto-check passed
  • Material You Slides

    yzlnew/infra-skills

    Create presentation slides using Material You (Material Design 3) style.

    149 GitHub stars~3.7k tokensUpdated 3 mo ago
    Auto-check passed

Questions about Slime User

What does Slime User do?

Guide for using SLIME (LLM post-training framework for RL Scaling). Slime User is an agent skill from yzlnew/infra-skills. Guide for using SLIME (LLM post-training framework for RL Scaling).

When should I use Slime User?

Slime User fits situations like: working with SLIME for reinforcement learning training of language models; including setup; training execution; multi-turn interactions.

How do I install Slime User in Claude Code?

Run `npx skills add yzlnew/infra-skills --skill slime-user -a claude-code`. Or copy the skill folder (slime-user in yzlnew/infra-skills) into .claude/skills/slime-user in your project. Claude Code loads it when a task matches its description.

How do I install Slime User in Codex?

Run `npx skills add yzlnew/infra-skills --skill slime-user -a codex`. Or copy the skill folder (slime-user in yzlnew/infra-skills) into .agents/skills/slime-user in your project. Codex loads it when a task matches its description.

Can I use Slime User in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add yzlnew/infra-skills --skill slime-user -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/slime-user, .gemini/skills/slime-user, .github/skills/slime-user and .opencode/skills/slime-user in your project.

What does Slime User need to run?

Going by SKILL.md and its folder, Slime User needs the command-line tools its instructions call (hf, python, bash and docker). Our summary lists: Python 3; Docker.

Does Slime User access the network?

SKILL.md names 1 domain. As links in the text: github.com. This is read from the text; nothing was executed.

Is Slime User safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Slime User use?

No licence was found for Slime User or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Slime User use?

About 3.2k tokens (SKILL.md is roughly 13k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 6k tokens, read only when the agent opens those files.

What are the alternatives to Slime User?

Skills that share tags, products or a category with Slime User: Graphsignal (graphsignal/graphsignal, 257 stars), SGLang Structured Serving (Orchestra-Research/AI-Research-SKILLs, 13k stars), Spark Environment Setup (wshobson/agents, 40k stars) and Serving Systems (uw-syfi/vibesys, 103 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Slime User?

yzlnew (a GitHub user) maintains it in yzlnew/infra-skills, which has 149 GitHub stars. The repository holds 8 skills in this directory. The repository was last updated on July 9, 2026.

Source: yzlnew/infra-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.