Search

SGLang · Reinforcement learning

3 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Guides reinforcement-learning post-training of LLMs with slime, which pairs Megatron-LM training with SGLang rollouts, including GRPO runs on GLM, Qwen3 and Llama 3 models.

Orchestra-Research/AI-Research-SKILLs13k4 repos~2.8kAutomated safety check: PassMIT3 mo ago
2

Trains LLMs with reinforcement learning using verl, from ByteDance's Seed team, with GRPO, PPO and other algorithms and swappable training and rollout backends.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2.4kAutomated safety check: PassMIT3 mo ago
3

Guide for using SLIME (LLM post-training framework for RL Scaling).

yzlnew/infra-skills149—~3.2kAutomated safety check: PassNo licence3 mo ago