Search

Qwen · Reinforcement learning

4 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Fix a GitHub issue on OpenPipe/ART and open a PR. An agent skill from OpenPipe/ART.

OpenPipe/ART11k—~840Automated safety check: NotesApache-2.0today
2

RL training reference for the ART framework. An agent skill from OpenPipe/ART.

OpenPipe/ART11k—~2.4kAutomated safety check: PassApache-2.0today
3

Guides reinforcement-learning post-training of LLMs with slime, which pairs Megatron-LM training with SGLang rollouts, including GRPO runs on GLM, Qwen3 and Llama 3 models.

Orchestra-Research/AI-Research-SKILLs13k4 repos~2.8kAutomated safety check: PassMIT3 mo ago
4

Provides guidance for enterprise-grade RL training using miles, a production-ready fork of slime.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2.2kAutomated safety check: PassMIT3 mo ago