Search

DeepSeek · Orchestra-Research/AI-Research-SKILLs

5 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Provides PyTorch-native distributed LLM pretraining using torchtitan with 4D parallelism (FSDP2, TP, PP, CP).

Orchestra-Research/AI-Research-SKILLs13k4 repos~2.2kAutomated safety check: PassMIT3 mo ago
2

Walks through aligning language models with SimPO, a reference-free preference optimization method, using accelerate configs for Mistral 7B, Llama 3 8B and math-focused models.

Orchestra-Research/AI-Research-SKILLs13k4 repos~1.5kAutomated safety check: PassMIT3 mo ago
3

Guides reinforcement-learning post-training of LLMs with slime, which pairs Megatron-LM training with SGLang rollouts, including GRPO runs on GLM, Qwen3 and Llama 3 models.

Orchestra-Research/AI-Research-SKILLs13k4 repos~2.8kAutomated safety check: PassMIT3 mo ago
4

Provides guidance for enterprise-grade RL training using miles, a production-ready fork of slime.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2.2kAutomated safety check: PassMIT3 mo ago
5

Train Mixture of Experts (MoE) models using DeepSpeed or HuggingFace.

Orchestra-Research/AI-Research-SKILLs13k2 repos~3.7kAutomated safety check: PassMIT3 mo ago