Search
AI & LLM Engineering · DeepSeek · By Orchestra-Research
5 skills found.
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Provides PyTorch-native distributed LLM pretraining using torchtitan with 4D parallelism (FSDP2, TP, PP, CP). | Orchestra-Research/ | 13k | 4 repos | ~2.2k | Automated safety check: Pass | MIT | 3 mo ago |
| 2 | Walks through aligning language models with SimPO, a reference-free preference optimization method, using accelerate configs for Mistral 7B, Llama 3 8B and math-focused models. | Orchestra-Research/ | 13k | 4 repos | ~1.5k | Automated safety check: Pass | MIT | 3 mo ago |
| 3 | Guides reinforcement-learning post-training of LLMs with slime, which pairs Megatron-LM training with SGLang rollouts, including GRPO runs on GLM, Qwen3 and Llama 3 models. | Orchestra-Research/ | 13k | 4 repos | ~2.8k | Automated safety check: Pass | MIT | 3 mo ago |
| 4 | Provides guidance for enterprise-grade RL training using miles, a production-ready fork of slime. | Orchestra-Research/ | 13k | 2 repos | ~2.2k | Automated safety check: Pass | MIT | 3 mo ago |
| 5 | Train Mixture of Experts (MoE) models using DeepSpeed or HuggingFace. | Orchestra-Research/ | 13k | 2 repos | ~3.7k | Automated safety check: Pass | MIT | 3 mo ago |