Search
Docker · Fine-tuning
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Integrate a benchmark or custom environment into SAfactory using fixed adapter templates and local contract tests, optionally run Docker/RJob evaluation, or prepare GRPO/RL training. | AI45Lab/ | 236 | — | ~1.8k | Automated safety check: Pass | No licence | 17 days ago |
| 2 | Set up the NVIDIA "Build an Agent" DevX workshop as a working JupyterLab environment from INSIDE a locked-down OpenShell/NemoClaw sandbox, and hand the user the token URL + access commands. | brevdev/ | 146 | — | ~5.2k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 3 | Guides reinforcement-learning post-training of LLMs with slime, which pairs Megatron-LM training with SGLang rollouts, including GRPO runs on GLM, Qwen3 and Llama 3 models. | Orchestra-Research/ | 13k | 4 repos | ~2.8k | Automated safety check: Pass | MIT | 3 mo ago |
| 4 | A skill your agent uses when designing, reviewing, or implementing OpenEnv-style environment interfaces for agentic RL with TRL, including reset/step/state contracts, tasksets, Docker or… | burtenshaw/ | 153 | — | ~381 | Automated safety check: Pass | Apache-2.0 | 28 days ago |
| 5 | Build an end-to-end UAV object detection and telemetry overlay application on Intel hardware using DL Streamer Pipeline Server with MAVLink telemetry. | open-edge-platform/ | 140 | — | ~2.6k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 6 | Run GPU workloads on Modal — training, fine-tuning, inference, batch processing. | AI4Scientist/ | 128 | 3 repos | ~3.1k | Automated safety check: Notes | No licence | 4 mo ago |
| 7 | Train a character/identity LoRA locally on FLUX.1-dev via the comfyui-mcp train tools (GPU Docker + ostris ai-toolkit). | artokun/ | 803 | — | ~1.3k | Automated safety check: Pass | MIT | 6 days ago |
| 8 | Check progress for a detached KERMT run (pretrain, finetune, or any kermtrundetached invocation). | NVIDIA/ | 3.6k | 1 repo | ~1.8k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 9 | DINOv3 continual self-supervised pre-training. An agent skill from NVIDIA/skills. | NVIDIA/ | 3.6k | — | ~2.2k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 10 | 10.Runpod A skill your agent uses when running GPU compute on RunPod and deciding between Pods (hourly, always-on) and Serverless (per-second, autoscaling) for training, fine-tuning or inference — serverless… | ericrisco/ | 180 | — | ~2.8k | Automated safety check: Pass | MIT | yesterday |
| 11 | Verl 分布式训练服务一键拉起与配置。触发场景:(1) 用户要启动 Verl 训练任务或部署 RLHF/DAPO 训练环境 (2) 在 NPU 集群上拉起 Verl 训练容器 (3) 配置 Ray 集群和 SwanLab 监控 (4) 根据 7 位二进制掩码灵活配置加速特性。支持 Qwen3-8B 等 Megatron 模型的 DAPO 训练全流程。 | ascend-ai-coding/ | 174 | — | ~2k | Automated safety check: Pass | No licence | yesterday |
| 12 | Generates an executable, end-to-end VERL reinforcement learning quickstart runbook for Ascend/NPU (docker image, dataset preprocessing, model setup, mainppo training, and examples/run.sh flow). | ascend-ai-coding/ | 174 | — | ~592 | Automated safety check: Pass | No licence | yesterday |