Search

Docker · Fine-tuning

12 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Integrate a benchmark or custom environment into SAfactory using fixed adapter templates and local contract tests, optionally run Docker/RJob evaluation, or prepare GRPO/RL training.

AI45Lab/SAfactory236—~1.8kAutomated safety check: PassNo licence17 days ago
2

Set up the NVIDIA "Build an Agent" DevX workshop as a working JupyterLab environment from INSIDE a locked-down OpenShell/NemoClaw sandbox, and hand the user the token URL + access commands.

brevdev/workshop-build-an-agent146—~5.2kAutomated safety check: PassApache-2.03 days ago
3

Guides reinforcement-learning post-training of LLMs with slime, which pairs Megatron-LM training with SGLang rollouts, including GRPO runs on GLM, Qwen3 and Llama 3 models.

Orchestra-Research/AI-Research-SKILLs13k4 repos~2.8kAutomated safety check: PassMIT3 mo ago
4

A skill your agent uses when designing, reviewing, or implementing OpenEnv-style environment interfaces for agentic RL with TRL, including reset/step/state contracts, tasksets, Docker or…

burtenshaw/training-agents153—~381Automated safety check: PassApache-2.028 days ago
5

Build an end-to-end UAV object detection and telemetry overlay application on Intel hardware using DL Streamer Pipeline Server with MAVLink telemetry.

open-edge-platform/edge-ai-suites140—~2.6kAutomated safety check: NotesApache-2.0yesterday
6

Run GPU workloads on Modal — training, fine-tuning, inference, batch processing.

AI4Scientist/nano-scientist1283 repos~3.1kAutomated safety check: NotesNo licence4 mo ago
7

Train a character/identity LoRA locally on FLUX.1-dev via the comfyui-mcp train tools (GPU Docker + ostris ai-toolkit).

artokun/comfyui-mcp803—~1.3kAutomated safety check: PassMIT6 days ago
8
8.Kermt MonitorOfficial

Check progress for a detached KERMT run (pretrain, finetune, or any kermtrundetached invocation).

NVIDIA/skills3.6k1 repo~1.8kAutomated safety check: PassApache-2.0yesterday
9

DINOv3 continual self-supervised pre-training. An agent skill from NVIDIA/skills.

NVIDIA/skills3.6k—~2.2kAutomated safety check: NotesApache-2.0yesterday
10

A skill your agent uses when running GPU compute on RunPod and deciding between Pods (hourly, always-on) and Serverless (per-second, autoscaling) for training, fine-tuning or inference — serverless…

ericrisco/rsc-harness180—~2.8kAutomated safety check: PassMITyesterday
11

Verl 分布式训练服务一键拉起与配置。触发场景:(1) 用户要启动 Verl 训练任务或部署 RLHF/DAPO 训练环境 (2) 在 NPU 集群上拉起 Verl 训练容器 (3) 配置 Ray 集群和 SwanLab 监控 (4) 根据 7 位二进制掩码灵活配置加速特性。支持 Qwen3-8B 等 Megatron 模型的 DAPO 训练全流程。

ascend-ai-coding/awesome-ascend-skills174—~2kAutomated safety check: PassNo licenceyesterday
12

Generates an executable, end-to-end VERL reinforcement learning quickstart runbook for Ascend/NPU (docker image, dataset preprocessing, model setup, mainppo training, and examples/run.sh flow).

ascend-ai-coding/awesome-ascend-skills174—~592Automated safety check: PassNo licenceyesterday