Search

Docker · Reinforcement learning

8 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Builds a Verifiers (PrimeIntellect) variant of an RL environment.

adithya-s-k/FineEnvs4611 repo~2.3kAutomated safety check: PassApache-2.03 days ago
2

Integrate a benchmark or custom environment into SAfactory using fixed adapter templates and local contract tests, optionally run Docker/RJob evaluation, or prepare GRPO/RL training.

AI45Lab/SAfactory236—~1.8kAutomated safety check: PassNo licence17 days ago
3

Builds a NeMo Gym (NVIDIA) variant of an RL environment. An agent skill from adithya-s-k/FineEnvs.

adithya-s-k/FineEnvs461—~2.1kAutomated safety check: PassApache-2.03 days ago
4

Builds an OpenEnv (Hugging Face) variant of an RL environment.

adithya-s-k/FineEnvs461—~2.4kAutomated safety check: PassApache-2.03 days ago
5

Guides reinforcement-learning post-training of LLMs with slime, which pairs Megatron-LM training with SGLang rollouts, including GRPO runs on GLM, Qwen3 and Llama 3 models.

Orchestra-Research/AI-Research-SKILLs13k4 repos~2.8kAutomated safety check: PassMIT3 mo ago
6

Builds an Open Reward Standard (ORS) variant of an RL environment using the official openreward Python package.

adithya-s-k/FineEnvs461—~2.3kAutomated safety check: NotesApache-2.03 days ago
7

Deploy the article to a Hugging Face Space. An agent skill from adithya-s-k/FineEnvs.

adithya-s-k/FineEnvs461—~864Automated safety check: PassApache-2.03 days ago
8

Generates an executable, end-to-end VERL reinforcement learning quickstart runbook for Ascend/NPU (docker image, dataset preprocessing, model setup, mainppo training, and examples/run.sh flow).

ascend-ai-coding/awesome-ascend-skills174—~592Automated safety check: PassNo licenceyesterday