Topic · AI & LLM Engineering

Best reinforcement learning skills, page 2

Skills #49–66 of 66, ranked by score.

Reinforcement learning skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Reinforcement learning skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
49

Applies the reasoning of Pieter Abbeel, robotics and reinforcement learning expert, UC Berkeley professor, and co-founder of Covariant.

K-Dense-AI/mimeo282—~1.5kAutomated safety check: PassMIT1 mo ago
50

Reach for this skill whenever you are discussing reinforcement learning, agentic AI systems, AI alignment, continual learning, or the philosophical limits of large language models.

K-Dense-AI/mimeo282—~1.8kAutomated safety check: PassMIT1 mo ago
51

Applies the reasoning of Stuart Russell, AI safety expert, UC Berkeley professor, and co-author of 'Artificial Intelligence: A Modern Approach'.

K-Dense-AI/mimeo282—~1.6kAutomated safety check: PassMIT1 mo ago
52

Train and optimize AI agents using Microsoft's Agent Lightning framework with reinforcement learning.

coco-research/coco482—~2kAutomated safety check: PassUnknowntoday
53
53.Grpo

Reference for the GRPO (Group Relative Policy Optimization) algorithm.

benchflow-ai/skillsbench1.8k—~1.1kAutomated safety check: PassApache-2.02 mo ago
54

Diagnostic guide for RL-based post-training of language models (GRPO, PPO, REINFORCE, DPO).

benchflow-ai/skillsbench1.8k—~1.6kAutomated safety check: PassApache-2.02 mo ago
55

Build practice environments and evaluation gyms where agents can try, fail, and learn from feedback.

LearnPrompt/andrej-karpathy-skills109—~2.1kAutomated safety check: PassMIT2 mo ago
56

Brev instance operating guidance for NeMo-RL agents working in /home/ubuntu/RL with limited workspace disk, a larger /ephemeral volume, and optional /home/ubuntu/RL/.env secrets.

NVIDIA/skills3.5k—~1.3kAutomated safety check: NotesApache-2.0today
57

Generates A/B test plans and optimization checklists for your App Store product page — icon, screenshots, and app previews.

gustavscirulis/snapgrid1171 repo~1.9kAutomated safety check: PassUnknown5 mo ago
58
58.Trl

Reference for the TRL (Transformer Reinforcement Learning) library codebase.

benchflow-ai/skillsbench1.8k—~989Automated safety check: PassApache-2.02 mo ago
59

A skill your agent uses when implementing staking contracts, reward distribution systems, or yield farming.

ccashwell/evm-cortex131—~1.7kAutomated safety check: PassMIT8 days ago
60

Reinforcement learning fundamentals, algorithms, and research

wentorai/research-plugins2981 repo~2.4kAutomated safety check: PassMIT3 mo ago
61

Vectorized multi-agent reinforcement learning simulator. An agent skill from wentorai/research-plugins.

wentorai/research-plugins2981 repo~960Automated safety check: PassMIT3 mo ago
62

Build and review OpenRLHF supervised/preference training plans for SFT, reward models, DPO, IPO, and cDPO.

VectorSpaceLab/AREX-Skill328—~1kAutomated safety check: PassApache-2.01 mo ago
63

Use TorchRL for TensorDict-first reinforcement-learning environments, collectors, replay buffers, modules, objectives, LLM/RLHF/VLA workflows, services, rendering, and maintainer-safe repository…

VectorSpaceLab/AREX-Skill328—~1.5kAutomated safety check: PassMIT1 mo ago
64

Generates an executable, end-to-end VERL reinforcement learning quickstart runbook for Ascend/NPU (docker image, dataset preprocessing, model setup, mainppo training, and examples/run.sh flow).

ascend-ai-coding/awesome-ascend-skills174—~592Automated safety check: PassNo licencetoday
65

A skill your agent uses when positioning an AAMAS submission against multiagent, game-theory, and reinforcement-learning literature spread across AAMAS, AAAI, IJCAI, NeurIPS, ICML, EC, and JAAMAS…

brycewang-stanford/Awesome-Journal-Skills1.2k—~970Automated safety check: PassMIT10 days ago
66

Your AI research and engineering brain trust. An agent skill from majiayu000/claude-skill-registry.

majiayu000/claude-skill-registry6661 repo~3.7kAutomated safety check: PassMITtoday