Topic · AI & LLM Engineering
Best reinforcement learning skills, page 2
Reinforcement learning skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 49 | Applies the reasoning of Pieter Abbeel, robotics and reinforcement learning expert, UC Berkeley professor, and co-founder of Covariant. | K-Dense-AI/ | 282 | — | ~1.5k | Automated safety check: Pass | MIT | 1 mo ago |
| 50 | Reach for this skill whenever you are discussing reinforcement learning, agentic AI systems, AI alignment, continual learning, or the philosophical limits of large language models. | K-Dense-AI/ | 282 | — | ~1.8k | Automated safety check: Pass | MIT | 1 mo ago |
| 51 | Applies the reasoning of Stuart Russell, AI safety expert, UC Berkeley professor, and co-author of 'Artificial Intelligence: A Modern Approach'. | K-Dense-AI/ | 282 | — | ~1.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 52 | Train and optimize AI agents using Microsoft's Agent Lightning framework with reinforcement learning. | coco-research/ | 482 | — | ~2k | Automated safety check: Pass | Unknown | today |
| 53 | 53.Grpo Reference for the GRPO (Group Relative Policy Optimization) algorithm. | benchflow-ai/ | 1.8k | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 54 | Diagnostic guide for RL-based post-training of language models (GRPO, PPO, REINFORCE, DPO). | benchflow-ai/ | 1.8k | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 55 | Build practice environments and evaluation gyms where agents can try, fail, and learn from feedback. | LearnPrompt/ | 109 | — | ~2.1k | Automated safety check: Pass | MIT | 2 mo ago |
| 56 | Brev instance operating guidance for NeMo-RL agents working in /home/ubuntu/RL with limited workspace disk, a larger /ephemeral volume, and optional /home/ubuntu/RL/.env secrets. | NVIDIA/ | 3.5k | — | ~1.3k | Automated safety check: Notes | Apache-2.0 | today |
| 57 | Generates A/B test plans and optimization checklists for your App Store product page — icon, screenshots, and app previews. | gustavscirulis/ | 117 | 1 repo | ~1.9k | Automated safety check: Pass | Unknown | 5 mo ago |
| 58 | 58.Trl Reference for the TRL (Transformer Reinforcement Learning) library codebase. | benchflow-ai/ | 1.8k | — | ~989 | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 59 | A skill your agent uses when implementing staking contracts, reward distribution systems, or yield farming. | ccashwell/ | 131 | — | ~1.7k | Automated safety check: Pass | MIT | 8 days ago |
| 60 | Reinforcement learning fundamentals, algorithms, and research | wentorai/ | 298 | 1 repo | ~2.4k | Automated safety check: Pass | MIT | 3 mo ago |
| 61 | Vectorized multi-agent reinforcement learning simulator. An agent skill from wentorai/research-plugins. | wentorai/ | 298 | 1 repo | ~960 | Automated safety check: Pass | MIT | 3 mo ago |
| 62 | Build and review OpenRLHF supervised/preference training plans for SFT, reward models, DPO, IPO, and cDPO. | VectorSpaceLab/ | 328 | — | ~1k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 63 | 63.Torchrl Use TorchRL for TensorDict-first reinforcement-learning environments, collectors, replay buffers, modules, objectives, LLM/RLHF/VLA workflows, services, rendering, and maintainer-safe repository… | VectorSpaceLab/ | 328 | — | ~1.5k | Automated safety check: Pass | MIT | 1 mo ago |
| 64 | Generates an executable, end-to-end VERL reinforcement learning quickstart runbook for Ascend/NPU (docker image, dataset preprocessing, model setup, mainppo training, and examples/run.sh flow). | ascend-ai-coding/ | 174 | — | ~592 | Automated safety check: Pass | No licence | today |
| 65 | A skill your agent uses when positioning an AAMAS submission against multiagent, game-theory, and reinforcement-learning literature spread across AAMAS, AAAI, IJCAI, NeurIPS, ICML, EC, and JAAMAS… | brycewang-stanford/ | 1.2k | — | ~970 | Automated safety check: Pass | MIT | 10 days ago |
| 66 | Your AI research and engineering brain trust. An agent skill from majiayu000/claude-skill-registry. | majiayu000/ | 666 | 1 repo | ~3.7k | Automated safety check: Pass | MIT | today |
Explore related skills
More topics in AI & LLM Engineering
- Building AI agents563
- Deep learning415
- Embeddings386
- LLM inference and serving372
- Prompt engineering360
- Retrieval-augmented generation358
- Fine-tuning309
- LLM evaluation308
- Speech recognition and synthesis308
- Structured output and tool calling276
- LLM cost and token optimization259
- LLM API integration255
- Model routing and gateways255
- LLM observability240
- LLM guardrails221
- Computer vision203
- Model hubs and datasets180
- GPU and accelerator computing176
- Diffusion and image models166
- Natural language processing131
- AI interpretability23