Topic · AI & LLM Engineering
Best fine-tuning skills for Claude Code, Codex and other agents.
- skills
- 313
- official
- 50
Fine-tuning skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Parameter-efficient fine-tuning for LLMs using LoRA, QLoRA, and 25+ methods. | Orchestra-Research/ | 13k | 9 repos | ~3.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 2 | Trains or fine-tunes language and vision models with TRL or Unsloth on Hugging Face Jobs cloud GPUs, then converts the results to GGUF. | huggingface/ | 11k | 3 repos | ~7.2k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 3 | Routes a sentence-transformers training task to the right model type and required reference docs and example scripts, covering bi-encoders, rerankers, sparse and multi-vector models. | huggingface/ | 11k | 1 repo | ~2.6k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 4 | Fix a GitHub issue on OpenPipe/ART and open a PR. An agent skill from OpenPipe/ART. | OpenPipe/ | 11k | — | ~840 | Automated safety check: Notes | Apache-2.0 | yesterday |
| 5 | Validates dataset formatting and quality for SageMaker model fine-tuning (SFT, DPO, or RLVR). | awslabs/ | 912 | 2 repos | ~1.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 6 | 6.Train Rl RL training reference for the ART framework. An agent skill from OpenPipe/ART. | OpenPipe/ | 11k | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 7 | Prepare, validate, launch-plan, monitor, resume, and stop configurable Qwopus 27B reinforcement-learning workflows for GRPO or GSPO. | R6410418/ | 1.7k | — | ~830 | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 8 | Guides LLM fine-tuning with LoRA and QLoRA through Hugging Face PEFT, from dataset validation and training checks to adapter merging, quantization and deployment. | Jeffallan/ | 12k | 1 repo | ~1.7k | Automated safety check: Pass | MIT | 4 days ago |
| 9 | Generates code that transforms datasets between ML schemas for model training or evaluation. | awslabs/ | 912 | 2 repos | ~3.5k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 10 | 10.Train Sft SFT training reference for the ART framework. An agent skill from OpenPipe/ART. | OpenPipe/ | 11k | — | ~2.9k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 11 | Check Kiln's fine-tunable model list for deprecated or unsupported base models. | Kiln-AI/ | 5.2k | — | ~1.9k | Automated safety check: Notes | Unknown | today |
| 12 | Fine-tune LLMs using reinforcement learning with TRL - SFT for instruction tuning, DPO for preference alignment, PPO/GRPO for reward optimization, and reward model training. | Orchestra-Research/ | 13k | 7 repos | ~2.9k | Automated safety check: Pass | MIT | 3 mo ago |
| 13 | 13.Optim Agent A skill your agent uses when the user wants to optimize configurable system parameters against a measurable scalar objective, especially for model training, inference, quantitative strategies… | Optim-Agent/ | 800 | — | ~1.3k | Automated safety check: Pass | MIT | 1 mo ago |
| 14 | Integrate a benchmark or custom environment into SAfactory using fixed adapter templates and local contract tests, optionally run Docker/RJob evaluation, or prepare GRPO/RL training. | AI45Lab/ | 236 | — | ~1.8k | Automated safety check: Pass | No licence | 13 days ago |
| 15 | Trigger this skill when the user wants to train, fine-tune, or adapt Gemma models (e.g. | google-gemma/ | 1k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | today |
| 16 | Selects a fine-tuning technique (SFT, DPO, RLVR, or RLAIF) for the user's use case and validates it against the selected model's available recipes. | awslabs/ | 912 | 1 repo | ~604 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 17 | Trains and evaluates several WiFi-signal-based pose and sensing models, from unsupervised pose estimation to domain adaptation and publishing. | ruvnet/ | 97k | — | ~1.3k | Automated safety check: Notes | MIT | today |
| 18 | 18.Finetuning This skill should be used when picking or diagnosing a training move (SFT, LoRA, DPO/KTO/ORPO, RFT, GRPO/PPO/RLOO, RLHF), or when the user mentions fine-tuning, post-training, training recipe… | evo-hq/ | 1.5k | — | ~4.5k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 19 | 19.Add Pipeline Router for adding a diffusion or omni pipeline to verl-omni. | verl-project/ | 1.2k | — | ~1k | Automated safety check: Pass | Apache-2.0 | today |
| 20 | 20.Flux Image Generate images with FLUX models (Black Forest Labs) via inference.sh CLI. | danielmeppiel/ | 157 | 2 repos | ~778 | Automated safety check: Pass | Unknown | 3 mo ago |
| 21 | Add a cross-cutting decision pattern under src/nemotron/steps/patterns/. | NVIDIA-NeMo/ | 2.1k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 22 | Trains and fine-tunes object detection, image classification and SAM or SAM2 segmentation models on Hugging Face Jobs cloud GPUs and saves the results to the Hub. | huggingface/ | 11k | 1 repo | ~7.5k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 23 | Runs and configures the anomalib tiled-ensemble pipeline, which trains/evaluates one model per image tile and merges results (with optional seam smoothing) for high-resolution anomaly detection. | open-edge-platform/ | 6.2k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | today |
| 24 | 24.Setup Connect this agent to a running Guaardvark (self-hosted AI studio) and check what it can do right now. | guaardvark/ | 251 | 1 repo | ~1.2k | Automated safety check: Pass | MIT | today |
| 25 | Guide for adding a new reward scorer to verl-omni and wiring it into a run. | verl-project/ | 1.2k | — | ~648 | Automated safety check: Pass | Apache-2.0 | today |
| 26 | Guides building on the 0G Compute Network, a decentralized GPU marketplace for AI inference and fine-tuning, with SDK patterns and CLI commands. | internet-court/ | 6.4k | 1 repo | ~1.9k | Automated safety check: Pass | Unknown | 1 mo ago |
| 27 | 27.Deeplabcut Toolbox for markerless animal pose estimation with DeepLabCut. | NeuroAIHub/ | 1k | — | ~1.7k | Automated safety check: Pass | AGPL-3.0 | 5 days ago |
| 28 | Add a new step under src/nemotron/steps/<category/<stepid/ — manifest (step.toml), runner glue, configs, and per-step README.md. | NVIDIA-NeMo/ | 2.1k | — | ~1.7k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 29 | 29.Swift Mlx Lm MLX Swift LM - Run LLMs and VLMs on Apple Silicon using MLX. | kellyvv/ | 1.3k | — | ~3.7k | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 30 | Train or fine-tune language models with TRL or Unsloth on Hugging Face Jobs, including SFT, DPO, GRPO, reward models, and GGUF conversion. | waybarrios/ | 533 | — | ~3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 31 | Selectively monitor important long-running or resource-intensive commands with Haoleme by prefixing them with hao, so status, output, and completion notifications sync to the mobile app. | HaolemeApp/ | 157 | — | ~1.3k | Automated safety check: Pass | AGPL-3.0 | 1 mo ago |
| 32 | 32.Unimol A standardized CLI wrapper for Uni-Mol molecular ML workflows that handles representation extraction (embeddings), model training (regression/classification), and property prediction with built-in… | jinzhezenggroup/ | 148 | 1 repo | ~1.5k | Automated safety check: Pass | LGPL-3.0-or-later | yesterday |
| 33 | Expert guidance for fast fine-tuning with Unsloth - 2-5x faster training, 50-80% less memory, LoRA/QLoRA optimization | Orchestra-Research/ | 13k | 11 repos | ~577 | Automated safety check: Pass | MIT | 3 mo ago |
| 34 | Lilly community-research skill. An agent skill from ssaaffaakk/Lilly. | ssaaffaakk/ | 171 | — | ~1.4k | Automated safety check: Pass | MIT | yesterday |
| 35 | Audit imaging acquisition, reconstruction, series eligibility, quantitative transforms and protocol drift; not model training. | huang-sir1/ | 1.9k | — | ~2.8k | Automated safety check: Pass | Unknown | 16 days ago |
| 36 | Train custom LoRAs with ostris AI-Toolkit. An agent skill from artokun/comfyui-mcp. | artokun/ | 793 | — | ~2.7k | Automated safety check: Pass | MIT | 2 days ago |
| 37 | Quick-reference help for fine-tuning language models with Axolotl, covering YAML configs, FSDP, context parallelism, compressed saves and dataset formats. | Orchestra-Research/ | 13k | 9 repos | ~1.2k | Automated safety check: Pass | MIT | 3 mo ago |
| 38 | Run, configure, retry, and validate AReno SFT, DPO, GSPO, GRPO, PPO, and agentic training. | inclusionAI/ | 323 | — | ~782 | Automated safety check: Pass | Apache-2.0 | 13 days ago |
| 39 | A skill your agent uses whenever the user asks to generate, collect, inspect, or prepare early-experience training data (Implicit World Modeling or Self-Reflection, in the sense of arXiv:2510.08558)… | OSU-NLP-Group/ | 102 | — | ~4.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 40 | 40.Trl Sft A skill your agent uses when designing, implementing, reviewing, or debugging supervised fine-tuning with TRL SFTTrainer or trl sft, especially for agentic models trained on chat messages… | burtenshaw/ | 153 | — | ~685 | Automated safety check: Pass | Apache-2.0 | 24 days ago |
| 41 | Configure and launch SparkDiffusion sparse finetuning for Wan 2.1 or Wan 2.2. | AlibabaResearch/ | 490 | — | ~904 | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 42 | Train object-detection, image-classification, or SAM segmentation models on Hugging Face Jobs. | waybarrios/ | 533 | — | ~2.7k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 43 | Build, deploy, evaluate, optimize, fine-tune, and manage Microsoft Foundry agents, models, and resources end to end. | microsoft/ | 255 | 1 repo | ~6.7k | Automated safety check: Pass | MIT | today |
| 44 | Set up the NVIDIA "Build an Agent" DevX workshop as a working JupyterLab environment from INSIDE a locked-down OpenShell/NemoClaw sandbox, and hand the user the token URL + access commands. | brevdev/ | 143 | — | ~5.2k | Automated safety check: Pass | Apache-2.0 | today |
| 45 | Walks through aligning language models with SimPO, a reference-free preference optimization method, using accelerate configs for Mistral 7B, Llama 3 8B and math-focused models. | Orchestra-Research/ | 13k | 5 repos | ~1.5k | Automated safety check: Pass | MIT | 3 mo ago |
| 46 | Guides reinforcement-learning post-training of LLMs with slime, which pairs Megatron-LM training with SGLang rollouts, including GRPO runs on GLM, Qwen3 and Llama 3 models. | Orchestra-Research/ | 13k | 5 repos | ~2.8k | Automated safety check: Pass | MIT | 3 mo ago |
| 47 | Fine-tune or transfer-learn AlphaGenome-PyTorch on custom genomic data — pick a mode (linear probe, LoRA, Locon, full), train on BigWig tracks with agt finetune, use adapters, delta checkpoints… | genomicsxai/ | 162 | — | ~1k | Automated safety check: Pass | Apache-2.0 | 22 days ago |
| 48 | Complete CLI reference for the ADS AQUA command-line interface (ads aqua). | oracle/ | 125 | — | ~2.1k | Automated safety check: Pass | UPL-1.0 | 1 mo ago |
Questions, answered from the data.
What is the best fine-tuning skill?
Peft Fine Tuning from Orchestra-Research/AI-Research-SKILLs ranks first of the 313 fine-tuning skills listed here, with the highest score: its repository has 13k GitHub stars, 9 other GitHub owners carry a copy, its SKILL.md loads about 3.1k tokens and it passes the automated safety check with no findings. Next come Hugging Face LLM Trainer and Sentence-Transformers Training Router.
Which fine-tuning skills are official?
50 of the 313 fine-tuning skills are official, published by the vendor's own GitHub organization: Hugging Face LLM Trainer, Sentence-Transformers Training Router, Dataset Evaluation, Dataset Transformation, Finetuning Technique and 45 more.
How are these skills ranked?
By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.
Explore related skills
More topics in AI & LLM Engineering
- Building AI agents525
- Deep learning408
- Embeddings381
- LLM inference and serving364
- Retrieval-augmented generation360
- Prompt engineering350
- LLM evaluation303
- Speech recognition and synthesis272
- Structured output and tool calling271
- LLM cost and token optimization256
- LLM API integration218
- LLM observability217
- Model routing and gateways217
- LLM guardrails208
- Computer vision206
- Model hubs and datasets180
- GPU and accelerator computing171
- Diffusion and image models167
- Natural language processing141
- Reinforcement learning67
- AI interpretability23