Topic · AI & LLM Engineering

Best fine-tuning skills for Claude Code, Codex and other agents.

Skills that fine-tune language models with LoRA, SFT and preference optimisation.
skills
313
official
50

Fine-tuning skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Fine-tuning skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Parameter-efficient fine-tuning for LLMs using LoRA, QLoRA, and 25+ methods.

Orchestra-Research/AI-Research-SKILLs13k9 repos~3.1kAutomated safety check: PassMIT3 mo ago
2

Trains or fine-tunes language and vision models with TRL or Unsloth on Hugging Face Jobs cloud GPUs, then converts the results to GGUF.

huggingface/skills11k3 repos~7.2kAutomated safety check: PassApache-2.06 days ago
3

Routes a sentence-transformers training task to the right model type and required reference docs and example scripts, covering bi-encoders, rerankers, sparse and multi-vector models.

huggingface/skills11k1 repo~2.6kAutomated safety check: PassApache-2.06 days ago
4

Fix a GitHub issue on OpenPipe/ART and open a PR. An agent skill from OpenPipe/ART.

OpenPipe/ART11k—~840Automated safety check: NotesApache-2.0yesterday
5

Validates dataset formatting and quality for SageMaker model fine-tuning (SFT, DPO, or RLVR).

awslabs/agent-plugins9122 repos~1.3kAutomated safety check: PassApache-2.0yesterday
6

RL training reference for the ART framework. An agent skill from OpenPipe/ART.

OpenPipe/ART11k—~2.4kAutomated safety check: PassApache-2.0yesterday
7

Prepare, validate, launch-plan, monitor, resume, and stop configurable Qwopus 27B reinforcement-learning workflows for GRPO or GSPO.

R6410418/Jackrong-llm-finetuning-guide1.7k—~830Automated safety check: PassApache-2.02 mo ago
8

Guides LLM fine-tuning with LoRA and QLoRA through Hugging Face PEFT, from dataset validation and training checks to adapter merging, quantization and deployment.

Jeffallan/claude-skills12k1 repo~1.7kAutomated safety check: PassMIT4 days ago
9

Generates code that transforms datasets between ML schemas for model training or evaluation.

awslabs/agent-plugins9122 repos~3.5kAutomated safety check: PassApache-2.0yesterday
10

SFT training reference for the ART framework. An agent skill from OpenPipe/ART.

OpenPipe/ART11k—~2.9kAutomated safety check: PassApache-2.0yesterday
11

Check Kiln's fine-tunable model list for deprecated or unsupported base models.

Kiln-AI/Kiln5.2k—~1.9kAutomated safety check: NotesUnknowntoday
12

Fine-tune LLMs using reinforcement learning with TRL - SFT for instruction tuning, DPO for preference alignment, PPO/GRPO for reward optimization, and reward model training.

Orchestra-Research/AI-Research-SKILLs13k7 repos~2.9kAutomated safety check: PassMIT3 mo ago
13

A skill your agent uses when the user wants to optimize configurable system parameters against a measurable scalar objective, especially for model training, inference, quantitative strategies…

Optim-Agent/optim-agent800—~1.3kAutomated safety check: PassMIT1 mo ago
14

Integrate a benchmark or custom environment into SAfactory using fixed adapter templates and local contract tests, optionally run Docker/RJob evaluation, or prepare GRPO/RL training.

AI45Lab/SAfactory236—~1.8kAutomated safety check: PassNo licence13 days ago
15

Trigger this skill when the user wants to train, fine-tune, or adapt Gemma models (e.g.

google-gemma/gemma-skills1k—~1.9kAutomated safety check: PassApache-2.0today
16

Selects a fine-tuning technique (SFT, DPO, RLVR, or RLAIF) for the user's use case and validates it against the selected model's available recipes.

awslabs/agent-plugins9121 repo~604Automated safety check: PassApache-2.0yesterday
17

Trains and evaluates several WiFi-signal-based pose and sensing models, from unsupervised pose estimation to domain adaptation and publishing.

ruvnet/RuView97k—~1.3kAutomated safety check: NotesMITtoday
18

This skill should be used when picking or diagnosing a training move (SFT, LoRA, DPO/KTO/ORPO, RFT, GRPO/PPO/RLOO, RLHF), or when the user mentions fine-tuning, post-training, training recipe…

evo-hq/evo1.5k—~4.5kAutomated safety check: PassApache-2.02 days ago
19

Router for adding a diffusion or omni pipeline to verl-omni.

verl-project/verl-omni1.2k—~1kAutomated safety check: PassApache-2.0today
20

Generate images with FLUX models (Black Forest Labs) via inference.sh CLI.

danielmeppiel/agentic-sdlc-handbook1572 repos~778Automated safety check: PassUnknown3 mo ago
21

Add a cross-cutting decision pattern under src/nemotron/steps/patterns/.

NVIDIA-NeMo/Nemotron2.1k—~1.4kAutomated safety check: PassApache-2.0yesterday
22

Trains and fine-tunes object detection, image classification and SAM or SAM2 segmentation models on Hugging Face Jobs cloud GPUs and saves the results to the Hub.

huggingface/skills11k1 repo~7.5kAutomated safety check: PassApache-2.06 days ago
23

Runs and configures the anomalib tiled-ensemble pipeline, which trains/evaluates one model per image tile and merges results (with optional seam smoothing) for high-resolution anomaly detection.

open-edge-platform/anomalib6.2k—~1.4kAutomated safety check: PassApache-2.0today
24

Connect this agent to a running Guaardvark (self-hosted AI studio) and check what it can do right now.

guaardvark/guaardvark2511 repo~1.2kAutomated safety check: PassMITtoday
25

Guide for adding a new reward scorer to verl-omni and wiring it into a run.

verl-project/verl-omni1.2k—~648Automated safety check: PassApache-2.0today
26

Guides building on the 0G Compute Network, a decentralized GPU marketplace for AI inference and fine-tuning, with SDK patterns and CLI commands.

internet-court/internet-court-skill6.4k1 repo~1.9kAutomated safety check: PassUnknown1 mo ago
27

Toolbox for markerless animal pose estimation with DeepLabCut.

NeuroAIHub/BrainPilot1k—~1.7kAutomated safety check: PassAGPL-3.05 days ago
28

Add a new step under src/nemotron/steps/<category/<stepid/ — manifest (step.toml), runner glue, configs, and per-step README.md.

NVIDIA-NeMo/Nemotron2.1k—~1.7kAutomated safety check: PassApache-2.0yesterday
29

MLX Swift LM - Run LLMs and VLMs on Apple Silicon using MLX.

kellyvv/PhoneClaw1.3k—~3.7kAutomated safety check: PassApache-2.02 mo ago
30

Train or fine-tune language models with TRL or Unsloth on Hugging Face Jobs, including SFT, DPO, GRPO, reward models, and GGUF conversion.

waybarrios/opencode-power-pack533—~3kAutomated safety check: PassApache-2.0yesterday
31

Selectively monitor important long-running or resource-intensive commands with Haoleme by prefixing them with hao, so status, output, and completion notifications sync to the mobile app.

HaolemeApp/Haoleme157—~1.3kAutomated safety check: PassAGPL-3.01 mo ago
32

A standardized CLI wrapper for Uni-Mol molecular ML workflows that handles representation extraction (embeddings), model training (regression/classification), and property prediction with built-in…

jinzhezenggroup/computational-chemistry-agent-skills1481 repo~1.5kAutomated safety check: PassLGPL-3.0-or-lateryesterday
33

Expert guidance for fast fine-tuning with Unsloth - 2-5x faster training, 50-80% less memory, LoRA/QLoRA optimization

Orchestra-Research/AI-Research-SKILLs13k11 repos~577Automated safety check: PassMIT3 mo ago
34

Lilly community-research skill. An agent skill from ssaaffaakk/Lilly.

ssaaffaakk/Lilly171—~1.4kAutomated safety check: PassMITyesterday
35

Audit imaging acquisition, reconstruction, series eligibility, quantitative transforms and protocol drift; not model training.

huang-sir1/radiology-skills1.9k—~2.8kAutomated safety check: PassUnknown16 days ago
36

Train custom LoRAs with ostris AI-Toolkit. An agent skill from artokun/comfyui-mcp.

artokun/comfyui-mcp793—~2.7kAutomated safety check: PassMIT2 days ago
37

Quick-reference help for fine-tuning language models with Axolotl, covering YAML configs, FSDP, context parallelism, compressed saves and dataset formats.

Orchestra-Research/AI-Research-SKILLs13k9 repos~1.2kAutomated safety check: PassMIT3 mo ago
38

Run, configure, retry, and validate AReno SFT, DPO, GSPO, GRPO, PPO, and agentic training.

inclusionAI/AReno323—~782Automated safety check: PassApache-2.013 days ago
39

A skill your agent uses whenever the user asks to generate, collect, inspect, or prepare early-experience training data (Implicit World Modeling or Self-Reflection, in the sense of arXiv:2510.08558)…

OSU-NLP-Group/EarlyExperience102—~4.1kAutomated safety check: PassMIT3 mo ago
40

A skill your agent uses when designing, implementing, reviewing, or debugging supervised fine-tuning with TRL SFTTrainer or trl sft, especially for agentic models trained on chat messages…

burtenshaw/training-agents153—~685Automated safety check: PassApache-2.024 days ago
41

Configure and launch SparkDiffusion sparse finetuning for Wan 2.1 or Wan 2.2.

AlibabaResearch/SparkDiffusion490—~904Automated safety check: PassApache-2.02 days ago
42

Train object-detection, image-classification, or SAM segmentation models on Hugging Face Jobs.

waybarrios/opencode-power-pack533—~2.7kAutomated safety check: PassApache-2.0yesterday
43

Build, deploy, evaluate, optimize, fine-tune, and manage Microsoft Foundry agents, models, and resources end to end.

microsoft/GitHub-Copilot-for-Azure2551 repo~6.7kAutomated safety check: PassMITtoday
44

Set up the NVIDIA "Build an Agent" DevX workshop as a working JupyterLab environment from INSIDE a locked-down OpenShell/NemoClaw sandbox, and hand the user the token URL + access commands.

brevdev/workshop-build-an-agent143—~5.2kAutomated safety check: PassApache-2.0today
45

Walks through aligning language models with SimPO, a reference-free preference optimization method, using accelerate configs for Mistral 7B, Llama 3 8B and math-focused models.

Orchestra-Research/AI-Research-SKILLs13k5 repos~1.5kAutomated safety check: PassMIT3 mo ago
46

Guides reinforcement-learning post-training of LLMs with slime, which pairs Megatron-LM training with SGLang rollouts, including GRPO runs on GLM, Qwen3 and Llama 3 models.

Orchestra-Research/AI-Research-SKILLs13k5 repos~2.8kAutomated safety check: PassMIT3 mo ago
47

Fine-tune or transfer-learn AlphaGenome-PyTorch on custom genomic data — pick a mode (linear probe, LoRA, Locon, full), train on BigWig tracks with agt finetune, use adapters, delta checkpoints…

genomicsxai/alphagenome-pytorch162—~1kAutomated safety check: PassApache-2.022 days ago
48
48.Aqua CLIOfficial

Complete CLI reference for the ADS AQUA command-line interface (ads aqua).

oracle/accelerated-data-science125—~2.1kAutomated safety check: PassUPL-1.01 mo ago

Questions, answered from the data.

What is the best fine-tuning skill?

Peft Fine Tuning from Orchestra-Research/AI-Research-SKILLs ranks first of the 313 fine-tuning skills listed here, with the highest score: its repository has 13k GitHub stars, 9 other GitHub owners carry a copy, its SKILL.md loads about 3.1k tokens and it passes the automated safety check with no findings. Next come Hugging Face LLM Trainer and Sentence-Transformers Training Router.

Which fine-tuning skills are official?

50 of the 313 fine-tuning skills are official, published by the vendor's own GitHub organization: Hugging Face LLM Trainer, Sentence-Transformers Training Router, Dataset Evaluation, Dataset Transformation, Finetuning Technique and 45 more.

How are these skills ranked?

By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.