Topic · AI & LLM Engineering

Best fine-tuning skills, page 2

Skills #49–96 of 313, ranked by score.

Fine-tuning skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Fine-tuning skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
49

Rules for adding a new model or model family to the finetuning pipeline, or changing finetuning behavior for an existing one — engine-agnostic customization via family hooks instead of if/else in…

overmind-core/overmind544—~3.1kAutomated safety check: PassAGPL-3.0yesterday
50
50.Quax

A skill your agent uses when writing, reviewing, or debugging JAX code that involves quax — custom array-ish objects (physical units, LoRA, sparse, symbolic zero, named axes), quax.quaxify…

nstarman/quax143—~5.5kAutomated safety check: PassApache-2.03 days ago
51

Interact with a Meshtastic LoRa mesh network through MESH-API — list nodes, read messages, send texts, and check connection status.

mr-tbot/mesh-api179—~1.8kAutomated safety check: PassGPL-3.02 mo ago
52

Run and extend embodiment-aware GR00T N1.7 post-training workflows from successful robot-policy collection through semantic recording, LeRobot conversion, statistics, fine-tuning, open-loop…

nvidia-isaac/video_to_data847—~2.6kAutomated safety check: PassUnknownyesterday
53

Analyze a job description (pasted text OR a URL) and find the AI/ML/GenAI topics it requires that this learn-ai course does NOT yet cover.

starkyru/learn-ai105—~1.9kAutomated safety check: PassMIT2 mo ago
54

A skill your agent uses when the user wants to set up LLM training for the first time, or when traininghub is not yet installed/configured in the current environment.

Red-Hat-AI-Innovation-Team/training_hub100—~959Automated safety check: PassApache-2.05 days ago
55

Produces ranked, evidence-grounded next steps when an ML experiment has stalled, drawing on a Leeroopedia knowledge base or on fetched docs and issues.

Leeroo-AI/superml195—~4.8kAutomated safety check: PassApache-2.06 mo ago
56
56.Aqua DeploymentOfficial

Deploy LLM models on OCI using AI Quick Actions (AQUA) - single model, multi-model, stacked (LoRA), with GPU shape selection, vLLM configuration, streaming, and tool calling.

oracle/accelerated-data-science125—~2.4kAutomated safety check: PassUPL-1.01 mo ago
57

Discover Civitai models with the BUILT-IN downloadmodel action:"searchcivitai" and install/generate them locally.

artokun/comfyui-mcp793—~1.1kAutomated safety check: PassMIT2 days ago
58

Reference desk for NVIDIA Nemotron 3 Super — architecture, training data, recipes (pretrain/SFT/RL/eval/quantization), and deployment notes.

NVIDIA-NeMo/Nemotron2.1k—~2.4kAutomated safety check: PassApache-2.0yesterday
59
59.Cast

Build consistent characters, environments and props in Guaardvark's Cast Library and train LoRAs for them locally (reference photos → vision bible → sample plan → approved samples → training).

guaardvark/guaardvark251—~673Automated safety check: PassMITyesterday
60

Add or modify an AReno algorithm, trainer, loss, advantage calculation, role model, or algorithm-specific configuration.

inclusionAI/AReno323—~501Automated safety check: PassApache-2.013 days ago
61
61.Gptq

Post-training 4-bit quantization for LLMs with minimal accuracy loss.

Orchestra-Research/AI-Research-SKILLs13k3 repos~2.9kAutomated safety check: PassMIT3 mo ago
62

Half-Quadratic Quantization for LLMs without calibration data.

Orchestra-Research/AI-Research-SKILLs13k3 repos~2.9kAutomated safety check: PassMIT3 mo ago
63

Implements and trains LLMs using Lightning AI's LitGPT with 20+ pretrained architectures (Llama, Gemma, Phi, Qwen, Mistral).

Orchestra-Research/AI-Research-SKILLs13k3 repos~2.8kAutomated safety check: PassMIT3 mo ago
64

High-performance RLHF framework with Ray+vLLM acceleration. An agent skill from Orchestra-Research/AI-Research-SKILLs.

Orchestra-Research/AI-Research-SKILLs13k3 repos~2.1kAutomated safety check: NotesMIT3 mo ago
65

Loads large language models in 8-bit or 4-bit with bitsandbytes so they fit smaller GPUs, and sets up QLoRA fine-tuning on a 4-bit base model.

Orchestra-Research/AI-Research-SKILLs13k3 repos~2.5kAutomated safety check: PassMIT3 mo ago
66

Trains LLMs with reinforcement learning using verl, from ByteDance's Seed team, with GRPO, PPO and other algorithms and swappable training and rollout backends.

Orchestra-Research/AI-Research-SKILLs13k3 repos~2.4kAutomated safety check: PassMIT3 mo ago
67

《AI Engineering: Building Applications with Foundation Models》(Chip Huyen, O'Reilly 2025) 方法论。

SpaceZephyr/career.skill160—~1.2kAutomated safety check: PassNo licence3 mo ago
68

Train or fine-tune SentenceTransformer bi-encoders, CrossEncoder rerankers, or SparseEncoder models, including losses, negatives, evaluation, distillation, LoRA, and Matryoshka.

waybarrios/opencode-power-pack533—~2.2kAutomated safety check: PassApache-2.0yesterday
69

Guides users through LLM post-training with Training Hub, including installation, algorithm selection (SFT, OSFT, LoRA), hyperparameter tuning, troubleshooting OOM errors, interpreting loss curves…

Red-Hat-AI-Innovation-Team/training_hub100—~2.8kAutomated safety check: PassApache-2.05 days ago
70

A skill your agent uses when building, reviewing, or editing TRL post-training workflows for agentic applications, including SFT, DPO, GRPO, RLOO, reward modeling, dataset formats, chat templates…

burtenshaw/training-agents153—~599Automated safety check: PassApache-2.024 days ago
71

Universal ARA Compiler. An agent skill from ARA-Labs/Agent-Native-Research-Artifact.

ARA-Labs/Agent-Native-Research-Artifact690—~6.4kAutomated safety check: PassMITtoday
72

Checks training code, configs and math against documented framework behavior before an expensive run, citing a knowledge base or official docs for every claim.

Leeroo-AI/superml195—~3.8kAutomated safety check: PassApache-2.06 mo ago
73

Guide users through Cosmos3 supervised fine-tuning (SFT) post-training: preparing the example dataset and Wan2.2 VAE, converting the base checkpoint to DCP, launching distributed training (paired…

NVIDIA/cosmos-framework556—~2.7kAutomated safety check: PassUnknown11 days ago
74

Explains how to train a model to be harmless with self-critique, revision and AI-generated preference feedback, with Hugging Face and TRL code for each stage.

Orchestra-Research/AI-Research-SKILLs13k3 repos~2kAutomated safety check: PassMIT3 mo ago
75

Loads pre-trained Hugging Face Transformers models for text, vision and audio tasks, runs inference with pipelines and fine-tunes on custom datasets.

davila7/claude-code-templates32k12 repos~1.2kAutomated safety check: PassMITtoday
76

Drive and verify Knit on a device or emulator through the headless debug bridge (am broadcast to app.getknit.knit.debug.<ACTION, replies as JSON) — send a message on one phone and confirm it landed…

getknit/knit130—~2.2kAutomated safety check: PassGPL-3.03 days ago
77

PyTorch training reference: architecture choice by data type, scaling rules, a training loop, optimizer and learning-rate choices, and fixes for loss spikes or OOM.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2.8kAutomated safety check: PassMIT3 mo ago
78

Expert guidance for fine-tuning LLMs with LLaMA-Factory - WebUI no-code, 100+ models, 2/3/4/5/6/8-bit QLoRA, multimodal support

Orchestra-Research/AI-Research-SKILLs13k3 repos~618Automated safety check: PassMIT3 mo ago
79

Walks through nanoGPT, Karpathy's compact GPT implementation: training on Shakespeare, reproducing GPT-2, fine-tuning GPT-2 checkpoints and training on your own text.

Orchestra-Research/AI-Research-SKILLs13k3 repos~1.7kAutomated safety check: PassMIT3 mo ago
80

Rigor Improve implementation leaf skill for auditable candidate implementation in deep learning research repositories.

lllllllama/RigorPilot-Skills4971 repo~648Automated safety check: PassMIT14 days ago
81

Guides GRPO reinforcement-learning fine-tuning of language models with TRL, centered on designing reward functions for formats, verifiable tasks and reasoning.

Orchestra-Research/AI-Research-SKILLs13k4 repos~4.3kAutomated safety check: PassMIT3 mo ago
82

Drives fine-tuning on Alibaba Cloud Model Studio with the bl CLI: validate and upload data, create a job, watch it, export results and deploy the model.

modelstudioai/cli541—~1.8kAutomated safety check: PassApache-2.07 days ago
83

Orchestrates defect image generation for PCBA, metal surface and glass inspection with NVIDIA Cosmos AnomalyGen on OSMO, from cold-start Day 0 to real-photo Day 1 labeling.

NVIDIA/skills3.5k—~5kAutomated safety check: NotesApache-2.0today
84

Train custom TTS voices for Piper (ONNX format) using fine-tuning or from-scratch approaches.

sammcj/agentic-coding162—~1.4kAutomated safety check: PassApache-2.02 days ago
85

Run the Nemotron-3.5 Lightning Text2SQL LoRA fine-tuning tutorial (NeMo Megatron-Bridge) end-to-end for the user on a single node: data prep, checkpoint conversion, LoRA fine-tuning of the 30B-A3B…

NVIDIA-NeMo/Nemotron2.1k—~1.9kAutomated safety check: PassApache-2.0yesterday
86

Qwen-Image-2.1 LoRA line (NOT Anima) — running cache/train through the daemon, make gui-qwen, the CacheRequest/TrainRequest flag surface and how to add a field, model-dir resolution, cache layout…

sorryhyun/anima_lora125—~1.9kAutomated safety check: NotesMITyesterday
87

A skill your agent uses when the user wants to estimate GPU memory (VRAM) requirements for a training configuration, check if a model will fit on their GPUs, or plan GPU allocation for training.

Red-Hat-AI-Innovation-Team/training_hub100—~393Automated safety check: PassApache-2.05 days ago
88

Reference for post-training language models with TRL: which trainer and dataset format to use for SFT, DPO, GRPO, KTO and reward models, and how to add LoRA.

huggingface/skills11k1 repo~1.1kAutomated safety check: PassApache-2.06 days ago
89

Calculate training costs for Tinker fine-tuning jobs. An agent skill from sundial-org/skills.

sundial-org/skills152—~1.2kAutomated safety check: PassNo licence2 mo ago
90

Fine-tunes and evaluates OpenVLA-OFT and OFT+ robot policies with LoRA and continuous action heads on LIBERO simulation and ALOHA real-robot setups.

Orchestra-Research/AI-Research-SKILLs13k1 repo~3.7kAutomated safety check: PassMIT3 mo ago
91

Fine-tunes and serves Physical Intelligence's pi0, pi0-fast and pi0.5 robot policies with JAX or PyTorch, including checkpoint conversion and policy servers.

Orchestra-Research/AI-Research-SKILLs13k1 repo~3.6kAutomated safety check: PassMIT3 mo ago
92

The validated recipe for training MiniMax-Music3 planner-LM style adapters (artist/album clones) with ace-train mm3-lm-train and the Training Studio.

scragnog/HOT-Step-CPP170—~4kAutomated safety check: PassMIT2 days ago
93

Run the Nemotron-3 Ultra Text2SQL LoRA fine-tuning tutorial (NeMo Megatron-Bridge) end-to-end for the user on their SLURM cluster: data prep, distributed checkpoint conversion, and packed LoRA…

NVIDIA-NeMo/Nemotron2.1k—~1.9kAutomated safety check: PassApache-2.0yesterday
94

Validate Draw Things LoRA training end to end with draw-things-cli, including tiny-dataset training, loss and scaler checks, checkpoint sanity, and base-versus-LoRA generation comparison.

drawthingsai/draw-things-community579—~2kAutomated safety check: PassGPL-3.0yesterday
95

A skill your agent uses for BetterTogether, prompt plus weight optimization, fine-tuning sequences, and strategy chains like p - w - p.

OmidZamani/dspy-skills1241 repo~756Automated safety check: PassMIT3 mo ago
96

A skill your agent uses when working with ComfyUI workflows in OpenMontage, including comfyuiimage/comfyuivideo/comfyuimusic, custom workflowjson/workflowpath inputs, outputnode selection, missing…

calesthio/OpenMontage65k—~2kAutomated safety check: PassAGPL-3.04 days ago