Search

Fine-tuning

305 skills found, page 2.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
49

Run and extend embodiment-aware GR00T N1.7 post-training workflows from successful robot-policy collection through semantic recording, LeRobot conversion, statistics, fine-tuning, open-loop…

nvidia-isaac/video_to_data861—~2.6kAutomated safety check: PassUnknown2 days ago
50

Analyze a job description (pasted text OR a URL) and find the AI/ML/GenAI topics it requires that this learn-ai course does NOT yet cover.

starkyru/learn-ai107—~1.9kAutomated safety check: PassMIT2 mo ago
51

Walks through aligning language models with SimPO, a reference-free preference optimization method, using accelerate configs for Mistral 7B, Llama 3 8B and math-focused models.

Orchestra-Research/AI-Research-SKILLs13k4 repos~1.5kAutomated safety check: PassMIT3 mo ago
52

Guides reinforcement-learning post-training of LLMs with slime, which pairs Megatron-LM training with SGLang rollouts, including GRPO runs on GLM, Qwen3 and Llama 3 models.

Orchestra-Research/AI-Research-SKILLs13k4 repos~2.8kAutomated safety check: PassMIT3 mo ago
53

A skill your agent uses when the user wants to set up LLM training for the first time, or when traininghub is not yet installed/configured in the current environment.

Red-Hat-AI-Innovation-Team/training_hub100—~959Automated safety check: PassApache-2.03 days ago
54

Produces ranked, evidence-grounded next steps when an ML experiment has stalled, drawing on a Leeroopedia knowledge base or on fetched docs and issues.

Leeroo-AI/superml195—~4.8kAutomated safety check: PassApache-2.06 mo ago
55
55.Aqua DeploymentOfficial

Deploy LLM models on OCI using AI Quick Actions (AQUA) - single model, multi-model, stacked (LoRA), with GPU shape selection, vLLM configuration, streaming, and tool calling.

oracle/accelerated-data-science125—~2.4kAutomated safety check: PassUPL-1.01 mo ago
56

Discover Civitai models with the BUILT-IN downloadmodel action:"searchcivitai" and install/generate them locally.

artokun/comfyui-mcp803—~1.1kAutomated safety check: PassMIT6 days ago
57

Reference desk for NVIDIA Nemotron 3 Super — architecture, training data, recipes (pretrain/SFT/RL/eval/quantization), and deployment notes.

NVIDIA-NeMo/Nemotron2.1k—~2.4kAutomated safety check: PassApache-2.04 days ago
58

Add or modify an AReno algorithm, trainer, loss, advantage calculation, role model, or algorithm-specific configuration.

inclusionAI/AReno323—~501Automated safety check: PassApache-2.0yesterday
59

《AI Engineering: Building Applications with Foundation Models》(Chip Huyen, O'Reilly 2025) 方法论。

SpaceZephyr/career.skill160—~1.2kAutomated safety check: PassNo licence3 mo ago
60

Train or fine-tune SentenceTransformer bi-encoders, CrossEncoder rerankers, or SparseEncoder models, including losses, negatives, evaluation, distillation, LoRA, and Matryoshka.

waybarrios/opencode-power-pack534—~2.2kAutomated safety check: PassApache-2.05 days ago
61

Guides users through LLM post-training with Training Hub, including installation, algorithm selection (SFT, OSFT, LoRA), hyperparameter tuning, troubleshooting OOM errors, interpreting loss curves…

Red-Hat-AI-Innovation-Team/training_hub100—~2.8kAutomated safety check: PassApache-2.03 days ago
62

A skill your agent uses when building, reviewing, or editing TRL post-training workflows for agentic applications, including SFT, DPO, GRPO, RLOO, reward modeling, dataset formats, chat templates…

burtenshaw/training-agents153—~599Automated safety check: PassApache-2.028 days ago
63

Universal ARA Compiler. An agent skill from ARA-Labs/Agent-Native-Research-Artifact.

ARA-Labs/Agent-Native-Research-Artifact691—~6.4kAutomated safety check: PassMIT4 days ago
64

Checks training code, configs and math against documented framework behavior before an expensive run, citing a knowledge base or official docs for every claim.

Leeroo-AI/superml195—~3.8kAutomated safety check: PassApache-2.06 mo ago
65

Build this skill automates the adaptation of pre-trained machine learning models using transfer learning techniques.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.1kAutomated safety check: PassMITtoday
66

Guide users through Cosmos3 supervised fine-tuning (SFT) post-training: preparing the example dataset and Wan2.2 VAE, converting the base checkpoint to DCP, launching distributed training (paired…

NVIDIA/cosmos-framework560—~2.7kAutomated safety check: PassUnknownyesterday
67

Drive and verify Knit on a device or emulator through the headless debug bridge (am broadcast to app.getknit.knit.debug.<ACTION, replies as JSON) — send a message on one phone and confirm it landed…

getknit/knit133—~2.3kAutomated safety check: PassGPL-3.0yesterday
68
68.Gptq

Post-training 4-bit quantization for LLMs with minimal accuracy loss.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2.9kAutomated safety check: PassMIT3 mo ago
69

Half-Quadratic Quantization for LLMs without calibration data.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2.9kAutomated safety check: PassMIT3 mo ago
70

Implements and trains LLMs using Lightning AI's LitGPT with 20+ pretrained architectures (Llama, Gemma, Phi, Qwen, Mistral).

Orchestra-Research/AI-Research-SKILLs13k2 repos~2.8kAutomated safety check: PassMIT3 mo ago
71

High-performance RLHF framework with Ray+vLLM acceleration. An agent skill from Orchestra-Research/AI-Research-SKILLs.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2.1kAutomated safety check: NotesMIT3 mo ago
72

Loads large language models in 8-bit or 4-bit with bitsandbytes so they fit smaller GPUs, and sets up QLoRA fine-tuning on a 4-bit base model.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2.5kAutomated safety check: PassMIT3 mo ago
73

Trains LLMs with reinforcement learning using verl, from ByteDance's Seed team, with GRPO, PPO and other algorithms and swappable training and rollout backends.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2.4kAutomated safety check: PassMIT3 mo ago
74

Loads pre-trained Hugging Face Transformers models for text, vision and audio tasks, runs inference with pipelines and fine-tunes on custom datasets.

davila7/claude-code-templates33k11 repos~1.2kAutomated safety check: PassMITtoday
75

Drives fine-tuning on Alibaba Cloud Model Studio with the bl CLI: validate and upload data, create a job, watch it, export results and deploy the model.

modelstudioai/cli542—~1.8kAutomated safety check: PassApache-2.02 days ago
76

Orchestrates defect image generation for PCBA, metal surface and glass inspection with NVIDIA Cosmos AnomalyGen on OSMO, from cold-start Day 0 to real-photo Day 1 labeling.

NVIDIA/skills3.6k—~5kAutomated safety check: NotesApache-2.0yesterday
77

A skill your agent uses when engineering or selecting the best features for single-response classification or regression in MATLAB, whatever the data's modality — for non-tabular data it routes…

matlab/matlab-agentic-toolkit1.1k—~4.8kAutomated safety check: PassUnknown2 days ago
78

Explains how to train a model to be harmless with self-critique, revision and AI-generated preference feedback, with Hugging Face and TRL code for each stage.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2kAutomated safety check: PassMIT3 mo ago
79

Train custom TTS voices for Piper (ONNX format) using fine-tuning or from-scratch approaches.

sammcj/agentic-coding162—~1.4kAutomated safety check: PassApache-2.02 days ago
80

Guides GRPO reinforcement-learning fine-tuning of language models with TRL, centered on designing reward functions for formats, verifiable tasks and reasoning.

Orchestra-Research/AI-Research-SKILLs13k3 repos~4.3kAutomated safety check: PassMIT3 mo ago
81

Run the Nemotron-3.5 Lightning Text2SQL LoRA fine-tuning tutorial (NeMo Megatron-Bridge) end-to-end for the user on a single node: data prep, checkpoint conversion, LoRA fine-tuning of the 30B-A3B…

NVIDIA-NeMo/Nemotron2.1k—~1.9kAutomated safety check: PassApache-2.04 days ago
82

Qwen-Image-2.1 LoRA line (NOT Anima) — running cache/train through the daemon, make gui-qwen, the CacheRequest/TrainRequest flag surface and how to add a field, model-dir resolution, cache layout…

sorryhyun/anima_lora125—~1.9kAutomated safety check: NotesMITtoday
83

A skill your agent uses when the user wants to estimate GPU memory (VRAM) requirements for a training configuration, check if a model will fit on their GPUs, or plan GPU allocation for training.

Red-Hat-AI-Innovation-Team/training_hub100—~393Automated safety check: PassApache-2.03 days ago
84

Expert guidance for fine-tuning LLMs with LLaMA-Factory - WebUI no-code, 100+ models, 2/3/4/5/6/8-bit QLoRA, multimodal support

Orchestra-Research/AI-Research-SKILLs13k2 repos~618Automated safety check: PassMIT3 mo ago
85

Walks through nanoGPT, Karpathy's compact GPT implementation: training on Shakespeare, reproducing GPT-2, fine-tuning GPT-2 checkpoints and training on your own text.

Orchestra-Research/AI-Research-SKILLs13k2 repos~1.7kAutomated safety check: PassMIT3 mo ago
86

Reference for post-training language models with TRL: which trainer and dataset format to use for SFT, DPO, GRPO, KTO and reward models, and how to add LoRA.

huggingface/skills11k1 repo~1.1kAutomated safety check: PassApache-2.02 days ago
87

Calculate training costs for Tinker fine-tuning jobs. An agent skill from sundial-org/skills.

sundial-org/skills153—~1.2kAutomated safety check: PassNo licence2 mo ago
88

PyTorch training reference: architecture choice by data type, scaling rules, a training loop, optimizer and learning-rate choices, and fixes for loss spikes or OOM.

Orchestra-Research/AI-Research-SKILLs13k1 repo~2.8kAutomated safety check: PassMIT3 mo ago
89

The validated recipe for training MiniMax-Music3 planner-LM style adapters (artist/album clones) with ace-train mm3-lm-train and the Training Studio.

scragnog/HOT-Step-CPP174—~4kAutomated safety check: PassMIT2 days ago
90

Connect this agent to a running Guaardvark (self-hosted AI studio) and check what it can do right now.

guaardvark/guaardvark258—~1.2kAutomated safety check: PassMITtoday
91

Run the Nemotron-3 Ultra Text2SQL LoRA fine-tuning tutorial (NeMo Megatron-Bridge) end-to-end for the user on their SLURM cluster: data prep, distributed checkpoint conversion, and packed LoRA…

NVIDIA-NeMo/Nemotron2.1k—~1.9kAutomated safety check: PassApache-2.04 days ago
92

Validate Draw Things LoRA training end to end with draw-things-cli, including tiny-dataset training, loss and scaler checks, checkpoint sanity, and base-versus-LoRA generation comparison.

drawthingsai/draw-things-community584—~2kAutomated safety check: PassGPL-3.0yesterday
93

A skill your agent uses when working with ComfyUI workflows in OpenMontage, including comfyuiimage/comfyuivideo/comfyuimusic, custom workflowjson/workflowpath inputs, outputnode selection, missing…

calesthio/OpenMontage66k—~2kAutomated safety check: PassAGPL-3.07 days ago
94
94.Aqua FinetuningOfficial

Fine-tune LLM models using LoRA on OCI AI Quick Actions (AQUA).

oracle/accelerated-data-science125—~1.7kAutomated safety check: PassUPL-1.01 mo ago
95

Use sub-agents as test subjects to iteratively improve .claude/ instruction files.

adam-s/intercept189—~4.6kAutomated safety check: PassMIT2 mo ago
96

Reference desk for NVIDIA Nemotron 3 Ultra (550B-A55B) — architecture, NVFP4 pretraining, SFT, MOPD (multi-teacher on-policy distillation), MTP boosting, quantization, inference.

NVIDIA-NeMo/Nemotron2.1k—~1.8kAutomated safety check: PassApache-2.04 days ago