Search
Fine-tuning
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 49 | Run and extend embodiment-aware GR00T N1.7 post-training workflows from successful robot-policy collection through semantic recording, LeRobot conversion, statistics, fine-tuning, open-loop… | nvidia-isaac/ | 861 | — | ~2.6k | Automated safety check: Pass | Unknown | 2 days ago |
| 50 | Analyze a job description (pasted text OR a URL) and find the AI/ML/GenAI topics it requires that this learn-ai course does NOT yet cover. | starkyru/ | 107 | — | ~1.9k | Automated safety check: Pass | MIT | 2 mo ago |
| 51 | Walks through aligning language models with SimPO, a reference-free preference optimization method, using accelerate configs for Mistral 7B, Llama 3 8B and math-focused models. | Orchestra-Research/ | 13k | 4 repos | ~1.5k | Automated safety check: Pass | MIT | 3 mo ago |
| 52 | Guides reinforcement-learning post-training of LLMs with slime, which pairs Megatron-LM training with SGLang rollouts, including GRPO runs on GLM, Qwen3 and Llama 3 models. | Orchestra-Research/ | 13k | 4 repos | ~2.8k | Automated safety check: Pass | MIT | 3 mo ago |
| 53 | 53.Setup Guide A skill your agent uses when the user wants to set up LLM training for the first time, or when traininghub is not yet installed/configured in the current environment. | Red-Hat-AI-Innovation-Team/ | 100 | — | ~959 | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 54 | Produces ranked, evidence-grounded next steps when an ML experiment has stalled, drawing on a Leeroopedia knowledge base or on fetched docs and issues. | Leeroo-AI/ | 195 | — | ~4.8k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 55 | Deploy LLM models on OCI using AI Quick Actions (AQUA) - single model, multi-model, stacked (LoRA), with GPU shape selection, vLLM configuration, streaming, and tool calling. | oracle/ | 125 | — | ~2.4k | Automated safety check: Pass | UPL-1.0 | 1 mo ago |
| 56 | 56.Civitai Discover Civitai models with the BUILT-IN downloadmodel action:"searchcivitai" and install/generate them locally. | artokun/ | 803 | — | ~1.1k | Automated safety check: Pass | MIT | 6 days ago |
| 57 | Reference desk for NVIDIA Nemotron 3 Super — architecture, training data, recipes (pretrain/SFT/RL/eval/quantization), and deployment notes. | NVIDIA-NeMo/ | 2.1k | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | 4 days ago |
| 58 | Add or modify an AReno algorithm, trainer, loss, advantage calculation, role model, or algorithm-specific configuration. | inclusionAI/ | 323 | — | ~501 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 59 | 《AI Engineering: Building Applications with Foundation Models》(Chip Huyen, O'Reilly 2025) 方法论。 | SpaceZephyr/ | 160 | — | ~1.2k | Automated safety check: Pass | No licence | 3 mo ago |
| 60 | Train or fine-tune SentenceTransformer bi-encoders, CrossEncoder rerankers, or SparseEncoder models, including losses, negatives, evaluation, distillation, LoRA, and Matryoshka. | waybarrios/ | 534 | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 61 | Guides users through LLM post-training with Training Hub, including installation, algorithm selection (SFT, OSFT, LoRA), hyperparameter tuning, troubleshooting OOM errors, interpreting loss curves… | Red-Hat-AI-Innovation-Team/ | 100 | — | ~2.8k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 62 | A skill your agent uses when building, reviewing, or editing TRL post-training workflows for agentic applications, including SFT, DPO, GRPO, RLOO, reward modeling, dataset formats, chat templates… | burtenshaw/ | 153 | — | ~599 | Automated safety check: Pass | Apache-2.0 | 28 days ago |
| 63 | 63.Compiler Universal ARA Compiler. An agent skill from ARA-Labs/Agent-Native-Research-Artifact. | ARA-Labs/ | 691 | — | ~6.4k | Automated safety check: Pass | MIT | 4 days ago |
| 64 | Checks training code, configs and math against documented framework behavior before an expensive run, citing a knowledge base or official docs for every claim. | Leeroo-AI/ | 195 | — | ~3.8k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 65 | Build this skill automates the adaptation of pre-trained machine learning models using transfer learning techniques. | jeremylongshore/ | 2.8k | — | ~1.1k | Automated safety check: Pass | MIT | today |
| 66 | Guide users through Cosmos3 supervised fine-tuning (SFT) post-training: preparing the example dataset and Wan2.2 VAE, converting the base checkpoint to DCP, launching distributed training (paired… | NVIDIA/ | 560 | — | ~2.7k | Automated safety check: Pass | Unknown | yesterday |
| 67 | 67.Debug Bridge Drive and verify Knit on a device or emulator through the headless debug bridge (am broadcast to app.getknit.knit.debug.<ACTION, replies as JSON) — send a message on one phone and confirm it landed… | getknit/ | 133 | — | ~2.3k | Automated safety check: Pass | GPL-3.0 | yesterday |
| 68 | 68.Gptq Post-training 4-bit quantization for LLMs with minimal accuracy loss. | Orchestra-Research/ | 13k | 2 repos | ~2.9k | Automated safety check: Pass | MIT | 3 mo ago |
| 69 | Half-Quadratic Quantization for LLMs without calibration data. | Orchestra-Research/ | 13k | 2 repos | ~2.9k | Automated safety check: Pass | MIT | 3 mo ago |
| 70 | Implements and trains LLMs using Lightning AI's LitGPT with 20+ pretrained architectures (Llama, Gemma, Phi, Qwen, Mistral). | Orchestra-Research/ | 13k | 2 repos | ~2.8k | Automated safety check: Pass | MIT | 3 mo ago |
| 71 | High-performance RLHF framework with Ray+vLLM acceleration. An agent skill from Orchestra-Research/AI-Research-SKILLs. | Orchestra-Research/ | 13k | 2 repos | ~2.1k | Automated safety check: Notes | MIT | 3 mo ago |
| 72 | Loads large language models in 8-bit or 4-bit with bitsandbytes so they fit smaller GPUs, and sets up QLoRA fine-tuning on a 4-bit base model. | Orchestra-Research/ | 13k | 2 repos | ~2.5k | Automated safety check: Pass | MIT | 3 mo ago |
| 73 | Trains LLMs with reinforcement learning using verl, from ByteDance's Seed team, with GRPO, PPO and other algorithms and swappable training and rollout backends. | Orchestra-Research/ | 13k | 2 repos | ~2.4k | Automated safety check: Pass | MIT | 3 mo ago |
| 74 | Loads pre-trained Hugging Face Transformers models for text, vision and audio tasks, runs inference with pipelines and fine-tunes on custom datasets. | davila7/ | 33k | 11 repos | ~1.2k | Automated safety check: Pass | MIT | today |
| 75 | Drives fine-tuning on Alibaba Cloud Model Studio with the bl CLI: validate and upload data, create a job, watch it, export results and deploy the model. | modelstudioai/ | 542 | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 76 | Orchestrates defect image generation for PCBA, metal surface and glass inspection with NVIDIA Cosmos AnomalyGen on OSMO, from cold-start Day 0 to real-photo Day 1 labeling. | NVIDIA/ | 3.6k | — | ~5k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 77 | A skill your agent uses when engineering or selecting the best features for single-response classification or regression in MATLAB, whatever the data's modality — for non-tabular data it routes… | matlab/ | 1.1k | — | ~4.8k | Automated safety check: Pass | Unknown | 2 days ago |
| 78 | Explains how to train a model to be harmless with self-critique, revision and AI-generated preference feedback, with Hugging Face and TRL code for each stage. | Orchestra-Research/ | 13k | 2 repos | ~2k | Automated safety check: Pass | MIT | 3 mo ago |
| 79 | Train custom TTS voices for Piper (ONNX format) using fine-tuning or from-scratch approaches. | sammcj/ | 162 | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 80 | Guides GRPO reinforcement-learning fine-tuning of language models with TRL, centered on designing reward functions for formats, verifiable tasks and reasoning. | Orchestra-Research/ | 13k | 3 repos | ~4.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 81 | Run the Nemotron-3.5 Lightning Text2SQL LoRA fine-tuning tutorial (NeMo Megatron-Bridge) end-to-end for the user on a single node: data prep, checkpoint conversion, LoRA fine-tuning of the 30B-A3B… | NVIDIA-NeMo/ | 2.1k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | 4 days ago |
| 82 | 82.Qwen21 Qwen-Image-2.1 LoRA line (NOT Anima) — running cache/train through the daemon, make gui-qwen, the CacheRequest/TrainRequest flag surface and how to add a field, model-dir resolution, cache layout… | sorryhyun/ | 125 | — | ~1.9k | Automated safety check: Notes | MIT | today |
| 83 | A skill your agent uses when the user wants to estimate GPU memory (VRAM) requirements for a training configuration, check if a model will fit on their GPUs, or plan GPU allocation for training. | Red-Hat-AI-Innovation-Team/ | 100 | — | ~393 | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 84 | Expert guidance for fine-tuning LLMs with LLaMA-Factory - WebUI no-code, 100+ models, 2/3/4/5/6/8-bit QLoRA, multimodal support | Orchestra-Research/ | 13k | 2 repos | ~618 | Automated safety check: Pass | MIT | 3 mo ago |
| 85 | Walks through nanoGPT, Karpathy's compact GPT implementation: training on Shakespeare, reproducing GPT-2, fine-tuning GPT-2 checkpoints and training on your own text. | Orchestra-Research/ | 13k | 2 repos | ~1.7k | Automated safety check: Pass | MIT | 3 mo ago |
| 86 | Reference for post-training language models with TRL: which trainer and dataset format to use for SFT, DPO, GRPO, KTO and reward models, and how to add LoRA. | huggingface/ | 11k | 1 repo | ~1.1k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 87 | Calculate training costs for Tinker fine-tuning jobs. An agent skill from sundial-org/skills. | sundial-org/ | 153 | — | ~1.2k | Automated safety check: Pass | No licence | 2 mo ago |
| 88 | PyTorch training reference: architecture choice by data type, scaling rules, a training loop, optimizer and learning-rate choices, and fixes for loss spikes or OOM. | Orchestra-Research/ | 13k | 1 repo | ~2.8k | Automated safety check: Pass | MIT | 3 mo ago |
| 89 | The validated recipe for training MiniMax-Music3 planner-LM style adapters (artist/album clones) with ace-train mm3-lm-train and the Training Studio. | scragnog/ | 174 | — | ~4k | Automated safety check: Pass | MIT | 2 days ago |
| 90 | 90.Setup Connect this agent to a running Guaardvark (self-hosted AI studio) and check what it can do right now. | guaardvark/ | 258 | — | ~1.2k | Automated safety check: Pass | MIT | today |
| 91 | Run the Nemotron-3 Ultra Text2SQL LoRA fine-tuning tutorial (NeMo Megatron-Bridge) end-to-end for the user on their SLURM cluster: data prep, distributed checkpoint conversion, and packed LoRA… | NVIDIA-NeMo/ | 2.1k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | 4 days ago |
| 92 | 92.Train Lora Validate Draw Things LoRA training end to end with draw-things-cli, including tiny-dataset training, loss and scaler checks, checkpoint sanity, and base-versus-LoRA generation comparison. | drawthingsai/ | 584 | — | ~2k | Automated safety check: Pass | GPL-3.0 | yesterday |
| 93 | 93.Comfyui A skill your agent uses when working with ComfyUI workflows in OpenMontage, including comfyuiimage/comfyuivideo/comfyuimusic, custom workflowjson/workflowpath inputs, outputnode selection, missing… | calesthio/ | 66k | — | ~2k | Automated safety check: Pass | AGPL-3.0 | 7 days ago |
| 94 | Fine-tune LLM models using LoRA on OCI AI Quick Actions (AQUA). | oracle/ | 125 | — | ~1.7k | Automated safety check: Pass | UPL-1.0 | 1 mo ago |
| 95 | Use sub-agents as test subjects to iteratively improve .claude/ instruction files. | adam-s/ | 189 | — | ~4.6k | Automated safety check: Pass | MIT | 2 mo ago |
| 96 | Reference desk for NVIDIA Nemotron 3 Ultra (550B-A55B) — architecture, NVFP4 pretraining, SFT, MOPD (multi-teacher on-policy distillation), MTP boosting, quantization, inference. | NVIDIA-NeMo/ | 2.1k | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | 4 days ago |