Topic · AI & LLM Engineering
Best fine-tuning skills, page 2
Fine-tuning skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 49 | Rules for adding a new model or model family to the finetuning pipeline, or changing finetuning behavior for an existing one — engine-agnostic customization via family hooks instead of if/else in… | overmind-core/ | 544 | — | ~3.1k | Automated safety check: Pass | AGPL-3.0 | yesterday |
| 50 | 50.Quax A skill your agent uses when writing, reviewing, or debugging JAX code that involves quax — custom array-ish objects (physical units, LoRA, sparse, symbolic zero, named axes), quax.quaxify… | nstarman/ | 143 | — | ~5.5k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 51 | 51.Mesh API Interact with a Meshtastic LoRa mesh network through MESH-API — list nodes, read messages, send texts, and check connection status. | mr-tbot/ | 179 | — | ~1.8k | Automated safety check: Pass | GPL-3.0 | 2 mo ago |
| 52 | Run and extend embodiment-aware GR00T N1.7 post-training workflows from successful robot-policy collection through semantic recording, LeRobot conversion, statistics, fine-tuning, open-loop… | nvidia-isaac/ | 847 | — | ~2.6k | Automated safety check: Pass | Unknown | yesterday |
| 53 | Analyze a job description (pasted text OR a URL) and find the AI/ML/GenAI topics it requires that this learn-ai course does NOT yet cover. | starkyru/ | 105 | — | ~1.9k | Automated safety check: Pass | MIT | 2 mo ago |
| 54 | 54.Setup Guide A skill your agent uses when the user wants to set up LLM training for the first time, or when traininghub is not yet installed/configured in the current environment. | Red-Hat-AI-Innovation-Team/ | 100 | — | ~959 | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 55 | Produces ranked, evidence-grounded next steps when an ML experiment has stalled, drawing on a Leeroopedia knowledge base or on fetched docs and issues. | Leeroo-AI/ | 195 | — | ~4.8k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 56 | Deploy LLM models on OCI using AI Quick Actions (AQUA) - single model, multi-model, stacked (LoRA), with GPU shape selection, vLLM configuration, streaming, and tool calling. | oracle/ | 125 | — | ~2.4k | Automated safety check: Pass | UPL-1.0 | 1 mo ago |
| 57 | 57.Civitai Discover Civitai models with the BUILT-IN downloadmodel action:"searchcivitai" and install/generate them locally. | artokun/ | 793 | — | ~1.1k | Automated safety check: Pass | MIT | 2 days ago |
| 58 | Reference desk for NVIDIA Nemotron 3 Super — architecture, training data, recipes (pretrain/SFT/RL/eval/quantization), and deployment notes. | NVIDIA-NeMo/ | 2.1k | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 59 | 59.Cast Build consistent characters, environments and props in Guaardvark's Cast Library and train LoRAs for them locally (reference photos → vision bible → sample plan → approved samples → training). | guaardvark/ | 251 | — | ~673 | Automated safety check: Pass | MIT | yesterday |
| 60 | Add or modify an AReno algorithm, trainer, loss, advantage calculation, role model, or algorithm-specific configuration. | inclusionAI/ | 323 | — | ~501 | Automated safety check: Pass | Apache-2.0 | 13 days ago |
| 61 | 61.Gptq Post-training 4-bit quantization for LLMs with minimal accuracy loss. | Orchestra-Research/ | 13k | 3 repos | ~2.9k | Automated safety check: Pass | MIT | 3 mo ago |
| 62 | Half-Quadratic Quantization for LLMs without calibration data. | Orchestra-Research/ | 13k | 3 repos | ~2.9k | Automated safety check: Pass | MIT | 3 mo ago |
| 63 | Implements and trains LLMs using Lightning AI's LitGPT with 20+ pretrained architectures (Llama, Gemma, Phi, Qwen, Mistral). | Orchestra-Research/ | 13k | 3 repos | ~2.8k | Automated safety check: Pass | MIT | 3 mo ago |
| 64 | High-performance RLHF framework with Ray+vLLM acceleration. An agent skill from Orchestra-Research/AI-Research-SKILLs. | Orchestra-Research/ | 13k | 3 repos | ~2.1k | Automated safety check: Notes | MIT | 3 mo ago |
| 65 | Loads large language models in 8-bit or 4-bit with bitsandbytes so they fit smaller GPUs, and sets up QLoRA fine-tuning on a 4-bit base model. | Orchestra-Research/ | 13k | 3 repos | ~2.5k | Automated safety check: Pass | MIT | 3 mo ago |
| 66 | Trains LLMs with reinforcement learning using verl, from ByteDance's Seed team, with GRPO, PPO and other algorithms and swappable training and rollout backends. | Orchestra-Research/ | 13k | 3 repos | ~2.4k | Automated safety check: Pass | MIT | 3 mo ago |
| 67 | 《AI Engineering: Building Applications with Foundation Models》(Chip Huyen, O'Reilly 2025) 方法论。 | SpaceZephyr/ | 160 | — | ~1.2k | Automated safety check: Pass | No licence | 3 mo ago |
| 68 | Train or fine-tune SentenceTransformer bi-encoders, CrossEncoder rerankers, or SparseEncoder models, including losses, negatives, evaluation, distillation, LoRA, and Matryoshka. | waybarrios/ | 533 | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 69 | Guides users through LLM post-training with Training Hub, including installation, algorithm selection (SFT, OSFT, LoRA), hyperparameter tuning, troubleshooting OOM errors, interpreting loss curves… | Red-Hat-AI-Innovation-Team/ | 100 | — | ~2.8k | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 70 | A skill your agent uses when building, reviewing, or editing TRL post-training workflows for agentic applications, including SFT, DPO, GRPO, RLOO, reward modeling, dataset formats, chat templates… | burtenshaw/ | 153 | — | ~599 | Automated safety check: Pass | Apache-2.0 | 24 days ago |
| 71 | 71.Compiler Universal ARA Compiler. An agent skill from ARA-Labs/Agent-Native-Research-Artifact. | ARA-Labs/ | 690 | — | ~6.4k | Automated safety check: Pass | MIT | today |
| 72 | Checks training code, configs and math against documented framework behavior before an expensive run, citing a knowledge base or official docs for every claim. | Leeroo-AI/ | 195 | — | ~3.8k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 73 | Guide users through Cosmos3 supervised fine-tuning (SFT) post-training: preparing the example dataset and Wan2.2 VAE, converting the base checkpoint to DCP, launching distributed training (paired… | NVIDIA/ | 556 | — | ~2.7k | Automated safety check: Pass | Unknown | 11 days ago |
| 74 | Explains how to train a model to be harmless with self-critique, revision and AI-generated preference feedback, with Hugging Face and TRL code for each stage. | Orchestra-Research/ | 13k | 3 repos | ~2k | Automated safety check: Pass | MIT | 3 mo ago |
| 75 | Loads pre-trained Hugging Face Transformers models for text, vision and audio tasks, runs inference with pipelines and fine-tunes on custom datasets. | davila7/ | 32k | 12 repos | ~1.2k | Automated safety check: Pass | MIT | today |
| 76 | 76.Debug Bridge Drive and verify Knit on a device or emulator through the headless debug bridge (am broadcast to app.getknit.knit.debug.<ACTION, replies as JSON) — send a message on one phone and confirm it landed… | getknit/ | 130 | — | ~2.2k | Automated safety check: Pass | GPL-3.0 | 3 days ago |
| 77 | PyTorch training reference: architecture choice by data type, scaling rules, a training loop, optimizer and learning-rate choices, and fixes for loss spikes or OOM. | Orchestra-Research/ | 13k | 2 repos | ~2.8k | Automated safety check: Pass | MIT | 3 mo ago |
| 78 | Expert guidance for fine-tuning LLMs with LLaMA-Factory - WebUI no-code, 100+ models, 2/3/4/5/6/8-bit QLoRA, multimodal support | Orchestra-Research/ | 13k | 3 repos | ~618 | Automated safety check: Pass | MIT | 3 mo ago |
| 79 | Walks through nanoGPT, Karpathy's compact GPT implementation: training on Shakespeare, reproducing GPT-2, fine-tuning GPT-2 checkpoints and training on your own text. | Orchestra-Research/ | 13k | 3 repos | ~1.7k | Automated safety check: Pass | MIT | 3 mo ago |
| 80 | 80.Explore Code Rigor Improve implementation leaf skill for auditable candidate implementation in deep learning research repositories. | lllllllama/ | 497 | 1 repo | ~648 | Automated safety check: Pass | MIT | 14 days ago |
| 81 | Guides GRPO reinforcement-learning fine-tuning of language models with TRL, centered on designing reward functions for formats, verifiable tasks and reasoning. | Orchestra-Research/ | 13k | 4 repos | ~4.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 82 | Drives fine-tuning on Alibaba Cloud Model Studio with the bl CLI: validate and upload data, create a job, watch it, export results and deploy the model. | modelstudioai/ | 541 | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | 7 days ago |
| 83 | Orchestrates defect image generation for PCBA, metal surface and glass inspection with NVIDIA Cosmos AnomalyGen on OSMO, from cold-start Day 0 to real-photo Day 1 labeling. | NVIDIA/ | 3.5k | — | ~5k | Automated safety check: Notes | Apache-2.0 | today |
| 84 | Train custom TTS voices for Piper (ONNX format) using fine-tuning or from-scratch approaches. | sammcj/ | 162 | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 85 | Run the Nemotron-3.5 Lightning Text2SQL LoRA fine-tuning tutorial (NeMo Megatron-Bridge) end-to-end for the user on a single node: data prep, checkpoint conversion, LoRA fine-tuning of the 30B-A3B… | NVIDIA-NeMo/ | 2.1k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 86 | 86.Qwen21 Qwen-Image-2.1 LoRA line (NOT Anima) — running cache/train through the daemon, make gui-qwen, the CacheRequest/TrainRequest flag surface and how to add a field, model-dir resolution, cache layout… | sorryhyun/ | 125 | — | ~1.9k | Automated safety check: Notes | MIT | yesterday |
| 87 | A skill your agent uses when the user wants to estimate GPU memory (VRAM) requirements for a training configuration, check if a model will fit on their GPUs, or plan GPU allocation for training. | Red-Hat-AI-Innovation-Team/ | 100 | — | ~393 | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 88 | Reference for post-training language models with TRL: which trainer and dataset format to use for SFT, DPO, GRPO, KTO and reward models, and how to add LoRA. | huggingface/ | 11k | 1 repo | ~1.1k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 89 | Calculate training costs for Tinker fine-tuning jobs. An agent skill from sundial-org/skills. | sundial-org/ | 152 | — | ~1.2k | Automated safety check: Pass | No licence | 2 mo ago |
| 90 | Fine-tunes and evaluates OpenVLA-OFT and OFT+ robot policies with LoRA and continuous action heads on LIBERO simulation and ALOHA real-robot setups. | Orchestra-Research/ | 13k | 1 repo | ~3.7k | Automated safety check: Pass | MIT | 3 mo ago |
| 91 | Fine-tunes and serves Physical Intelligence's pi0, pi0-fast and pi0.5 robot policies with JAX or PyTorch, including checkpoint conversion and policy servers. | Orchestra-Research/ | 13k | 1 repo | ~3.6k | Automated safety check: Pass | MIT | 3 mo ago |
| 92 | The validated recipe for training MiniMax-Music3 planner-LM style adapters (artist/album clones) with ace-train mm3-lm-train and the Training Studio. | scragnog/ | 170 | — | ~4k | Automated safety check: Pass | MIT | 2 days ago |
| 93 | Run the Nemotron-3 Ultra Text2SQL LoRA fine-tuning tutorial (NeMo Megatron-Bridge) end-to-end for the user on their SLURM cluster: data prep, distributed checkpoint conversion, and packed LoRA… | NVIDIA-NeMo/ | 2.1k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 94 | 94.Train Lora Validate Draw Things LoRA training end to end with draw-things-cli, including tiny-dataset training, loss and scaler checks, checkpoint sanity, and base-versus-LoRA generation comparison. | drawthingsai/ | 579 | — | ~2k | Automated safety check: Pass | GPL-3.0 | yesterday |
| 95 | A skill your agent uses for BetterTogether, prompt plus weight optimization, fine-tuning sequences, and strategy chains like p - w - p. | OmidZamani/ | 124 | 1 repo | ~756 | Automated safety check: Pass | MIT | 3 mo ago |
| 96 | 96.Comfyui A skill your agent uses when working with ComfyUI workflows in OpenMontage, including comfyuiimage/comfyuivideo/comfyuimusic, custom workflowjson/workflowpath inputs, outputnode selection, missing… | calesthio/ | 65k | — | ~2k | Automated safety check: Pass | AGPL-3.0 | 4 days ago |
Explore related skills
More topics in AI & LLM Engineering
- Building AI agents525
- Deep learning408
- Embeddings381
- LLM inference and serving364
- Retrieval-augmented generation360
- Prompt engineering350
- LLM evaluation303
- Speech recognition and synthesis272
- Structured output and tool calling271
- LLM cost and token optimization256
- LLM API integration218
- LLM observability217
- Model routing and gateways217
- LLM guardrails208
- Computer vision206
- Model hubs and datasets180
- GPU and accelerator computing171
- Diffusion and image models167
- Natural language processing141
- Reinforcement learning67
- AI interpretability23