Topic · AI & LLM Engineering
Best fine-tuning skills, page 3
Fine-tuning skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 97 | Fine-tune LLM models using LoRA on OCI AI Quick Actions (AQUA). | oracle/ | 125 | — | ~1.7k | Automated safety check: Pass | UPL-1.0 | 1 mo ago |
| 98 | Use sub-agents as test subjects to iteratively improve .claude/ instruction files. | adam-s/ | 189 | — | ~4.6k | Automated safety check: Pass | MIT | 2 mo ago |
| 99 | Reference desk for NVIDIA Nemotron 3 Ultra (550B-A55B) — architecture, NVFP4 pretraining, SFT, MOPD (multi-teacher on-policy distillation), MTP boosting, quantization, inference. | NVIDIA-NeMo/ | 2.1k | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 100 | 100.Evo Memory Manages persistent research memory across ideation and experimentation cycles. | EvoScientist/ | 474 | 3 repos | ~4.8k | Automated safety check: Pass | Apache-2.0 | 7 days ago |
| 101 | Add or tighten Draw Things LoRA trainer support for generative models available in the Draw Things app / CLI, covering LoRA builders, trainer dispatch, tokenizer and fixed-encoder wiring, checkpoint… | drawthingsai/ | 579 | — | ~2.6k | Automated safety check: Pass | GPL-3.0 | yesterday |
| 102 | A skill your agent uses when designing, reviewing, or implementing OpenEnv-style environment interfaces for agentic RL with TRL, including reset/step/state contracts, tasksets, Docker or… | burtenshaw/ | 153 | — | ~381 | Automated safety check: Pass | Apache-2.0 | 24 days ago |
| 103 | 103.Research Research workflow for distilling ML/AI papers into modelless inference primitives, freeze/thaw runtime patterns, latent-space operations, AND model-based training plans across the multi-repo stack. | katopz/ | 134 | — | ~19k | Automated safety check: Pass | MIT | today |
| 104 | Audit supervised fine-tuning datasets against the behavior and task they are meant to teach. | tokenbender/ | 367 | — | ~2.7k | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 105 | 105.02 Rl Reward Build RL reward signals using the OpenJudge framework. An agent skill from agentscope-ai/OpenJudge. | agentscope-ai/ | 867 | 1 repo | ~1.6k | Automated safety check: Pass | Apache-2.0 | 26 days ago |
| 106 | Build an end-to-end UAV object detection and telemetry overlay application on Intel hardware using DL Streamer Pipeline Server with MAVLink telemetry. | open-edge-platform/ | 140 | — | ~2.6k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 107 | Builds or modifies ComfyUI workflow JSON templates in MooshieUI's Rust backend (src-tauri/src/templates). | Mooshieblob1/ | 207 | — | ~640 | Automated safety check: Pass | AGPL-3.0 | 2 days ago |
| 108 | 108.Slime User Guide for using SLIME (LLM post-training framework for RL Scaling). | yzlnew/ | 149 | — | ~3.2k | Automated safety check: Pass | No licence | 3 mo ago |
| 109 | Oracle Enterprise AI guidance for building, deploying, securing, estimating cost for, and integrating AI models, agents, RAG, Responses API workflows, custom or imported models, fine-tuning, model… | oracle/ | 872 | — | ~1.4k | Automated safety check: Pass | UPL-1.0 | yesterday |
| 110 | Agent skill for sona-learning-optimizer - invoke with $agent-sona-learning-optimizer | ruvnet/ | 74k | 3 repos | ~516 | Automated safety check: Pass | MIT | yesterday |
| 111 | 111.Lora Manager Author ComfyUI-LoRA-Manager nodes from the panel. An agent skill from artokun/comfyui-mcp. | artokun/ | 793 | — | ~726 | Automated safety check: Pass | MIT | 2 days ago |
| 112 | Fine-tune models on Microsoft Foundry using SFT (supervised), DPO (preference), or RFT (reinforcement with graders). | microsoft/ | 255 | 1 repo | ~1.4k | Automated safety check: Pass | MIT | yesterday |
| 113 | 113.Hermes Traj Capture Claude Code interaction trajectories in training-friendly formats. | AlexAI-MCP/ | 135 | — | ~1.6k | Automated safety check: Pass | MIT | 6 mo ago |
| 114 | 114.Model Training Agent-driven YOLO fine-tuning — annotate, train, export, deploy | SharpAI/ | 3.1k | — | ~985 | Automated safety check: Pass | MIT | 21 days ago |
| 115 | Generates code that fine-tunes a base model using SageMaker serverless training jobs. | awslabs/ | 912 | 1 repo | ~2.3k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 116 | 116.Comfyui Manager Manage ComfyUI server, models, workflows, LoRAs, queues, dependencies and CLI workflow execution via comfyui-skill. | ShiroEirin/ | 476 | — | ~5.2k | Automated safety check: Pass | GPL-3.0 | 3 mo ago |
| 117 | 117.Arbor Applies Arbor Hypothesis Tree Refinement to research artifacts with repeatable evaluators, including model training, agent harnesses, data synthesis and benchmark optimization. | K-Dense-AI/ | 48k | 1 repo | ~4.3k | Automated safety check: Notes | MIT | 2 days ago |
| 118 | 118.Modal Modal is a serverless cloud platform for running Python on demand, including on-demand GPUs. | K-Dense-AI/ | 48k | 1 repo | ~4.5k | Automated safety check: Notes | Apache-2.0 | 2 days ago |
| 119 | 119.Waypoint Bio Supports work with Outpost Bio's open microbiome foundation models - the Waypoint checkpoints (Waypoint-6m, Waypoint-45m, Waypoint-170m), the Atlas pretraining corpus, the Compass eight-task… | K-Dense-AI/ | 48k | 1 repo | ~4.2k | Automated safety check: Pass | MIT | 2 days ago |
| 120 | Guidelines for creating high-quality datasets for LLM post-training (SFT/DPO/RLHF). | sundial-org/ | 152 | — | ~1.4k | Automated safety check: Pass | No licence | 2 mo ago |
| 121 | 121.Trl Training Train and fine-tune transformer language models using TRL (Transformers Reinforcement Learning). | waybarrios/ | 533 | 3 repos | ~2.1k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 122 | Extract arbitrary, custom entity types from clinical or biomedical text with no fine-tuning using OpenMed's GLiNER / GLiNER2 zero-shot support. | maziyarpanahi/ | 5.5k | — | ~1.7k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 123 | 123.LLM Finetuning LLM fine-tuning expert for LoRA, QLoRA, dataset preparation, and training optimization | RightNow-AI/ | 18k | — | ~986 | Automated safety check: Pass | Apache-2.0 | 3 mo ago |
| 124 | 124.NLP Alignment Best practices for LLM alignment techniques including RLHF, DPO, and instruction tuning. | aiming-lab/ | 15k | — | ~300 | Automated safety check: Pass | MIT | 1 mo ago |
| 125 | 125.NLP Pretraining Best practices for language model pretraining and fine-tuning. | aiming-lab/ | 15k | — | ~280 | Automated safety check: Pass | MIT | 1 mo ago |
| 126 | Generates code that deploys fine-tuned models from SageMaker Serverless Model Customization to SageMaker endpoints or Bedrock. | awslabs/ | 912 | 1 repo | ~1.5k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 127 | 127.LLM Ops LLM Operations -- RAG, embeddings, vector databases, fine-tuning, prompt engineering avancado, custos de LLM, evals de qualidade e arquiteturas de IA para producao. | davila7/ | 32k | 4 repos | ~2k | Automated safety check: Pass | MIT | yesterday |
| 128 | Train or fine-tune language and vision models using TRL (Transformer Reinforcement Learning) or Unsloth with Hugging Face Jobs infrastructure. | sickn33/ | 47k | 1 repo | ~1.1k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 129 | 129.Transformers Hugging Face Transformers for loading Hub models, running pipeline inference, text generation, and Trainer fine-tuning on NLP, vision, audio, and multimodal tasks. | K-Dense-AI/ | 48k | 1 repo | ~2.8k | Automated safety check: Notes | Apache-2.0 | 2 days ago |
| 130 | 130.Rgthree Configure and author rgthree-comfy nodes — Fast Groups Bypasser/Muter (group toggles), Power Lora Loader, Context/Context Big, Seed, Any Switch. | artokun/ | 793 | — | ~3.3k | Automated safety check: Pass | MIT | 2 days ago |
| 131 | 131.Scenario A skill your agent uses when connecting an AI agent to Scenario (scenario.com) through MCP, or when a task involves generating images, video, 3D, audio, sprites, textures, or game assets. | scenario-labs/ | 898 | 1 repo | ~6.2k | Automated safety check: Pass | MIT | yesterday |
| 132 | A skill your agent uses when one look must hold across Scenario generations: one character across scenes, a turnaround, or a video animated from its references, one product across angles, one style… | scenario-labs/ | 898 | 1 repo | ~3.7k | Automated safety check: Pass | MIT | yesterday |
| 133 | Discovers user intent and generates a structured, step-by-step plan for model customization workflows. | awslabs/ | 912 | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 134 | Fine-tune a DPA3 model in DeePMD-kit using the PyTorch backend. | jinzhezenggroup/ | 148 | 1 repo | ~3.1k | Automated safety check: Pass | LGPL-3.0-or-later | 2 days ago |
| 135 | 135.Llamafactory Fine-tune LLMs with LlamaFactory — register datasets, train via YAML configs, merge LoRA adapters and serve the result. | Prism-Shadow/ | 2.5k | — | ~855 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 136 | 136.Dataset Curation Prepare, format, and validate datasets for supervised fine-tuning and preference training. | wshobson/ | 40k | — | ~2k | Automated safety check: Pass | MIT | 3 days ago |
| 137 | Build the evaluation harness that gates every fine-tuning run — golden sets, per-failure-mode graders, judge calibration, and base-model baselines. | wshobson/ | 40k | — | ~2k | Automated safety check: Pass | MIT | 3 days ago |
| 138 | Decide whether to fine-tune at all, and route to the right method (SFT, DPO/ORPO/KTO, GRPO/RLVR, continued pretraining) and base model. | wshobson/ | 40k | — | ~2k | Automated safety check: Pass | MIT | 3 days ago |
| 139 | Train reasoning and verifiable-task behavior with GRPO and reinforcement learning from verifiable rewards (RLVR). | wshobson/ | 40k | — | ~1.9k | Automated safety check: Pass | MIT | 3 days ago |
| 140 | Configure LoRA and QLoRA supervised fine-tuning with current best-practice hyperparameters. | wshobson/ | 40k | — | ~1.9k | Automated safety check: Pass | MIT | 3 days ago |
| 141 | Align a fine-tuned model with preference data using DPO, ORPO, KTO, or SimPO. | wshobson/ | 40k | — | ~2k | Automated safety check: Pass | MIT | 3 days ago |
| 142 | Set up a working ML training/inference environment on NVIDIA DGX Spark (GB10, aarch64, CUDA 13). | wshobson/ | 40k | — | ~2k | Automated safety check: Pass | MIT | 3 days ago |
| 143 | 143.Vision Sft Fine-tune vision-language models (VLMs) with supervised learning on image+text data. | wshobson/ | 40k | — | ~2k | Automated safety check: Pass | MIT | 3 days ago |
| 144 | Chief AI Officer advisory for startups: model build-vs-buy decisions (API vs fine-tune vs in-house), AI risk classification under EU AI Act + US state patchwork, AI cost economics… | alirezarezvani/ | 28k | — | ~3.5k | Automated safety check: Pass | MIT | 1 mo ago |
Explore related skills
Category
More topics in AI & LLM Engineering
- Building AI agents525
- Deep learning408
- Embeddings381
- LLM inference and serving364
- Retrieval-augmented generation360
- Prompt engineering350
- LLM evaluation303
- Speech recognition and synthesis272
- Structured output and tool calling271
- LLM cost and token optimization256
- LLM API integration218
- LLM observability217
- Model routing and gateways217
- LLM guardrails208
- Computer vision206
- Model hubs and datasets180
- GPU and accelerator computing171
- Diffusion and image models167
- Natural language processing141
- Reinforcement learning67
- AI interpretability23