Search
Fine-tuning
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Routes a sentence-transformers training task to the right model type and required reference docs and example scripts, covering bi-encoders, rerankers, sparse and multi-vector models. | huggingface/ | 11k | 1 repo | ~2.6k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 2 | Fix a GitHub issue on OpenPipe/ART and open a PR. An agent skill from OpenPipe/ART. | OpenPipe/ | 11k | — | ~840 | Automated safety check: Notes | Apache-2.0 | today |
| 3 | 3.Train Rl RL training reference for the ART framework. An agent skill from OpenPipe/ART. | OpenPipe/ | 11k | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | today |
| 4 | Prepare, validate, launch-plan, monitor, resume, and stop configurable Qwopus 27B reinforcement-learning workflows for GRPO or GSPO. | R6410418/ | 1.7k | — | ~830 | Automated safety check: Pass | Apache-2.0 | 3 mo ago |
| 5 | Validates dataset formatting and quality for SageMaker model fine-tuning (SFT, DPO, or RLVR). | awslabs/ | 916 | 1 repo | ~1.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 6 | SFT training reference for the ART framework. An agent skill from OpenPipe/ART. | OpenPipe/ | 11k | — | ~2.9k | Automated safety check: Pass | Apache-2.0 | today |
| 7 | Fine-tune LLMs using reinforcement learning with TRL - SFT for instruction tuning, DPO for preference alignment, PPO/GRPO for reward optimization, and reward model training. | Orchestra-Research/ | 13k | 6 repos | ~2.9k | Automated safety check: Pass | MIT | 3 mo ago |
| 8 | Check Kiln's fine-tunable model list for deprecated or unsupported base models. | Kiln-AI/ | 5.2k | — | ~1.9k | Automated safety check: Notes | Unknown | yesterday |
| 9 | Guides building on the 0G Compute Network, a decentralized GPU marketplace for AI inference and fine-tuning, with SDK patterns and CLI commands. | internet-court/ | 6.6k | 1 repo | ~1.9k | Automated safety check: Pass | Unknown | 1 mo ago |
| 10 | 10.Optim Agent A skill your agent uses when the user wants to optimize configurable system parameters against a measurable scalar objective, especially for model training, inference, quantitative strategies… | Optim-Agent/ | 801 | — | ~1.3k | Automated safety check: Pass | MIT | 1 mo ago |
| 11 | Generates code that transforms datasets between ML schemas for model training or evaluation. | awslabs/ | 916 | 1 repo | ~3.5k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 12 | Integrate a benchmark or custom environment into SAfactory using fixed adapter templates and local contract tests, optionally run Docker/RJob evaluation, or prepare GRPO/RL training. | AI45Lab/ | 236 | — | ~1.8k | Automated safety check: Pass | No licence | 17 days ago |
| 13 | Trigger this skill when the user wants to train, fine-tune, or adapt Gemma models (e.g. | google-gemma/ | 1k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | 4 days ago |
| 14 | Trains and evaluates several WiFi-signal-based pose and sensing models, from unsupervised pose estimation to domain adaptation and publishing. | ruvnet/ | 97k | — | ~1.3k | Automated safety check: Notes | MIT | today |
| 15 | Trains or fine-tunes language and vision models with TRL or Unsloth on Hugging Face Jobs cloud GPUs, then converts the results to GGUF. | huggingface/ | 11k | 1 repo | ~7.2k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 16 | 16.Finetuning This skill should be used when picking or diagnosing a training move (SFT, LoRA, DPO/KTO/ORPO, RFT, GRPO/PPO/RLOO, RLHF), or when the user mentions fine-tuning, post-training, training recipe… | evo-hq/ | 1.5k | — | ~4.5k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 17 | Runs and configures the anomalib tiled-ensemble pipeline, which trains/evaluates one model per image tile and merges results (with optional seam smoothing) for high-resolution anomaly detection. | open-edge-platform/ | 6.2k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 18 | 18.Add Pipeline Router for adding a diffusion or omni pipeline to verl-omni. | verl-project/ | 1.2k | — | ~1k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 19 | Add a cross-cutting decision pattern under src/nemotron/steps/patterns/. | NVIDIA-NeMo/ | 2.1k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | 4 days ago |
| 20 | Parameter-efficient fine-tuning for LLMs using LoRA, QLoRA, and 25+ methods. | Orchestra-Research/ | 13k | 6 repos | ~3.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 21 | Trains and fine-tunes object detection, image classification and SAM or SAM2 segmentation models on Hugging Face Jobs cloud GPUs and saves the results to the Hub. | huggingface/ | 11k | 1 repo | ~7.5k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 22 | Guide for adding a new reward scorer to verl-omni and wiring it into a run. | verl-project/ | 1.2k | — | ~648 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 23 | 23.Deeplabcut Toolbox for markerless animal pose estimation with DeepLabCut. | NeuroAIHub/ | 1.1k | — | ~1.7k | Automated safety check: Pass | AGPL-3.0 | 8 days ago |
| 24 | Add a new step under src/nemotron/steps/<category/<stepid/ — manifest (step.toml), runner glue, configs, and per-step README.md. | NVIDIA-NeMo/ | 2.1k | — | ~1.7k | Automated safety check: Pass | Apache-2.0 | 4 days ago |
| 25 | 25.Swift Mlx Lm MLX Swift LM - Run LLMs and VLMs on Apple Silicon using MLX. | kellyvv/ | 1.3k | — | ~3.7k | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 26 | 26.Flux Image Generate images with FLUX models (Black Forest Labs) via inference.sh CLI. | danielmeppiel/ | 158 | 1 repo | ~778 | Automated safety check: Pass | Unknown | 3 mo ago |
| 27 | Train or fine-tune language models with TRL or Unsloth on Hugging Face Jobs, including SFT, DPO, GRPO, reward models, and GGUF conversion. | waybarrios/ | 534 | — | ~3k | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 28 | Selectively monitor important long-running or resource-intensive commands with Haoleme by prefixing them with hao, so status, output, and completion notifications sync to the mobile app. | HaolemeApp/ | 157 | — | ~1.3k | Automated safety check: Pass | AGPL-3.0 | 1 mo ago |
| 29 | Lilly community-research skill. An agent skill from ssaaffaakk/Lilly. | ssaaffaakk/ | 171 | — | ~1.4k | Automated safety check: Pass | MIT | 4 days ago |
| 30 | Audit imaging acquisition, reconstruction, series eligibility, quantitative transforms and protocol drift; not model training. | huang-sir1/ | 1.9k | — | ~2.8k | Automated safety check: Pass | Unknown | 20 days ago |
| 31 | Expert guidance for fast fine-tuning with Unsloth - 2-5x faster training, 50-80% less memory, LoRA/QLoRA optimization | Orchestra-Research/ | 13k | 10 repos | ~577 | Automated safety check: Pass | MIT | 3 mo ago |
| 32 | Train custom LoRAs with ostris AI-Toolkit. An agent skill from artokun/comfyui-mcp. | artokun/ | 803 | — | ~2.7k | Automated safety check: Pass | MIT | 6 days ago |
| 33 | Configure and launch SparkDiffusion sparse finetuning for Wan 2.1 or Wan 2.2. | AlibabaResearch/ | 542 | — | ~904 | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 34 | A skill your agent uses whenever the user asks to generate, collect, inspect, or prepare early-experience training data (Implicit World Modeling or Self-Reflection, in the sense of arXiv:2510.08558)… | OSU-NLP-Group/ | 103 | — | ~4.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 35 | Run, configure, retry, and validate AReno SFT, DPO, GSPO, GRPO, PPO, and agentic training. | inclusionAI/ | 323 | — | ~782 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 36 | 36.Trl Sft A skill your agent uses when designing, implementing, reviewing, or debugging supervised fine-tuning with TRL SFTTrainer or trl sft, especially for agentic models trained on chat messages… | burtenshaw/ | 153 | — | ~685 | Automated safety check: Pass | Apache-2.0 | 28 days ago |
| 37 | 37.Explore Code Rigor Improve implementation leaf skill for auditable candidate implementation in deep learning research repositories. | lllllllama/ | 497 | 1 repo | ~648 | Automated safety check: Pass | MIT | 17 days ago |
| 38 | Train object-detection, image-classification, or SAM segmentation models on Hugging Face Jobs. | waybarrios/ | 534 | — | ~2.7k | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 39 | Selects a fine-tuning technique (SFT, DPO, RLVR, or RLAIF) for the user's use case and validates it against the selected model's available recipes. | awslabs/ | 916 | — | ~604 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 40 | Quick-reference help for fine-tuning language models with Axolotl, covering YAML configs, FSDP, context parallelism, compressed saves and dataset formats. | Orchestra-Research/ | 13k | 8 repos | ~1.2k | Automated safety check: Pass | MIT | 3 mo ago |
| 41 | Build, deploy, evaluate, optimize, fine-tune, and manage Microsoft Foundry agents, models, and resources end to end. | microsoft/ | 255 | 1 repo | ~6.7k | Automated safety check: Pass | MIT | yesterday |
| 42 | Set up the NVIDIA "Build an Agent" DevX workshop as a working JupyterLab environment from INSIDE a locked-down OpenShell/NemoClaw sandbox, and hand the user the token URL + access commands. | brevdev/ | 146 | — | ~5.2k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 43 | Rules for adding a new model or model family to the finetuning pipeline, or changing finetuning behavior for an existing one — engine-agnostic customization via family hooks instead of if/else in… | overmind-core/ | 612 | — | ~3.2k | Automated safety check: Pass | AGPL-3.0 | today |
| 44 | Fine-tune or transfer-learn AlphaGenome-PyTorch on custom genomic data — pick a mode (linear probe, LoRA, Locon, full), train on BigWig tracks with agt finetune, use adapters, delta checkpoints… | genomicsxai/ | 162 | — | ~1k | Automated safety check: Pass | Apache-2.0 | 25 days ago |
| 45 | Complete CLI reference for the ADS AQUA command-line interface (ads aqua). | oracle/ | 125 | — | ~2.1k | Automated safety check: Pass | UPL-1.0 | 1 mo ago |
| 46 | 46.Cast Build consistent characters, environments and props in Guaardvark's Cast Library and train LoRAs for them locally (reference photos → vision bible → sample plan → approved samples → training). | guaardvark/ | 258 | — | ~673 | Automated safety check: Pass | MIT | today |
| 47 | 47.Quax A skill your agent uses when writing, reviewing, or debugging JAX code that involves quax — custom array-ish objects (physical units, LoRA, sparse, symbolic zero, named axes), quax.quaxify… | nstarman/ | 143 | — | ~5.5k | Automated safety check: Pass | Apache-2.0 | today |
| 48 | 48.Mesh API Interact with a Meshtastic LoRa mesh network through MESH-API — list nodes, read messages, send texts, and check connection status. | mr-tbot/ | 180 | — | ~1.8k | Automated safety check: Pass | GPL-3.0 | 2 mo ago |