Search

Fine-tuning

305 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Routes a sentence-transformers training task to the right model type and required reference docs and example scripts, covering bi-encoders, rerankers, sparse and multi-vector models.

huggingface/skills11k1 repo~2.6kAutomated safety check: PassApache-2.02 days ago
2

Fix a GitHub issue on OpenPipe/ART and open a PR. An agent skill from OpenPipe/ART.

OpenPipe/ART11k—~840Automated safety check: NotesApache-2.0today
3

RL training reference for the ART framework. An agent skill from OpenPipe/ART.

OpenPipe/ART11k—~2.4kAutomated safety check: PassApache-2.0today
4

Prepare, validate, launch-plan, monitor, resume, and stop configurable Qwopus 27B reinforcement-learning workflows for GRPO or GSPO.

R6410418/Jackrong-llm-finetuning-guide1.7k—~830Automated safety check: PassApache-2.03 mo ago
5

Validates dataset formatting and quality for SageMaker model fine-tuning (SFT, DPO, or RLVR).

awslabs/agent-plugins9161 repo~1.3kAutomated safety check: PassApache-2.0yesterday
6

SFT training reference for the ART framework. An agent skill from OpenPipe/ART.

OpenPipe/ART11k—~2.9kAutomated safety check: PassApache-2.0today
7

Fine-tune LLMs using reinforcement learning with TRL - SFT for instruction tuning, DPO for preference alignment, PPO/GRPO for reward optimization, and reward model training.

Orchestra-Research/AI-Research-SKILLs13k6 repos~2.9kAutomated safety check: PassMIT3 mo ago
8

Check Kiln's fine-tunable model list for deprecated or unsupported base models.

Kiln-AI/Kiln5.2k—~1.9kAutomated safety check: NotesUnknownyesterday
9

Guides building on the 0G Compute Network, a decentralized GPU marketplace for AI inference and fine-tuning, with SDK patterns and CLI commands.

internet-court/internet-court-skill6.6k1 repo~1.9kAutomated safety check: PassUnknown1 mo ago
10

A skill your agent uses when the user wants to optimize configurable system parameters against a measurable scalar objective, especially for model training, inference, quantitative strategies…

Optim-Agent/optim-agent801—~1.3kAutomated safety check: PassMIT1 mo ago
11

Generates code that transforms datasets between ML schemas for model training or evaluation.

awslabs/agent-plugins9161 repo~3.5kAutomated safety check: PassApache-2.0yesterday
12

Integrate a benchmark or custom environment into SAfactory using fixed adapter templates and local contract tests, optionally run Docker/RJob evaluation, or prepare GRPO/RL training.

AI45Lab/SAfactory236—~1.8kAutomated safety check: PassNo licence17 days ago
13

Trigger this skill when the user wants to train, fine-tune, or adapt Gemma models (e.g.

google-gemma/gemma-skills1k—~1.9kAutomated safety check: PassApache-2.04 days ago
14

Trains and evaluates several WiFi-signal-based pose and sensing models, from unsupervised pose estimation to domain adaptation and publishing.

ruvnet/RuView97k—~1.3kAutomated safety check: NotesMITtoday
15

Trains or fine-tunes language and vision models with TRL or Unsloth on Hugging Face Jobs cloud GPUs, then converts the results to GGUF.

huggingface/skills11k1 repo~7.2kAutomated safety check: PassApache-2.02 days ago
16

This skill should be used when picking or diagnosing a training move (SFT, LoRA, DPO/KTO/ORPO, RFT, GRPO/PPO/RLOO, RLHF), or when the user mentions fine-tuning, post-training, training recipe…

evo-hq/evo1.5k—~4.5kAutomated safety check: PassApache-2.06 days ago
17

Runs and configures the anomalib tiled-ensemble pipeline, which trains/evaluates one model per image tile and merges results (with optional seam smoothing) for high-resolution anomaly detection.

open-edge-platform/anomalib6.2k—~1.4kAutomated safety check: PassApache-2.0yesterday
18

Router for adding a diffusion or omni pipeline to verl-omni.

verl-project/verl-omni1.2k—~1kAutomated safety check: PassApache-2.0yesterday
19

Add a cross-cutting decision pattern under src/nemotron/steps/patterns/.

NVIDIA-NeMo/Nemotron2.1k—~1.4kAutomated safety check: PassApache-2.04 days ago
20

Parameter-efficient fine-tuning for LLMs using LoRA, QLoRA, and 25+ methods.

Orchestra-Research/AI-Research-SKILLs13k6 repos~3.1kAutomated safety check: PassMIT3 mo ago
21

Trains and fine-tunes object detection, image classification and SAM or SAM2 segmentation models on Hugging Face Jobs cloud GPUs and saves the results to the Hub.

huggingface/skills11k1 repo~7.5kAutomated safety check: PassApache-2.02 days ago
22

Guide for adding a new reward scorer to verl-omni and wiring it into a run.

verl-project/verl-omni1.2k—~648Automated safety check: PassApache-2.0yesterday
23

Toolbox for markerless animal pose estimation with DeepLabCut.

NeuroAIHub/BrainPilot1.1k—~1.7kAutomated safety check: PassAGPL-3.08 days ago
24

Add a new step under src/nemotron/steps/<category/<stepid/ — manifest (step.toml), runner glue, configs, and per-step README.md.

NVIDIA-NeMo/Nemotron2.1k—~1.7kAutomated safety check: PassApache-2.04 days ago
25

MLX Swift LM - Run LLMs and VLMs on Apple Silicon using MLX.

kellyvv/PhoneClaw1.3k—~3.7kAutomated safety check: PassApache-2.02 mo ago
26

Generate images with FLUX models (Black Forest Labs) via inference.sh CLI.

danielmeppiel/agentic-sdlc-handbook1581 repo~778Automated safety check: PassUnknown3 mo ago
27

Train or fine-tune language models with TRL or Unsloth on Hugging Face Jobs, including SFT, DPO, GRPO, reward models, and GGUF conversion.

waybarrios/opencode-power-pack534—~3kAutomated safety check: PassApache-2.05 days ago
28

Selectively monitor important long-running or resource-intensive commands with Haoleme by prefixing them with hao, so status, output, and completion notifications sync to the mobile app.

HaolemeApp/Haoleme157—~1.3kAutomated safety check: PassAGPL-3.01 mo ago
29

Lilly community-research skill. An agent skill from ssaaffaakk/Lilly.

ssaaffaakk/Lilly171—~1.4kAutomated safety check: PassMIT4 days ago
30

Audit imaging acquisition, reconstruction, series eligibility, quantitative transforms and protocol drift; not model training.

huang-sir1/radiology-skills1.9k—~2.8kAutomated safety check: PassUnknown20 days ago
31

Expert guidance for fast fine-tuning with Unsloth - 2-5x faster training, 50-80% less memory, LoRA/QLoRA optimization

Orchestra-Research/AI-Research-SKILLs13k10 repos~577Automated safety check: PassMIT3 mo ago
32

Train custom LoRAs with ostris AI-Toolkit. An agent skill from artokun/comfyui-mcp.

artokun/comfyui-mcp803—~2.7kAutomated safety check: PassMIT6 days ago
33

Configure and launch SparkDiffusion sparse finetuning for Wan 2.1 or Wan 2.2.

AlibabaResearch/SparkDiffusion542—~904Automated safety check: PassApache-2.02 days ago
34

A skill your agent uses whenever the user asks to generate, collect, inspect, or prepare early-experience training data (Implicit World Modeling or Self-Reflection, in the sense of arXiv:2510.08558)…

OSU-NLP-Group/EarlyExperience103—~4.1kAutomated safety check: PassMIT3 mo ago
35

Run, configure, retry, and validate AReno SFT, DPO, GSPO, GRPO, PPO, and agentic training.

inclusionAI/AReno323—~782Automated safety check: PassApache-2.0yesterday
36

A skill your agent uses when designing, implementing, reviewing, or debugging supervised fine-tuning with TRL SFTTrainer or trl sft, especially for agentic models trained on chat messages…

burtenshaw/training-agents153—~685Automated safety check: PassApache-2.028 days ago
37

Rigor Improve implementation leaf skill for auditable candidate implementation in deep learning research repositories.

lllllllama/RigorPilot-Skills4971 repo~648Automated safety check: PassMIT17 days ago
38

Train object-detection, image-classification, or SAM segmentation models on Hugging Face Jobs.

waybarrios/opencode-power-pack534—~2.7kAutomated safety check: PassApache-2.05 days ago
39

Selects a fine-tuning technique (SFT, DPO, RLVR, or RLAIF) for the user's use case and validates it against the selected model's available recipes.

awslabs/agent-plugins916—~604Automated safety check: PassApache-2.0yesterday
40

Quick-reference help for fine-tuning language models with Axolotl, covering YAML configs, FSDP, context parallelism, compressed saves and dataset formats.

Orchestra-Research/AI-Research-SKILLs13k8 repos~1.2kAutomated safety check: PassMIT3 mo ago
41

Build, deploy, evaluate, optimize, fine-tune, and manage Microsoft Foundry agents, models, and resources end to end.

microsoft/GitHub-Copilot-for-Azure2551 repo~6.7kAutomated safety check: PassMITyesterday
42

Set up the NVIDIA "Build an Agent" DevX workshop as a working JupyterLab environment from INSIDE a locked-down OpenShell/NemoClaw sandbox, and hand the user the token URL + access commands.

brevdev/workshop-build-an-agent146—~5.2kAutomated safety check: PassApache-2.03 days ago
43

Rules for adding a new model or model family to the finetuning pipeline, or changing finetuning behavior for an existing one — engine-agnostic customization via family hooks instead of if/else in…

overmind-core/overmind612—~3.2kAutomated safety check: PassAGPL-3.0today
44

Fine-tune or transfer-learn AlphaGenome-PyTorch on custom genomic data — pick a mode (linear probe, LoRA, Locon, full), train on BigWig tracks with agt finetune, use adapters, delta checkpoints…

genomicsxai/alphagenome-pytorch162—~1kAutomated safety check: PassApache-2.025 days ago
45
45.Aqua CLIOfficial

Complete CLI reference for the ADS AQUA command-line interface (ads aqua).

oracle/accelerated-data-science125—~2.1kAutomated safety check: PassUPL-1.01 mo ago
46
46.Cast

Build consistent characters, environments and props in Guaardvark's Cast Library and train LoRAs for them locally (reference photos → vision bible → sample plan → approved samples → training).

guaardvark/guaardvark258—~673Automated safety check: PassMITtoday
47
47.Quax

A skill your agent uses when writing, reviewing, or debugging JAX code that involves quax — custom array-ish objects (physical units, LoRA, sparse, symbolic zero, named axes), quax.quaxify…

nstarman/quax143—~5.5kAutomated safety check: PassApache-2.0today
48

Interact with a Meshtastic LoRa mesh network through MESH-API — list nodes, read messages, send texts, and check connection status.

mr-tbot/mesh-api180—~1.8kAutomated safety check: PassGPL-3.02 mo ago