Topic · AI & LLM Engineering

Best fine-tuning skills, page 3

Skills #97–144 of 313, ranked by score.

Fine-tuning skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Fine-tuning skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
97
97.Aqua FinetuningOfficial

Fine-tune LLM models using LoRA on OCI AI Quick Actions (AQUA).

oracle/accelerated-data-science125—~1.7kAutomated safety check: PassUPL-1.01 mo ago
98

Use sub-agents as test subjects to iteratively improve .claude/ instruction files.

adam-s/intercept189—~4.6kAutomated safety check: PassMIT2 mo ago
99

Reference desk for NVIDIA Nemotron 3 Ultra (550B-A55B) — architecture, NVFP4 pretraining, SFT, MOPD (multi-teacher on-policy distillation), MTP boosting, quantization, inference.

NVIDIA-NeMo/Nemotron2.1k—~1.8kAutomated safety check: PassApache-2.0yesterday
100

Manages persistent research memory across ideation and experimentation cycles.

EvoScientist/EvoSkills4743 repos~4.8kAutomated safety check: PassApache-2.07 days ago
101

Add or tighten Draw Things LoRA trainer support for generative models available in the Draw Things app / CLI, covering LoRA builders, trainer dispatch, tokenizer and fixed-encoder wiring, checkpoint…

drawthingsai/draw-things-community579—~2.6kAutomated safety check: PassGPL-3.0yesterday
102

A skill your agent uses when designing, reviewing, or implementing OpenEnv-style environment interfaces for agentic RL with TRL, including reset/step/state contracts, tasksets, Docker or…

burtenshaw/training-agents153—~381Automated safety check: PassApache-2.024 days ago
103

Research workflow for distilling ML/AI papers into modelless inference primitives, freeze/thaw runtime patterns, latent-space operations, AND model-based training plans across the multi-repo stack.

katopz/katgpt-rs134—~19kAutomated safety check: PassMITtoday
104

Audit supervised fine-tuning datasets against the behavior and task they are meant to teach.

tokenbender/agent-guides367—~2.7kAutomated safety check: PassApache-2.02 mo ago
105

Build RL reward signals using the OpenJudge framework. An agent skill from agentscope-ai/OpenJudge.

agentscope-ai/OpenJudge8671 repo~1.6kAutomated safety check: PassApache-2.026 days ago
106

Build an end-to-end UAV object detection and telemetry overlay application on Intel hardware using DL Streamer Pipeline Server with MAVLink telemetry.

open-edge-platform/edge-ai-suites140—~2.6kAutomated safety check: NotesApache-2.0yesterday
107

Builds or modifies ComfyUI workflow JSON templates in MooshieUI's Rust backend (src-tauri/src/templates).

Mooshieblob1/MooshieUI207—~640Automated safety check: PassAGPL-3.02 days ago
108

Guide for using SLIME (LLM post-training framework for RL Scaling).

yzlnew/infra-skills149—~3.2kAutomated safety check: PassNo licence3 mo ago
109
109.Enterprise AIOfficial

Oracle Enterprise AI guidance for building, deploying, securing, estimating cost for, and integrating AI models, agents, RAG, Responses API workflows, custom or imported models, fine-tuning, model…

oracle/skills872—~1.4kAutomated safety check: PassUPL-1.0yesterday
110

Agent skill for sona-learning-optimizer - invoke with $agent-sona-learning-optimizer

ruvnet/ruflo74k3 repos~516Automated safety check: PassMITyesterday
111

Author ComfyUI-LoRA-Manager nodes from the panel. An agent skill from artokun/comfyui-mcp.

artokun/comfyui-mcp793—~726Automated safety check: PassMIT2 days ago
112
112.FinetuningOfficial

Fine-tune models on Microsoft Foundry using SFT (supervised), DPO (preference), or RFT (reinforcement with graders).

microsoft/GitHub-Copilot-for-Azure2551 repo~1.4kAutomated safety check: PassMITyesterday
113

Capture Claude Code interaction trajectories in training-friendly formats.

AlexAI-MCP/hermes-CCC135—~1.6kAutomated safety check: PassMIT6 mo ago
114

Agent-driven YOLO fine-tuning — annotate, train, export, deploy

SharpAI/DeepCamera3.1k—~985Automated safety check: PassMIT21 days ago
115
115.FinetuningOfficial

Generates code that fine-tunes a base model using SageMaker serverless training jobs.

awslabs/agent-plugins9121 repo~2.3kAutomated safety check: PassApache-2.02 days ago
116

Manage ComfyUI server, models, workflows, LoRAs, queues, dependencies and CLI workflow execution via comfyui-skill.

ShiroEirin/comfyui-good-anima476—~5.2kAutomated safety check: PassGPL-3.03 mo ago
117
117.Arbor

Applies Arbor Hypothesis Tree Refinement to research artifacts with repeatable evaluators, including model training, agent harnesses, data synthesis and benchmark optimization.

K-Dense-AI/scientific-agent-skills48k1 repo~4.3kAutomated safety check: NotesMIT2 days ago
118
118.Modal

Modal is a serverless cloud platform for running Python on demand, including on-demand GPUs.

K-Dense-AI/scientific-agent-skills48k1 repo~4.5kAutomated safety check: NotesApache-2.02 days ago
119

Supports work with Outpost Bio's open microbiome foundation models - the Waypoint checkpoints (Waypoint-6m, Waypoint-45m, Waypoint-170m), the Atlas pretraining corpus, the Compass eight-task…

K-Dense-AI/scientific-agent-skills48k1 repo~4.2kAutomated safety check: PassMIT2 days ago
120

Guidelines for creating high-quality datasets for LLM post-training (SFT/DPO/RLHF).

sundial-org/skills152—~1.4kAutomated safety check: PassNo licence2 mo ago
121

Train and fine-tune transformer language models using TRL (Transformers Reinforcement Learning).

waybarrios/opencode-power-pack5333 repos~2.1kAutomated safety check: PassApache-2.02 days ago
122

Extract arbitrary, custom entity types from clinical or biomedical text with no fine-tuning using OpenMed's GLiNER / GLiNER2 zero-shot support.

maziyarpanahi/openmed5.5k—~1.7kAutomated safety check: PassApache-2.02 days ago
123

LLM fine-tuning expert for LoRA, QLoRA, dataset preparation, and training optimization

RightNow-AI/openfang18k—~986Automated safety check: PassApache-2.03 mo ago
124

Best practices for LLM alignment techniques including RLHF, DPO, and instruction tuning.

aiming-lab/AutoResearchClaw15k—~300Automated safety check: PassMIT1 mo ago
125

Best practices for language model pretraining and fine-tuning.

aiming-lab/AutoResearchClaw15k—~280Automated safety check: PassMIT1 mo ago
126
126.Model DeploymentOfficial

Generates code that deploys fine-tuned models from SageMaker Serverless Model Customization to SageMaker endpoints or Bedrock.

awslabs/agent-plugins9121 repo~1.5kAutomated safety check: PassApache-2.02 days ago
127

LLM Operations -- RAG, embeddings, vector databases, fine-tuning, prompt engineering avancado, custos de LLM, evals de qualidade e arquiteturas de IA para producao.

davila7/claude-code-templates32k4 repos~2kAutomated safety check: PassMITyesterday
128

Train or fine-tune language and vision models using TRL (Transformer Reinforcement Learning) or Unsloth with Hugging Face Jobs infrastructure.

sickn33/agentic-awesome-skills47k1 repo~1.1kAutomated safety check: PassApache-2.0yesterday
129

Hugging Face Transformers for loading Hub models, running pipeline inference, text generation, and Trainer fine-tuning on NLP, vision, audio, and multimodal tasks.

K-Dense-AI/scientific-agent-skills48k1 repo~2.8kAutomated safety check: NotesApache-2.02 days ago
130

Configure and author rgthree-comfy nodes — Fast Groups Bypasser/Muter (group toggles), Power Lora Loader, Context/Context Big, Seed, Any Switch.

artokun/comfyui-mcp793—~3.3kAutomated safety check: PassMIT2 days ago
131

A skill your agent uses when connecting an AI agent to Scenario (scenario.com) through MCP, or when a task involves generating images, video, 3D, audio, sprites, textures, or game assets.

scenario-labs/skills8981 repo~6.2kAutomated safety check: PassMITyesterday
132

A skill your agent uses when one look must hold across Scenario generations: one character across scenes, a turnaround, or a video animated from its references, one product across angles, one style…

scenario-labs/skills8981 repo~3.7kAutomated safety check: PassMITyesterday
133
133.PlanningOfficial

Discovers user intent and generates a structured, step-by-step plan for model customization workflows.

awslabs/agent-plugins912—~1.8kAutomated safety check: PassApache-2.02 days ago
134

Fine-tune a DPA3 model in DeePMD-kit using the PyTorch backend.

jinzhezenggroup/computational-chemistry-agent-skills1481 repo~3.1kAutomated safety check: PassLGPL-3.0-or-later2 days ago
135

Fine-tune LLMs with LlamaFactory — register datasets, train via YAML configs, merge LoRA adapters and serve the result.

Prism-Shadow/penguin-harness2.5k—~855Automated safety check: PassApache-2.0yesterday
136

Prepare, format, and validate datasets for supervised fine-tuning and preference training.

wshobson/agents40k—~2kAutomated safety check: PassMIT3 days ago
137

Build the evaluation harness that gates every fine-tuning run — golden sets, per-failure-mode graders, judge calibration, and base-model baselines.

wshobson/agents40k—~2kAutomated safety check: PassMIT3 days ago
138

Decide whether to fine-tune at all, and route to the right method (SFT, DPO/ORPO/KTO, GRPO/RLVR, continued pretraining) and base model.

wshobson/agents40k—~2kAutomated safety check: PassMIT3 days ago
139

Train reasoning and verifiable-task behavior with GRPO and reinforcement learning from verifiable rewards (RLVR).

wshobson/agents40k—~1.9kAutomated safety check: PassMIT3 days ago
140

Configure LoRA and QLoRA supervised fine-tuning with current best-practice hyperparameters.

wshobson/agents40k—~1.9kAutomated safety check: PassMIT3 days ago
141

Align a fine-tuned model with preference data using DPO, ORPO, KTO, or SimPO.

wshobson/agents40k—~2kAutomated safety check: PassMIT3 days ago
142

Set up a working ML training/inference environment on NVIDIA DGX Spark (GB10, aarch64, CUDA 13).

wshobson/agents40k—~2kAutomated safety check: PassMIT3 days ago
143

Fine-tune vision-language models (VLMs) with supervised learning on image+text data.

wshobson/agents40k—~2kAutomated safety check: PassMIT3 days ago
144

Chief AI Officer advisory for startups: model build-vs-buy decisions (API vs fine-tune vs in-house), AI risk classification under EU AI Act + US state patchwork, AI cost economics…

alirezarezvani/claude-skills28k—~3.5kAutomated safety check: PassMIT1 mo ago