Search

Python · Fine-tuning

28 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Check Kiln's fine-tunable model list for deprecated or unsupported base models.

Kiln-AI/Kiln5.2k—~1.9kAutomated safety check: NotesUnknownyesterday
2

Trains and fine-tunes object detection, image classification and SAM or SAM2 segmentation models on Hugging Face Jobs cloud GPUs and saves the results to the Hub.

huggingface/skills11k1 repo~7.5kAutomated safety check: PassApache-2.02 days ago
3

Selectively monitor important long-running or resource-intensive commands with Haoleme by prefixing them with hao, so status, output, and completion notifications sync to the mobile app.

HaolemeApp/Haoleme157—~1.3kAutomated safety check: PassAGPL-3.01 mo ago
4

Train custom LoRAs with ostris AI-Toolkit. An agent skill from artokun/comfyui-mcp.

artokun/comfyui-mcp803—~2.7kAutomated safety check: PassMIT6 days ago
5

Quick-reference help for fine-tuning language models with Axolotl, covering YAML configs, FSDP, context parallelism, compressed saves and dataset formats.

Orchestra-Research/AI-Research-SKILLs13k8 repos~1.2kAutomated safety check: PassMIT3 mo ago
6

Fine-tune or transfer-learn AlphaGenome-PyTorch on custom genomic data — pick a mode (linear probe, LoRA, Locon, full), train on BigWig tracks with agt finetune, use adapters, delta checkpoints…

genomicsxai/alphagenome-pytorch162—~1kAutomated safety check: PassApache-2.025 days ago
7
7.Aqua CLIOfficial

Complete CLI reference for the ADS AQUA command-line interface (ads aqua).

oracle/accelerated-data-science125—~2.1kAutomated safety check: PassUPL-1.01 mo ago
8

A skill your agent uses when writing, reviewing, or debugging JAX code that involves quax — custom array-ish objects (physical units, LoRA, sparse, symbolic zero, named axes), quax.quaxify…

nstarman/quax143—~5.5kAutomated safety check: PassApache-2.0today
9

Walks through aligning language models with SimPO, a reference-free preference optimization method, using accelerate configs for Mistral 7B, Llama 3 8B and math-focused models.

Orchestra-Research/AI-Research-SKILLs13k4 repos~1.5kAutomated safety check: PassMIT3 mo ago
10
10.Aqua DeploymentOfficial

Deploy LLM models on OCI using AI Quick Actions (AQUA) - single model, multi-model, stacked (LoRA), with GPU shape selection, vLLM configuration, streaming, and tool calling.

oracle/accelerated-data-science125—~2.4kAutomated safety check: PassUPL-1.01 mo ago
11

Loads large language models in 8-bit or 4-bit with bitsandbytes so they fit smaller GPUs, and sets up QLoRA fine-tuning on a 4-bit base model.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2.5kAutomated safety check: PassMIT3 mo ago
12

Trains LLMs with reinforcement learning using verl, from ByteDance's Seed team, with GRPO, PPO and other algorithms and swappable training and rollout backends.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2.4kAutomated safety check: PassMIT3 mo ago
13

Loads pre-trained Hugging Face Transformers models for text, vision and audio tasks, runs inference with pipelines and fine-tunes on custom datasets.

davila7/claude-code-templates33k11 repos~1.2kAutomated safety check: PassMITtoday
14

Walks through nanoGPT, Karpathy's compact GPT implementation: training on Shakespeare, reproducing GPT-2, fine-tuning GPT-2 checkpoints and training on your own text.

Orchestra-Research/AI-Research-SKILLs13k2 repos~1.7kAutomated safety check: PassMIT3 mo ago
15

Reference for post-training language models with TRL: which trainer and dataset format to use for SFT, DPO, GRPO, KTO and reward models, and how to add LoRA.

huggingface/skills11k1 repo~1.1kAutomated safety check: PassApache-2.02 days ago
16

Connect this agent to a running Guaardvark (self-hosted AI studio) and check what it can do right now.

guaardvark/guaardvark258—~1.2kAutomated safety check: PassMITtoday
17
17.Aqua FinetuningOfficial

Fine-tune LLM models using LoRA on OCI AI Quick Actions (AQUA).

oracle/accelerated-data-science125—~1.7kAutomated safety check: PassUPL-1.01 mo ago
18

Fine-tunes and evaluates OpenVLA-OFT and OFT+ robot policies with LoRA and continuous action heads on LIBERO simulation and ALOHA real-robot setups.

Orchestra-Research/AI-Research-SKILLs13k—~3.7kAutomated safety check: PassMIT3 mo ago
19

Modal is a serverless cloud platform for running Python on demand, including on-demand GPUs.

K-Dense-AI/scientific-agent-skills48k1 repo~4.5kAutomated safety check: NotesApache-2.06 days ago
20

Hugging Face Transformers for loading Hub models, running pipeline inference, text generation, and Trainer fine-tuning on NLP, vision, audio, and multimodal tasks.

K-Dense-AI/scientific-agent-skills48k1 repo~2.8kAutomated safety check: NotesApache-2.06 days ago
21

Manages and orchestrates prompts in Agent Platform. An agent skill from google/skills.

google/skills21k—~2.2kAutomated safety check: PassApache-2.0yesterday
22

Manages GenAI tuning jobs in Agent Platform. An agent skill from google/skills.

google/skills21k—~1.9kAutomated safety check: PassApache-2.0yesterday
23

Run the disk-backed DEFT AOI improvement loop for NVIDIA Cosmos Reason 3 / Cosmos3 models, using Nano by default and Edge or Super when explicitly requested: evaluate the base model on Proxy and…

NVIDIA/skills3.6k—~5kAutomated safety check: NotesApache-2.0yesterday
24

Launch, relaunch, or sweep STANDARD (non-agentic) SkyRL RL on CINECA Leonardo — GRPO on math/reasoning datasets (gsm8k, MATH/aime) and on-policy distillation (OPD, teacher→student) — via raw sbatch…

open-thoughts/OpenThoughts-Agent301—~3.2kAutomated safety check: PassApache-2.012 days ago
25

Launch SFT via python -m hpc.launch --jobtype sft on any cluster (JSC Jupiter GH200, CINECA Leonardo A100, TACC Vista GH200), with EITHER backend — LLaMA-Factory (default) or axolotl (--sftbackend…

open-thoughts/OpenThoughts-Agent301—~2.9kAutomated safety check: PassApache-2.012 days ago
26

A skill your agent uses when running GPU compute on RunPod and deciding between Pods (hourly, always-on) and Serverless (per-second, autoscaling) for training, fine-tuning or inference — serverless…

ericrisco/rsc-harness180—~2.8kAutomated safety check: PassMITyesterday
27

A skill your agent uses when training or debugging a neural net in PyTorch — the forward/loss/backward/step loop and its silent bugs, mixed precision (AMP), AdamW/LR schedules, DDP/FSDP/ZeRO…

ericrisco/rsc-harness180—~3.4kAutomated safety check: PassMITyesterday
28

Operate, configure, benchmark, and troubleshoot llama.cpp across CPU, Metal, CUDA, HIP/ROCm, Vulkan, SYCL, and hybrid or multi-GPU systems.

magnus919/agent-skills115—~2.3kAutomated safety check: PassMITyesterday