Search
Python · Fine-tuning
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Check Kiln's fine-tunable model list for deprecated or unsupported base models. | Kiln-AI/ | 5.2k | — | ~1.9k | Automated safety check: Notes | Unknown | yesterday |
| 2 | Trains and fine-tunes object detection, image classification and SAM or SAM2 segmentation models on Hugging Face Jobs cloud GPUs and saves the results to the Hub. | huggingface/ | 11k | 1 repo | ~7.5k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 3 | Selectively monitor important long-running or resource-intensive commands with Haoleme by prefixing them with hao, so status, output, and completion notifications sync to the mobile app. | HaolemeApp/ | 157 | — | ~1.3k | Automated safety check: Pass | AGPL-3.0 | 1 mo ago |
| 4 | Train custom LoRAs with ostris AI-Toolkit. An agent skill from artokun/comfyui-mcp. | artokun/ | 803 | — | ~2.7k | Automated safety check: Pass | MIT | 6 days ago |
| 5 | Quick-reference help for fine-tuning language models with Axolotl, covering YAML configs, FSDP, context parallelism, compressed saves and dataset formats. | Orchestra-Research/ | 13k | 8 repos | ~1.2k | Automated safety check: Pass | MIT | 3 mo ago |
| 6 | Fine-tune or transfer-learn AlphaGenome-PyTorch on custom genomic data — pick a mode (linear probe, LoRA, Locon, full), train on BigWig tracks with agt finetune, use adapters, delta checkpoints… | genomicsxai/ | 162 | — | ~1k | Automated safety check: Pass | Apache-2.0 | 25 days ago |
| 7 | Complete CLI reference for the ADS AQUA command-line interface (ads aqua). | oracle/ | 125 | — | ~2.1k | Automated safety check: Pass | UPL-1.0 | 1 mo ago |
| 8 | 8.Quax A skill your agent uses when writing, reviewing, or debugging JAX code that involves quax — custom array-ish objects (physical units, LoRA, sparse, symbolic zero, named axes), quax.quaxify… | nstarman/ | 143 | — | ~5.5k | Automated safety check: Pass | Apache-2.0 | today |
| 9 | Walks through aligning language models with SimPO, a reference-free preference optimization method, using accelerate configs for Mistral 7B, Llama 3 8B and math-focused models. | Orchestra-Research/ | 13k | 4 repos | ~1.5k | Automated safety check: Pass | MIT | 3 mo ago |
| 10 | Deploy LLM models on OCI using AI Quick Actions (AQUA) - single model, multi-model, stacked (LoRA), with GPU shape selection, vLLM configuration, streaming, and tool calling. | oracle/ | 125 | — | ~2.4k | Automated safety check: Pass | UPL-1.0 | 1 mo ago |
| 11 | Loads large language models in 8-bit or 4-bit with bitsandbytes so they fit smaller GPUs, and sets up QLoRA fine-tuning on a 4-bit base model. | Orchestra-Research/ | 13k | 2 repos | ~2.5k | Automated safety check: Pass | MIT | 3 mo ago |
| 12 | Trains LLMs with reinforcement learning using verl, from ByteDance's Seed team, with GRPO, PPO and other algorithms and swappable training and rollout backends. | Orchestra-Research/ | 13k | 2 repos | ~2.4k | Automated safety check: Pass | MIT | 3 mo ago |
| 13 | Loads pre-trained Hugging Face Transformers models for text, vision and audio tasks, runs inference with pipelines and fine-tunes on custom datasets. | davila7/ | 33k | 11 repos | ~1.2k | Automated safety check: Pass | MIT | today |
| 14 | Walks through nanoGPT, Karpathy's compact GPT implementation: training on Shakespeare, reproducing GPT-2, fine-tuning GPT-2 checkpoints and training on your own text. | Orchestra-Research/ | 13k | 2 repos | ~1.7k | Automated safety check: Pass | MIT | 3 mo ago |
| 15 | Reference for post-training language models with TRL: which trainer and dataset format to use for SFT, DPO, GRPO, KTO and reward models, and how to add LoRA. | huggingface/ | 11k | 1 repo | ~1.1k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 16 | 16.Setup Connect this agent to a running Guaardvark (self-hosted AI studio) and check what it can do right now. | guaardvark/ | 258 | — | ~1.2k | Automated safety check: Pass | MIT | today |
| 17 | Fine-tune LLM models using LoRA on OCI AI Quick Actions (AQUA). | oracle/ | 125 | — | ~1.7k | Automated safety check: Pass | UPL-1.0 | 1 mo ago |
| 18 | Fine-tunes and evaluates OpenVLA-OFT and OFT+ robot policies with LoRA and continuous action heads on LIBERO simulation and ALOHA real-robot setups. | Orchestra-Research/ | 13k | — | ~3.7k | Automated safety check: Pass | MIT | 3 mo ago |
| 19 | 19.Modal Modal is a serverless cloud platform for running Python on demand, including on-demand GPUs. | K-Dense-AI/ | 48k | 1 repo | ~4.5k | Automated safety check: Notes | Apache-2.0 | 6 days ago |
| 20 | 20.Transformers Hugging Face Transformers for loading Hub models, running pipeline inference, text generation, and Trainer fine-tuning on NLP, vision, audio, and multimodal tasks. | K-Dense-AI/ | 48k | 1 repo | ~2.8k | Automated safety check: Notes | Apache-2.0 | 6 days ago |
| 21 | Manages and orchestrates prompts in Agent Platform. An agent skill from google/skills. | google/ | 21k | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 22 | Manages GenAI tuning jobs in Agent Platform. An agent skill from google/skills. | google/ | 21k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 23 | Run the disk-backed DEFT AOI improvement loop for NVIDIA Cosmos Reason 3 / Cosmos3 models, using Nano by default and Edge or Super when explicitly requested: evaluate the base model on Proxy and… | NVIDIA/ | 3.6k | — | ~5k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 24 | Launch, relaunch, or sweep STANDARD (non-agentic) SkyRL RL on CINECA Leonardo — GRPO on math/reasoning datasets (gsm8k, MATH/aime) and on-policy distillation (OPD, teacher→student) — via raw sbatch… | open-thoughts/ | 301 | — | ~3.2k | Automated safety check: Pass | Apache-2.0 | 12 days ago |
| 25 | 25.Sft Launch Launch SFT via python -m hpc.launch --jobtype sft on any cluster (JSC Jupiter GH200, CINECA Leonardo A100, TACC Vista GH200), with EITHER backend — LLaMA-Factory (default) or axolotl (--sftbackend… | open-thoughts/ | 301 | — | ~2.9k | Automated safety check: Pass | Apache-2.0 | 12 days ago |
| 26 | 26.Runpod A skill your agent uses when running GPU compute on RunPod and deciding between Pods (hourly, always-on) and Serverless (per-second, autoscaling) for training, fine-tuning or inference — serverless… | ericrisco/ | 180 | — | ~2.8k | Automated safety check: Pass | MIT | yesterday |
| 27 | A skill your agent uses when training or debugging a neural net in PyTorch — the forward/loss/backward/step loop and its silent bugs, mixed precision (AMP), AdamW/LR schedules, DDP/FSDP/ZeRO… | ericrisco/ | 180 | — | ~3.4k | Automated safety check: Pass | MIT | yesterday |
| 28 | 28.Llama Cpp Operate, configure, benchmark, and troubleshoot llama.cpp across CPU, Metal, CUDA, HIP/ROCm, Vulkan, SYCL, and hybrid or multi-GPU systems. | magnus919/ | 115 | — | ~2.3k | Automated safety check: Pass | MIT | yesterday |