Search
Hugging Face · For data scientists
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Runs lm-evaluation-harness to benchmark language models on academic suites such as MMLU, GSM8K and HumanEval, compare models and track training checkpoints. | Orchestra-Research/ | 13k | 8 repos | ~3k | Automated safety check: Pass | MIT | 3 mo ago |
| 2 | Guide to using Meta's Segment Anything Model for zero-shot image segmentation with point, box or mask prompts, or automatic mask generation. | Orchestra-Research/ | 13k | 8 repos | ~3.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 3 | Shows how to store documents and embeddings in Chroma, query them by similarity with metadata filters, and persist them to disk for RAG and semantic search projects. | Orchestra-Research/ | 13k | 7 repos | ~2.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 4 | Chooses the right serving container and current image URI for deploying a Hugging Face model to a SageMaker endpoint, preferring Hugging Face images over generic ones. | huggingface/ | 11k | 1 repo | ~4.6k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 5 | Routes a sentence-transformers training task to the right model type and required reference docs and example scripts, covering bi-encoders, rerankers, sparse and multi-vector models. | huggingface/ | 11k | 1 repo | ~2.6k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 6 | Runs evaluations of Hugging Face Hub models on local hardware with inspect-ai or lighteval, and helps choose between vLLM, Transformers and accelerate backends. | huggingface/ | 11k | 2 repos | ~1.6k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 7 | A skill your agent uses when adding or migrating non-thumbnail images for a Hugging Face Blog post. | huggingface/ | 3.5k | — | ~1.1k | Automated safety check: Pass | No licence | 2 days ago |
| 8 | 8.Esmfold2 Biohub ESMFold2 / ESMFold2-Fast all-atom co-folding (Candido et al. | JimLiu/ | 228 | 4 repos | ~2.5k | Automated safety check: Pass | Apache-2.0 | 3 mo ago |
| 9 | Guide for adding a new model to the Archon engine. An agent skill from areal-project/AReaL. | areal-project/ | 5.8k | — | ~4.9k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 10 | Complete agent-ready workflow for Qwen-family MTP or nextn GGUF conversion and release. | R6410418/ | 1.7k | — | ~1.7k | Automated safety check: Pass | MIT | 3 mo ago |
| 11 | Shows how to load, train and use fast Hugging Face tokenizers, with BPE, WordPiece and Unigram models, padding, truncation and alignment tracking. | Orchestra-Research/ | 13k | 6 repos | ~3.4k | Automated safety check: Pass | MIT | 3 mo ago |
| 12 | Indexes research papers on the Hugging Face Hub from arXiv, links them to models and datasets, claims authorship and generates markdown research articles from templates. | huggingface/ | 11k | 4 repos | ~4.2k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 13 | Trains or fine-tunes language and vision models with TRL or Unsloth on Hugging Face Jobs cloud GPUs, then converts the results to GGUF. | huggingface/ | 11k | 1 repo | ~7.2k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 14 | Parameter-efficient fine-tuning for LLMs using LoRA, QLoRA, and 25+ methods. | Orchestra-Research/ | 13k | 6 repos | ~3.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 15 | Finds llama.cpp-compatible GGUF models on the Hugging Face Hub, picks a quantization for your hardware and launches them with llama-cli or llama-server. | huggingface/ | 11k | 3 repos | ~945 | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 16 | Trains and fine-tunes object detection, image classification and SAM or SAM2 segmentation models on Hugging Face Jobs cloud GPUs and saves the results to the Hub. | huggingface/ | 11k | 1 repo | ~7.5k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 17 | Prioritize and run Modal conference-deadline agents on high-ROI conferences only. | huggingface/ | 350 | — | ~1.3k | Automated safety check: Pass | MIT | 17 days ago |
| 18 | Add a new HuggingFace-supported VAE to the stage1 tokenizer pipeline. | End2End-Diffusion/ | 105 | — | ~1.1k | Automated safety check: Pass | No licence | 3 mo ago |
| 19 | A skill your agent uses when adding support for a new model to VeOmni. | ByteDance-Seed/ | 2.2k | — | ~2k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 20 | Integrate new evaluation datasets into FlagEvalMM as benchmark tasks. | flageval-baai/ | 108 | — | ~3.2k | Automated safety check: Pass | No licence | 5 mo ago |
| 21 | Builds reusable command line scripts that fetch, enrich or process data from the Hugging Face API, aimed at chained, repeated or automated tasks. | huggingface/ | 11k | 2 repos | ~1.5k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 22 | Train object-detection, image-classification, or SAM segmentation models on Hugging Face Jobs. | waybarrios/ | 534 | — | ~2.7k | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 23 | Logs and visualizes ML training metrics with Trackio, firing alerts for issues like loss spikes, and syncing a live dashboard to a Hugging Face Space. | huggingface/ | 11k | 2 repos | ~1.3k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 24 | Benchmarks code generation models with the BigCode Evaluation Harness across HumanEval, MBPP, MultiPL-E and other suites using pass@k metrics. | Orchestra-Research/ | 13k | 4 repos | ~2.9k | Automated safety check: Pass | MIT | 3 mo ago |
| 25 | Produces ranked, evidence-grounded next steps when an ML experiment has stalled, drawing on a Leeroopedia knowledge base or on fetched docs and issues. | Leeroo-AI/ | 195 | — | ~4.8k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 26 | Adds distributed and mixed-precision training to a PyTorch script with a few Accelerate lines, then launches it on one GPU, many GPUs or DeepSpeed and FSDP setups. | Orchestra-Research/ | 13k | 5 repos | ~2.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 27 | Explores Hugging Face datasets through the read-only Dataset Viewer API: list splits, preview and page through rows, search, filter, and fetch parquet links and statistics. | huggingface/ | 11k | 3 repos | ~1.1k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 28 | Fetches Hugging Face paper pages as markdown and reads paper metadata through the papers API when you share a paper URL, an arXiv link or an arXiv ID. | huggingface/ | 11k | 3 repos | ~2.3k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 29 | Uses Ray Data to read, transform and write large datasets across a cluster for ML training and batch inference, with streaming execution and optional GPU steps. | Orchestra-Research/ | 13k | 3 repos | ~1.8k | Automated safety check: Pass | MIT | 3 mo ago |
| 30 | Best practices for multi-step Python tasks including data analysis, HuggingFace datasets, token counting, and any task requiring state across multiple python() calls. | A-EVO-Lab/ | 809 | — | ~476 | Automated safety check: Pass | No licence | 1 mo ago |
| 31 | Checks training code, configs and math against documented framework behavior before an expensive run, citing a knowledge base or official docs for every claim. | Leeroo-AI/ | 195 | — | ~3.8k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 32 | Finds top models for a task from official Hugging Face benchmark leaderboards, filters them to what fits your hardware, and returns a comparison table with scores. | huggingface/ | 11k | 2 repos | ~1.5k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 33 | Builds and publishes a Gradio demo on Hugging Face Spaces for a LoRA, with the pipeline, UI and settings chosen to match that LoRA's task and model card. | huggingface/ | 11k | 2 repos | ~8.4k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 34 | Guides causal experiments on PyTorch models with pyvene, such as causal tracing, activation patching and interchange intervention training, to test how a model works. | Orchestra-Research/ | 13k | 2 repos | ~3.5k | Automated safety check: Pass | MIT | 3 mo ago |
| 35 | Loads large language models in 8-bit or 4-bit with bitsandbytes so they fit smaller GPUs, and sets up QLoRA fine-tuning on a 4-bit base model. | Orchestra-Research/ | 13k | 2 repos | ~2.5k | Automated safety check: Pass | MIT | 3 mo ago |
| 36 | Scales PyTorch, TensorFlow and Hugging Face training from a single GPU to multi-node clusters with Ray Train, including Ray Tune sweeps and checkpoint recovery. | Orchestra-Research/ | 13k | 2 repos | ~2.7k | Automated safety check: Pass | MIT | 3 mo ago |
| 37 | Generates text embeddings locally with the sentence-transformers library for RAG, semantic search, clustering and similarity, with model picks for general, multilingual and legal text. | Orchestra-Research/ | 13k | 2 repos | ~1.6k | Automated safety check: Pass | MIT | 3 mo ago |
| 38 | Explains three ways to speed up LLM inference: draft-model speculative decoding, Medusa heads and lookahead decoding with Jacobi iteration, and when each one fits. | Orchestra-Research/ | 13k | 2 repos | ~3.5k | Automated safety check: Pass | MIT | 3 mo ago |
| 39 | Guides reinforcement-learning research with torchforge, Meta's PyTorch-native library that keeps RL algorithms apart from infrastructure, including GRPO math-reasoning runs. | Orchestra-Research/ | 13k | 2 repos | ~2.5k | Automated safety check: Pass | MIT | 3 mo ago |
| 40 | Trains LLMs with reinforcement learning using verl, from ByteDance's Seed team, with GRPO, PPO and other algorithms and swappable training and rollout backends. | Orchestra-Research/ | 13k | 2 repos | ~2.4k | Automated safety check: Pass | MIT | 3 mo ago |
| 41 | Loads pre-trained Hugging Face Transformers models for text, vision and audio tasks, runs inference with pipelines and fine-tunes on custom datasets. | davila7/ | 33k | 11 repos | ~1.2k | Automated safety check: Pass | MIT | today |
| 42 | 42.Search Search all Japanese NLP resources (libraries, models, datasets, tutorials, dictionaries, Hugging Face). | taishi-i/ | 1k | — | ~4.3k | Automated safety check: Notes | CC0-1.0 | 4 days ago |
| 43 | Explains how to train a model to be harmless with self-critique, revision and AI-generated preference feedback, with Hugging Face and TRL code for each stage. | Orchestra-Research/ | 13k | 2 repos | ~2k | Automated safety check: Pass | MIT | 3 mo ago |
| 44 | Guides GRPO reinforcement-learning fine-tuning of language models with TRL, centered on designing reward functions for formats, verifiable tasks and reasoning. | Orchestra-Research/ | 13k | 3 repos | ~4.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 45 | Guide to using Mamba selective state-space models for linear-time sequence modeling, from the Mamba block and pretrained checkpoints to Mamba-2 and speed comparisons. | Orchestra-Research/ | 13k | 2 repos | ~1.8k | Automated safety check: Pass | MIT | 3 mo ago |
| 46 | Orchestrates video data augmentation and auto-labeling workflows on OSMO, from flow selection and preflight checks to submission, monitoring and output download. | NVIDIA/ | 3.6k | — | ~4.7k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 47 | Entry point for hosting a model on Amazon SageMaker: asks a few questions, picks a deployment pathway and hands off to the specialist skills. | huggingface/ | 11k | 1 repo | ~2.1k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 48 | Reference for post-training language models with TRL: which trainer and dataset format to use for SFT, DPO, GRPO, KTO and reward models, and how to add LoRA. | huggingface/ | 11k | 1 repo | ~1.1k | Automated safety check: Pass | Apache-2.0 | 2 days ago |