Topic · Data & Analytics
Best machine learning skills for Claude Code, Codex and other agents.
- skills
- 337
- official
- 13
Machine learning skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Machine learning in Python with scikit-learn. An agent skill from zLanqing/codex-claude-academic-skills. | zLanqing/ | 4.6k | 17 repos | ~3.9k | Automated safety check: Pass | BSD-3-Clause | 4 mo ago |
| 2 | Runs lm-evaluation-harness to benchmark language models on academic suites such as MMLU, GSM8K and HumanEval, compare models and track training checkpoints. | Orchestra-Research/ | 13k | 8 repos | ~3k | Automated safety check: Pass | MIT | 3 mo ago |
| 3 | Explains machine learning, LLMs, AI agents and speech modeling in a Hung-Yi Lee-inspired teaching style, drawing on a knowledge base built from his lectures and research references. | voidful/ | 1.3k | — | ~13k | Automated safety check: Pass | No licence | 1 mo ago |
| 4 | Evaluate and improve GenAI models and agents using the Google GenAI Evaluation SDK. | GoogleCloudPlatform/ | 791 | — | ~2k | Automated safety check: Pass | Apache-2.0 | today |
| 5 | World-class data science skill for statistical modeling, experimentation, causal inference, and advanced analytics. | Raidriar7170/ | 125 | 6 repos | ~1.4k | Automated safety check: Pass | MIT | 11 days ago |
| 6 | Trains and evaluates several WiFi-signal-based pose and sensing models, from unsupervised pose estimation to domain adaptation and publishing. | ruvnet/ | 97k | — | ~1.3k | Automated safety check: Notes | MIT | today |
| 7 | Takes a Kaggle competition from rules and validation design through baselines, ensembling and notebook architecture to a scored submission. | FrankS-IntelLab/ | 188 | — | ~4k | Automated safety check: Pass | MIT | 3 mo ago |
| 8 | Plans memory headroom, works through out-of-memory failures and watches temperature and power during long ML training jobs on NVIDIA DGX Spark. | wshobson/ | 40k | 1 repo | ~2k | Automated safety check: Pass | MIT | 2 days ago |
| 9 | Analyze user retention and churn using survival analysis, cohort analysis, and machine learning. | liangdabiao/ | 290 | 1 repo | ~1.3k | Automated safety check: Notes | No licence | 5 mo ago |
| 10 | Reference patterns for writing qiskit 2.x code for variational quantum machine learning: feature maps, VQC training, VQE for chemistry, MPS circuits and noise models. | aiming-lab/ | 15k | — | ~4.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 11 | Scaffold a brand-new Comet example in this repo from the canonical template under templates/integration-example/. | comet-ml/ | 175 | — | ~1k | Automated safety check: Pass | No licence | 1 mo ago |
| 12 | Designs ML pipeline infrastructure: experiment tracking with MLflow or Weights & Biases, Kubeflow and Airflow orchestration, Feast feature stores and model validation gates. | Jeffallan/ | 12k | 1 repo | ~1.9k | Automated safety check: Pass | MIT | 4 days ago |
| 13 | Routes agents to the right tangermeme reference for analyzing trained genomic deep learning models, from attributions and motif experiments to variant effects and design. | jmschrei/ | 311 | — | ~1.6k | Automated safety check: Pass | MIT | today |
| 14 | Shows how to organize PyTorch training with Lightning's LightningModule and Trainer, covering validation, DDP, callbacks and learning-rate scheduling. | Orchestra-Research/ | 13k | 7 repos | ~2.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 15 | 15.Geoml Working knowledge of the geoML Python package (github.com/italo-goncalves/geoML): variational Gaussian processes for spatial data, implicit geological modelling, block models, drillhole data… | italo-goncalves/ | 108 | — | ~4.2k | Automated safety check: Pass | GPL-3.0 | 5 days ago |
| 16 | 16.Radiomics ML A skill your agent uses when building or auditing a radiomics or tabular clinical-ML prediction model with a classical learner (LASSO, SVM, random forest, XGBoost and similar). | Aperivue/ | 329 | 1 repo | ~2.7k | Automated safety check: Pass | MIT | 2 days ago |
| 17 | Turns a plain-language model training request into a validated QuantMind training config file that can be imported from the Model Training page. | qusong0627/ | 1.7k | — | ~1.5k | Automated safety check: Pass | AGPL-3.0 | yesterday |
| 18 | 18.Unimol A standardized CLI wrapper for Uni-Mol molecular ML workflows that handles representation extraction (embeddings), model training (regression/classification), and property prediction with built-in… | jinzhezenggroup/ | 148 | 1 repo | ~1.5k | Automated safety check: Pass | LGPL-3.0-or-later | yesterday |
| 19 | 19.Alpha Evolve A skill your agent uses when the user wants to evolve an ML model/program through population-based search rather than a single sequential refine loop — a generational evolution where parallel… | gaasher/ | 174 | 1 repo | ~3.4k | Automated safety check: Pass | MIT | 3 mo ago |
| 20 | Guides an agent through tracking ML experiments with W&B: run logging, config capture, hyperparameter sweeps, artifacts and a model registry. | Orchestra-Research/ | 13k | 10 repos | ~3.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 21 | Build a new time-series analytics use case on top of the deployed Time Series Analytics microservice — bring it up with Docker Compose (from a repo clone, or by fetching the compose files from… | open-edge-platform/ | 168 | — | ~3.1k | Automated safety check: Pass | Apache-2.0 | today |
| 22 | A skill your agent uses for GPU-accelerated machine learning on tabular data using NVIDIA cuML. | wahyudesu/ | 114 | — | ~1.8k | Automated safety check: Pass | MIT | 5 mo ago |
| 23 | Entry point and controller for computational chemistry and materials workflows. | JCLiuGroup/ | 143 | 1 repo | ~1.4k | Automated safety check: Pass | Unknown | 10 days ago |
| 24 | Estimate a covariance / correlation / precision matrix incrementally with precise. | microprediction/ | 336 | — | ~535 | Automated safety check: Pass | MIT | yesterday |
| 25 | Serverless GPU cloud platform for running ML workloads. An agent skill from Orchestra-Research/AI-Research-SKILLs. | Orchestra-Research/ | 13k | 5 repos | ~2.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 26 | Keeps a persistent, append-only journal of ML experiment hypotheses and results across sessions, so no hyperparameter or architecture change runs without being logged first. | Leeroo-AI/ | 195 | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 27 | 27.Xybrid Init Generate model metadata for an ML model so it works with xybrid. | xybrid-ai/ | 465 | — | ~3k | Automated safety check: Pass | Apache-2.0 | today |
| 28 | Invoke ML models, run vector search, and connect to MCP servers from Databricks Apps. | databricks-solutions/ | 183 | — | ~1.7k | Automated safety check: Pass | Unknown | 3 days ago |
| 29 | CEO checks consistency of agent outputs during post-completion review | chekusu/ | 688 | — | ~824 | Automated safety check: Pass | Apache-2.0 | 3 mo ago |
| 30 | Produces ranked, evidence-grounded next steps when an ML experiment has stalled, drawing on a Leeroopedia knowledge base or on fetched docs and issues. | Leeroo-AI/ | 195 | — | ~4.8k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 31 | Uses Ray Data to read, transform and write large datasets across a cluster for ML training and batch inference, with streaming execution and optional GPU steps. | Orchestra-Research/ | 13k | 3 repos | ~1.8k | Automated safety check: Pass | MIT | 3 mo ago |
| 32 | Scales PyTorch, TensorFlow and Hugging Face training from a single GPU to multi-node clusters with Ray Train, including Ray Tune sweeps and checkpoint recovery. | Orchestra-Research/ | 13k | 3 repos | ~2.7k | Automated safety check: Pass | MIT | 3 mo ago |
| 33 | Run LAMMPS molecular dynamics with DeePMD-kit machine learning potentials. | Hello-QM/ | 205 | 1 repo | ~1k | Automated safety check: Pass | AGPL-3.0 | 15 days ago |
| 34 | Guides time series machine learning with the aeon toolkit: classification, regression, clustering, forecasting, anomaly detection, segmentation and similarity search. | davila7/ | 32k | 14 repos | ~2.6k | Automated safety check: Pass | MIT | today |
| 35 | 35.Andrew Ng Applies the reasoning, principles, and frameworks of Andrew Ng (machine learning pioneer, co-founder of Coursera and DeepLearning.AI, Stanford University, and former Google Brain lead). | K-Dense-AI/ | 282 | — | ~1.5k | Automated safety check: Pass | MIT | 1 mo ago |
| 36 | 36.Precise Online (incremental) covariance, correlation, and precision estimation in Python — the streaming complement to sklearn.covariance. | microprediction/ | 336 | — | ~782 | Automated safety check: Pass | MIT | yesterday |
| 37 | Declare the pipeline from data source to predictor as a skrub DataOps graph. | probabl-ai/ | 135 | — | ~4k | Automated safety check: Pass | BSD-3-Clause | today |
| 38 | Guides an agent through designing an MLOps pipeline that covers data preparation, training, validation and deployment, with DAG orchestration and reference guides. | wshobson/ | 40k | 12 repos | ~1.8k | Automated safety check: Pass | MIT | 2 days ago |
| 39 | 39.Core ML Core ML, Create ML, Vision framework, Natural Language framework, on-device ML integration. | gustavscirulis/ | 117 | 2 repos | ~3.8k | Automated safety check: Notes | Unknown | 5 mo ago |
| 40 | Works with genomic intervals using gtars, a Rust toolkit with Python bindings: overlap detection, coverage tracks, tokenization for ML models and reference sequences. | davila7/ | 32k | 12 repos | ~1.9k | Automated safety check: Pass | MIT | today |
| 41 | 41.Pathml Computational pathology toolkit for analyzing whole-slide images (WSI) and multiparametric imaging data. | davila7/ | 32k | 12 repos | ~1.9k | Automated safety check: Pass | MIT | today |
| 42 | Builds machine learning pipelines on clinical data with PyHealth: EHR datasets, prediction tasks, medical code mapping, healthcare models and evaluation. | davila7/ | 32k | 12 repos | ~4.4k | Automated safety check: Pass | MIT | today |
| 43 | Fits and evaluates survival models with scikit-survival: Cox models, Random Survival Forests, boosting, survival SVMs, concordance index, Brier score and competing risks. | davila7/ | 32k | 12 repos | ~3.7k | Automated safety check: Pass | MIT | today |
| 44 | Explains machine learning predictions with SHAP: picking the right explainer, computing Shapley values and drawing waterfall, beeswarm, bar and force plots. | davila7/ | 32k | 12 repos | ~4.6k | Automated safety check: Pass | MIT | today |
| 45 | 45.Deepchem Molecular machine learning toolkit. An agent skill from davila7/claude-code-templates. | davila7/ | 32k | 11 repos | ~4.4k | Automated safety check: Pass | MIT | today |
| 46 | Add new AI models to Kiln's mlmodellist.py and produce a Discord announcement. | Kiln-AI/ | 5.2k | — | ~15k | Automated safety check: Notes | Unknown | today |
| 47 | Applies the reasoning, architectural principles, and AI philosophy of Christopher Manning (natural language processing expert, Stanford University, director of Stanford AI Lab). | K-Dense-AI/ | 282 | — | ~1.8k | Automated safety check: Pass | MIT | 1 mo ago |
| 48 | Perform comprehensive regression analysis and predictive modeling using linear regression, decision trees, and random forests. | liangdabiao/ | 290 | 1 repo | ~1.7k | Automated safety check: Notes | No licence | 5 mo ago |
Questions, answered from the data.
What is the best machine learning skill?
Scikit Learn from zLanqing/codex-claude-academic-skills ranks first of the 337 machine learning skills listed here, with the highest score: its repository has 4.6k GitHub stars, 17 other GitHub owners carry a copy, its SKILL.md loads about 3.9k tokens and it passes the automated safety check with no findings. Next come LLM Benchmarking with lm-evaluation-harness and Hung-Yi Lee Teaching Style.
Which machine learning skills are official?
13 of the 337 machine learning skills are official, published by the vendor's own GitHub organization: Django Models, Transformers.js, Model Evaluation, Oci Data Science, Bigquery Bigframes and 8 more.
How are these skills ranked?
By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.