Library
scikit-learn agent skills for Claude Code, Codex and other agents.
- skills
- 53
- official
- 1
- Type
- Library
- Website
- scikit-learn.org
- Official GitHub
- scikit-learn
scikit-learn skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
Official (1 skill)
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Train ML models on Databricks. An agent skill from databricks/databricks-agent-skills. | databricks/ | 345 | — | ~4.6k | Automated safety check: Pass | Unknown | yesterday |
Community
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 2 | Machine learning in Python with scikit-learn. An agent skill from zLanqing/codex-claude-academic-skills. | zLanqing/ | 4.6k | 17 repos | ~3.9k | Automated safety check: Pass | BSD-3-Clause | 4 mo ago |
| 3 | Forecasts any univariate time series zero-shot with Google's TimesFM model, returning point forecasts and calibrated prediction intervals without training. | google-research/ | 34k | — | ~4.7k | Automated safety check: Pass | Apache-2.0 | 7 days ago |
| 4 | World-class data science skill for statistical modeling, experimentation, causal inference, and advanced analytics. | Raidriar7170/ | 125 | 6 repos | ~1.4k | Automated safety check: Pass | MIT | 11 days ago |
| 5 | Writes statistical analysis code for experimental data, runs it through a four-round review, and reports effect sizes, p-values and confidence intervals. | lingzhi227/ | 383 | — | ~886 | Automated safety check: Pass | No licence | 7 mo ago |
| 6 | Build a new time-series analytics use case on top of the deployed Time Series Analytics microservice — bring it up with Docker Compose (from a repo clone, or by fetching the compose files from… | open-edge-platform/ | 168 | — | ~3.1k | Automated safety check: Pass | Apache-2.0 | today |
| 7 | Six-phase process for reproducing a published paper's results from provided data, from variable mapping and sample filtering through regression tables and a written report. | xjtulyc/ | 617 | 1 repo | ~1.3k | Automated safety check: Pass | No licence | 6 mo ago |
| 8 | Estimate a covariance / correlation / precision matrix incrementally with precise. | microprediction/ | 336 | — | ~535 | Automated safety check: Pass | MIT | yesterday |
| 9 | Checks an LLM judge against human labels using train, dev and test splits, TPR and TNR, and a bias correction applied to production data. | ai-evals-course/ | 1.5k | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | 12 days ago |
| 10 | Guides time series machine learning with the aeon toolkit: classification, regression, clustering, forecasting, anomaly detection, segmentation and similarity search. | davila7/ | 32k | 14 repos | ~2.6k | Automated safety check: Pass | MIT | today |
| 11 | 11.Precise Online (incremental) covariance, correlation, and precision estimation in Python — the streaming complement to sklearn.covariance. | microprediction/ | 336 | — | ~782 | Automated safety check: Pass | MIT | yesterday |
| 12 | Declare the pipeline from data source to predictor as a skrub DataOps graph. | probabl-ai/ | 135 | — | ~4k | Automated safety check: Pass | BSD-3-Clause | today |
| 13 | Fits and evaluates survival models with scikit-survival: Cox models, Random Survival Forests, boosting, survival SVMs, concordance index, Brier score and competing risks. | davila7/ | 32k | 12 repos | ~3.7k | Automated safety check: Pass | MIT | today |
| 14 | Tracks ML experiments, versions models in the MLflow registry and covers deployment and reproducibility, with autologging for common frameworks. | Orchestra-Research/ | 13k | 2 repos | ~3.9k | Automated safety check: Pass | MIT | 3 mo ago |
| 15 | 15.Molfeat Molecular featurization for ML (100+ featurizers). An agent skill from davila7/claude-code-templates. | davila7/ | 32k | 10 repos | ~3.7k | Automated safety check: Pass | MIT | today |
| 16 | 16.Edit how to use the edit command properly | omegaml/ | 107 | — | ~206 | Automated safety check: Pass | Apache-2.0 | today |
| 17 | Fit, summarize, plot, and interpret a chosen CausalPy experiment. | pymc-labs/ | 1.2k | 1 repo | ~1.3k | Automated safety check: Pass | Apache-2.0 | today |
| 18 | Trains scikit-learn models with walk-forward validation on features from OHLCV data to predict return direction and turn the predictions into trading signals. | HKUDS/ | 35k | — | ~3.2k | Automated safety check: Pass | MIT | yesterday |
| 19 | World-class senior data scientist skill specialising in statistical modeling, experiment design, causal inference, and predictive analytics. | alirezarezvani/ | 28k | 2 repos | ~2.3k | Automated safety check: Pass | MIT | 1 mo ago |
| 20 | Turn a plain-English description ("a node that runs on the kernel and does XGBoost predictions", "a node that trims whitespace", "an ML clustering node") into a correct single-file Flowfile custom… | Edwardvaneechoud/ | 370 | — | ~8.6k | Automated safety check: Pass | MIT | today |
| 21 | Builds the code for a frozen research experiment test-first, with leakage controls, seed handling and saved evidence so results can be rerun and audited. | Light0305/ | 640 | — | ~2.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 22 | Quantitative strategy frameworks: pairs trading/cointegration, volatility regime strategies, seasonality/calendar effects, multi-factor models (IC/IR), factor research and screening, correlation… | helsome/ | 269 | 1 repo | ~1.6k | Automated safety check: Pass | MIT | 3 days ago |
| 23 | GPU-accelerates scientific Python on NVIDIA hardware and verifies that the result is correct and faster. | K-Dense-AI/ | 48k | 1 repo | ~3.4k | Automated safety check: Pass | MIT | 2 days ago |
| 24 | Processes paired-end 16S amplicon reads into QIIME 2 ASVs and taxonomy with retained artifact provenance. | K-Dense-AI/ | 48k | 1 repo | ~2.2k | Automated safety check: Pass | MIT | 2 days ago |
| 25 | 25.Scikit Learn Supports machine learning in Python with scikit-learn. An agent skill from K-Dense-AI/scientific-agent-skills. | K-Dense-AI/ | 48k | 1 repo | ~3.3k | Automated safety check: Notes | BSD-3-Clause | 2 days ago |
| 26 | 26.ML Engineer Machine learning engineer expert for PyTorch, scikit-learn, model evaluation, and MLOps | RightNow-AI/ | 18k | — | ~987 | Automated safety check: Pass | Apache-2.0 | 3 mo ago |
| 27 | 27.Umap Learn Applies UMAP-learn to nonlinear dimensionality reduction, 2D/3D embeddings, clustering preprocessing, supervised or semi-supervised UMAP, DensMAP, AlignedUMAP, and Parametric UMAP workflows. | K-Dense-AI/ | 48k | 1 repo | ~5.4k | Automated safety check: Pass | BSD-3-Clause | 2 days ago |
| 28 | Validates predictive models on omics and biomedical data with nested cross-validation, group/batch/temporal-aware splits, the full data-leakage taxonomy, probability calibration, decision-curve net… | GPTomics/ | 1.2k | 1 repo | ~5k | Automated safety check: Pass | MIT | 1 mo ago |
| 29 | Train ML models with scikit-learn, PyTorch, TensorFlow. An agent skill from secondsky/claude-skills. | secondsky/ | 227 | 1 repo | ~1.7k | Automated safety check: Pass | MIT | 9 days ago |
| 30 | Operates the QIIME2 framework as the glue for an amplicon analysis - the .qza/.qzv artifact model, semantic types (FeatureTable[Frequency], SampleData[PairedEndSequencesWithQuality]… | GPTomics/ | 1.2k | 1 repo | ~5.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 31 | Assigns taxonomy to amplicon ASVs/OTUs (16S, ITS, 18S) with a classifier conditioned on a reference database and primer region - DADA2 assignTaxonomy + addSpecies (RDP naive Bayes), DECIPHER IDTAXA… | GPTomics/ | 1.2k | 1 repo | ~6.1k | Automated safety check: Pass | MIT | 1 mo ago |
| 32 | Molecular featurization hub (100+ featurizers) for ML. An agent skill from jaechang-hits/SciAgent-Skills. | jaechang-hits/ | 370 | 1 repo | ~4.3k | Automated safety check: Pass | Apache-2.0 | 8 days ago |
| 33 | A comprehensive toolkit for survival analysis and time-to-event modeling in Python using scikit-survival; use it when you need to model censored time-to-event outcomes, fit Cox/RSF/GB models or… | aipoch/ | 2k | — | ~1.6k | Automated safety check: Pass | MIT | 20 days ago |
| 34 | Pick how to write a figure before custom plot code. An agent skill from probabl-ai/skills. | probabl-ai/ | 135 | — | ~785 | Automated safety check: Pass | BSD-3-Clause | today |
| 35 | 35.Scikit Learn Machine learning: clustering, PCA/t-SNE/UMAP, classification, prediction regression (Ridge/Lasso/ensemble), cross-validation, Pipelines. | brycewang-stanford/ | 4.5k | — | ~4.6k | Automated safety check: Pass | Unknown | 2 days ago |
| 36 | Define custom distance/similarity metrics for clustering and ML algorithms. | benchflow-ai/ | 1.8k | — | ~674 | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 37 | Python statistical modeling: regression (OLS, WLS, GLM), discrete (Logit, Poisson, NegBin), time series (ARIMA, SARIMAX, VAR), with rigorous inference, diagnostics, and hypothesis tests. | majiayu000/ | 666 | 2 repos | ~4.2k | Automated safety check: Pass | BSD-3-Clause | today |
| 38 | Classical ML in Python: classification, regression, clustering, dim reduction, evaluation, tuning, preprocessing pipelines. | jaechang-hits/ | 370 | 1 repo | ~4k | Automated safety check: Pass | BSD-3-Clause | 8 days ago |
| 39 | Time-to-event modeling with scikit-survival: Cox PH (elastic net), Random Survival Forests, Boosting, SVMs for censored data. | jaechang-hits/ | 370 | 1 repo | ~6.9k | Automated safety check: Pass | GPL-3.0 | 8 days ago |
| 40 | Builds classification models for omics data using RandomForest, XGBoost, and logistic regression with sklearn-compatible APIs. | majiayu000/ | 666 | 1 repo | ~923 | Automated safety check: Pass | MIT | today |
| 41 | Load when placing bulk RNA-seq samples on a single-cell reference's pseudotime axis (NNLS deconvolution + nearest-neighbour mapping). | TianGzlab/ | 161 | — | ~1k | Automated safety check: Pass | MIT | 2 mo ago |
| 42 | Best practices for scikit-learn machine learning, model development, evaluation, and deployment in Python | Kilo-Org/ | 189 | 1 repo | ~1.2k | Automated safety check: Pass | Apache-2.0 | 8 days ago |
| 43 | Turn free-text rows into calibrated numeric features with TypeSafe Jev, then model them on a leakage-safe development split. | PKU-YuanGroup/ | 608 | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 44 | GPU-accelerate Python code using CuPy, Numba CUDA, Warp, cuDF, cuML, cuGraph, KvikIO, cuCIM, cuxfilter, cuVS, cuSpatial, and RAFT. | majiayu000/ | 666 | 1 repo | ~8.5k | Automated safety check: Pass | MIT | today |
| 45 | Scikit-learn model training skill with cross-validation, hyperparameter tuning, pipeline construction, and model serialization. | majiayu000/ | 666 | 1 repo | ~2.1k | Automated safety check: Notes | MIT | today |
| 46 | 46.Capy Cortex Autonomous learning system - learns from mistakes, reflects on sessions, and gets smarter over time. | happycapy-ai/ | 137 | — | ~600 | Automated safety check: Pass | MIT | 1 mo ago |
| 47 | A skill your agent uses when training or debugging a neural net in PyTorch — the forward/loss/backward/step loop and its silent bugs, mixed precision (AMP), AdamW/LR schedules, DDP/FSDP/ZeRO… | ericrisco/ | 156 | — | ~3.4k | Automated safety check: Pass | MIT | today |
| 48 | A skill your agent uses when predicting a column from rows of tabular features with classic models — scikit-learn pipelines, RandomForest, XGBoost/LightGBM, leak-free cross-validation, metrics for… | ericrisco/ | 156 | — | ~4.2k | Automated safety check: Pass | MIT | today |
Questions, answered from the data.
What is the best scikit-learn skill?
Databricks ML Training (official) from databricks/databricks-agent-skills ranks first of the 53 scikit-learn skills listed here, with the highest score: its repository has 345 GitHub stars, its SKILL.md loads about 4.6k tokens and it passes the automated safety check with no findings. Next come Scikit Learn and TimesFM Forecasting.
Is there an official scikit-learn skill?
1 of the 53 scikit-learn skills are official, published by the vendor's own GitHub organization: Databricks ML Training.
How are these skills ranked?
By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.