AI model or service

Weights & Biases agent skills for Claude Code, Codex and other agents.

MLOps platform for experiment tracking, model registry and LLM evaluation with Weave.
skills
25
official
3
Type
AI model or service
Website
wandb.ai
Official GitHub
wandb
Reviews
See Weights & Biases on Enlisted

Weights & Biases skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Official (3 skills)

Official Weights & Biases skills
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Azure Weights & Biases SDK for .NET. An agent skill from microsoft/skills.

microsoft/skills3.1k6 repos~2.8kAutomated safety check: PassMITyesterday
2
2.Tao Run AutomlOfficial

Run container-backed AutoML / hyperparameter optimization (HPO) for NVIDIA TAO networks using AutoMLRunner.

NVIDIA/skills3.5k—~5kAutomated safety check: NotesApache-2.0yesterday
3

Cosmos-Embed1 video-text embedding for text-to-video retrieval, video-to-video search, semantic deduplication, and fine-tuning.

NVIDIA/skills3.5k—~3.5kAutomated safety check: NotesApache-2.0yesterday

Community

Community Weights & Biases skills
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
4

An opintionated skill to prepare a marimo notebook to make it ready for a scheduled run.

koaning/gitcharts1451 repo~819Automated safety check: NotesNo licence6 mo ago
5

Designs ML pipeline infrastructure: experiment tracking with MLflow or Weights & Biases, Kubeflow and Airflow orchestration, Feast feature stores and model validation gates.

Jeffallan/claude-skills12k1 repo~1.9kAutomated safety check: PassMIT4 days ago
6

Guides an agent through tracking ML experiments with W&B: run logging, config capture, hyperparameter sweeps, artifacts and a model registry.

Orchestra-Research/AI-Research-SKILLs13k10 repos~3.1kAutomated safety check: PassMIT3 mo ago
7

Same-epoch comparison of training runs across wandb, neptune, tensorboard, or mlflow.

fcakyon/phd-skills414—~1.2kAutomated safety check: PassMIT21 days ago
8

Produces ranked, evidence-grounded next steps when an ML experiment has stalled, drawing on a Leeroopedia knowledge base or on fetched docs and issues.

Leeroo-AI/superml195—~4.8kAutomated safety check: PassApache-2.06 mo ago
9

Manages biological datasets with LaminDB: versioned artifacts, run lineage, ontology-based annotation, schema validation and links to workflow managers and ML tools.

davila7/claude-code-templates32k12 repos~3.6kAutomated safety check: PassMITyesterday
10

Walks through nanoGPT, Karpathy's compact GPT implementation: training on Shakespeare, reproducing GPT-2, fine-tuning GPT-2 checkpoints and training on your own text.

Orchestra-Research/AI-Research-SKILLs13k3 repos~1.7kAutomated safety check: PassMIT3 mo ago
11

Bring up the Lego-RL training dashboard (webui/) on whatever machine you are on, adapting to that box's layout instead of assuming this repo's paths.

LegoX/Lego-RL108—~4.2kAutomated safety check: PassApache-2.03 days ago
12

Read, analyze, and manage Weights & Biases (wandb) experiment data for PithTrain runs.

mlc-ai/pith-train355—~1.1kAutomated safety check: WarnApache-2.03 days ago
13

WandB-specific PerforatedAI integration guardrail skill. An agent skill from PerforatedAI/PerforatedAI.

PerforatedAI/PerforatedAI237—~2.8kAutomated safety check: PassApache-2.0yesterday
14

Monitor running experiments, check progress, collect results.

AI4Scientist/nano-scientist1285 repos~1.1kAutomated safety check: PassNo licence4 mo ago
15

Durably ARCHIVE everything informative from a finished run / experiment before it's cleaned up or its cluster artifacts age out — ALL Harbor tracejobs (raw per-trial traces), ALL ray logs, ALL…

open-thoughts/OpenThoughts-Agent301—~1.2kAutomated safety check: PassApache-2.09 days ago
16

Build, test, and debug Hermes Agent RL environments for Atropos training.

Tommy-yw/RunbookHermes546—~3.3kAutomated safety check: PassMIT4 mo ago
17

Verify or complete a new internal Marin developer's local setup and access to GitHub, GCP, Iris, Weights & Biases, Hugging Face, and optional CoreWeave storage.

marin-community/marin3.9k—~1.1kAutomated safety check: PassApache-2.0today
18

Strategic guidance for operationalizing machine learning models from experimentation to production.

ancoleman/ai-design-components5261 repo~9.2kAutomated safety check: PassMIT10 mo ago
19

Harvest experiment issue reports and curate docs/reports/index.md only when explicitly requested.

marin-community/marin3.9k—~645Automated safety check: PassApache-2.0today
20

Periodically check WandB metrics during training to catch problems early (NaN, loss divergence, idle GPUs).

AI4Scientist/nano-scientist1284 repos~1.3kAutomated safety check: NotesNo licence4 mo ago
21

Launch SFT via python -m hpc.launch --jobtype sft on any cluster (JSC Jupiter GH200, CINECA Leonardo A100, TACC Vista GH200), with EITHER backend — LLaMA-Factory (default) or axolotl (--sftbackend…

open-thoughts/OpenThoughts-Agent301—~2.9kAutomated safety check: PassApache-2.09 days ago
22

Interactively monitor training metrics from the current Codex session, periodically checking WandB or fallback logs for NaN, divergence, plateaus, and broken runs.

AI4Scientist/nano-scientist1283 repos~1.1kAutomated safety check: NotesNo licence4 mo ago
23

Guide for experiment tracking tool setup (MLflow, Weights & Biases, etc.), reproducibility assurance, model registry, and experiment comparison methodology.

revfactory/harness-1001.3k—~1.4kAutomated safety check: PassApache-2.06 mo ago
24

MLflow, Weights & Biases 등 실험 추적 도구 설정, 재현성 보장, 모델 레지스트리, 실험 비교 방법론 가이드.

revfactory/harness-1001.3k—~1.2kAutomated safety check: PassApache-2.06 mo ago
25

W&B expert: experiment tracking, hyperparameter search, artifact management, sweep, team dashboards, performance visualization.

theneoai/awesome-skills183—~3.4kAutomated safety check: PassMIT4 mo ago

Questions, answered from the data.

What is the best Weights & Biases skill?

Azure Mgmt Weightsandbiases Dotnet (official) from microsoft/skills ranks first of the 25 Weights & Biases skills listed here, with the highest score: its repository has 3.1k GitHub stars, 6 other GitHub owners carry a copy, its SKILL.md loads about 2.8k tokens and it passes the automated safety check with no findings. Next come Tao Run Automl and Tao Finetune Cosmos Embed.

Is there an official Weights & Biases skill?

3 of the 25 Weights & Biases skills are official, published by the vendor's own GitHub organization: Azure Mgmt Weightsandbiases Dotnet, Tao Run Automl and Tao Finetune Cosmos Embed.

How are these skills ranked?

By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.