AI model or service
Weights & Biases agent skills for Claude Code, Codex and other agents.
- skills
- 25
- official
- 3
- Type
- AI model or service
- Website
- wandb.ai
- Official GitHub
- wandb
- Reviews
- See Weights & Biases on Enlisted
Weights & Biases skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
Official (3 skills)
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Azure Weights & Biases SDK for .NET. An agent skill from microsoft/skills. | microsoft/ | 3.1k | 6 repos | ~2.8k | Automated safety check: Pass | MIT | yesterday |
| 2 | Run container-backed AutoML / hyperparameter optimization (HPO) for NVIDIA TAO networks using AutoMLRunner. | NVIDIA/ | 3.5k | — | ~5k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 3 | Cosmos-Embed1 video-text embedding for text-to-video retrieval, video-to-video search, semantic deduplication, and fine-tuning. | NVIDIA/ | 3.5k | — | ~3.5k | Automated safety check: Notes | Apache-2.0 | yesterday |
Community
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 4 | An opintionated skill to prepare a marimo notebook to make it ready for a scheduled run. | koaning/ | 145 | 1 repo | ~819 | Automated safety check: Notes | No licence | 6 mo ago |
| 5 | Designs ML pipeline infrastructure: experiment tracking with MLflow or Weights & Biases, Kubeflow and Airflow orchestration, Feast feature stores and model validation gates. | Jeffallan/ | 12k | 1 repo | ~1.9k | Automated safety check: Pass | MIT | 4 days ago |
| 6 | Guides an agent through tracking ML experiments with W&B: run logging, config capture, hyperparameter sweeps, artifacts and a model registry. | Orchestra-Research/ | 13k | 10 repos | ~3.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 7 | 7.Compare Same-epoch comparison of training runs across wandb, neptune, tensorboard, or mlflow. | fcakyon/ | 414 | — | ~1.2k | Automated safety check: Pass | MIT | 21 days ago |
| 8 | Produces ranked, evidence-grounded next steps when an ML experiment has stalled, drawing on a Leeroopedia knowledge base or on fetched docs and issues. | Leeroo-AI/ | 195 | — | ~4.8k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 9 | Manages biological datasets with LaminDB: versioned artifacts, run lineage, ontology-based annotation, schema validation and links to workflow managers and ML tools. | davila7/ | 32k | 12 repos | ~3.6k | Automated safety check: Pass | MIT | yesterday |
| 10 | Walks through nanoGPT, Karpathy's compact GPT implementation: training on Shakespeare, reproducing GPT-2, fine-tuning GPT-2 checkpoints and training on your own text. | Orchestra-Research/ | 13k | 3 repos | ~1.7k | Automated safety check: Pass | MIT | 3 mo ago |
| 11 | 11.Dashboard Bring up the Lego-RL training dashboard (webui/) on whatever machine you are on, adapting to that box's layout instead of assuming this repo's paths. | LegoX/ | 108 | — | ~4.2k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 12 | Read, analyze, and manage Weights & Biases (wandb) experiment data for PithTrain runs. | mlc-ai/ | 355 | — | ~1.1k | Automated safety check: Warn | Apache-2.0 | 3 days ago |
| 13 | WandB-specific PerforatedAI integration guardrail skill. An agent skill from PerforatedAI/PerforatedAI. | PerforatedAI/ | 237 | — | ~2.8k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 14 | Monitor running experiments, check progress, collect results. | AI4Scientist/ | 128 | 5 repos | ~1.1k | Automated safety check: Pass | No licence | 4 mo ago |
| 15 | Durably ARCHIVE everything informative from a finished run / experiment before it's cleaned up or its cluster artifacts age out — ALL Harbor tracejobs (raw per-trial traces), ALL ray logs, ALL… | open-thoughts/ | 301 | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | 9 days ago |
| 16 | Build, test, and debug Hermes Agent RL environments for Atropos training. | Tommy-yw/ | 546 | — | ~3.3k | Automated safety check: Pass | MIT | 4 mo ago |
| 17 | Verify or complete a new internal Marin developer's local setup and access to GitHub, GCP, Iris, Weights & Biases, Hugging Face, and optional CoreWeave storage. | marin-community/ | 3.9k | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | today |
| 18 | Strategic guidance for operationalizing machine learning models from experimentation to production. | ancoleman/ | 526 | 1 repo | ~9.2k | Automated safety check: Pass | MIT | 10 mo ago |
| 19 | Harvest experiment issue reports and curate docs/reports/index.md only when explicitly requested. | marin-community/ | 3.9k | — | ~645 | Automated safety check: Pass | Apache-2.0 | today |
| 20 | Periodically check WandB metrics during training to catch problems early (NaN, loss divergence, idle GPUs). | AI4Scientist/ | 128 | 4 repos | ~1.3k | Automated safety check: Notes | No licence | 4 mo ago |
| 21 | 21.Sft Launch Launch SFT via python -m hpc.launch --jobtype sft on any cluster (JSC Jupiter GH200, CINECA Leonardo A100, TACC Vista GH200), with EITHER backend — LLaMA-Factory (default) or axolotl (--sftbackend… | open-thoughts/ | 301 | — | ~2.9k | Automated safety check: Pass | Apache-2.0 | 9 days ago |
| 22 | Interactively monitor training metrics from the current Codex session, periodically checking WandB or fallback logs for NaN, divergence, plateaus, and broken runs. | AI4Scientist/ | 128 | 3 repos | ~1.1k | Automated safety check: Notes | No licence | 4 mo ago |
| 23 | Guide for experiment tracking tool setup (MLflow, Weights & Biases, etc.), reproducibility assurance, model registry, and experiment comparison methodology. | revfactory/ | 1.3k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 24 | MLflow, Weights & Biases 등 실험 추적 도구 설정, 재현성 보장, 모델 레지스트리, 실험 비교 방법론 가이드. | revfactory/ | 1.3k | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 25 | 25.Wandb Expert W&B expert: experiment tracking, hyperparameter search, artifact management, sweep, team dashboards, performance visualization. | theneoai/ | 183 | — | ~3.4k | Automated safety check: Pass | MIT | 4 mo ago |
Questions, answered from the data.
What is the best Weights & Biases skill?
Azure Mgmt Weightsandbiases Dotnet (official) from microsoft/skills ranks first of the 25 Weights & Biases skills listed here, with the highest score: its repository has 3.1k GitHub stars, 6 other GitHub owners carry a copy, its SKILL.md loads about 2.8k tokens and it passes the automated safety check with no findings. Next come Tao Run Automl and Tao Finetune Cosmos Embed.
Is there an official Weights & Biases skill?
3 of the 25 Weights & Biases skills are official, published by the vendor's own GitHub organization: Azure Mgmt Weightsandbiases Dotnet, Tao Run Automl and Tao Finetune Cosmos Embed.
How are these skills ranked?
By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.