AI model or service
Amazon SageMaker agent skills for Claude Code, Codex and other agents.
- skills
- 29
- official
- 22
- Type
- AI model or service
- Website
- aws.amazon.com
- Official GitHub
- aws, awslabs
- Reviews
- See Amazon SageMaker on Enlisted
Amazon SageMaker skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
Official (22 skills)
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Finds or validates a usable SageMaker execution role before deploying or training, so scripts do not try to create IAM roles they lack permission to create. | huggingface/ | 11k | 1 repo | ~1.8k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 2 | Chooses the right serving container and current image URI for deploying a Hugging Face model to a SageMaker endpoint, preferring Hugging Face images over generic ones. | huggingface/ | 11k | 1 repo | ~4.6k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 3 | Validates dataset formatting and quality for SageMaker model fine-tuning (SFT, DPO, or RLVR). | awslabs/ | 912 | 2 repos | ~1.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 4 | Sets up an isolated Python environment with a supported interpreter and current boto3 before any SageMaker deployment, training or AWS automation code runs. | huggingface/ | 11k | 2 repos | ~1.7k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 5 | Generates code that transforms datasets between ML schemas for model training or evaluation. | awslabs/ | 912 | 2 repos | ~3.5k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 6 | Selects a fine-tuning technique (SFT, DPO, RLVR, or RLAIF) for the user's use case and validates it against the selected model's available recipes. | awslabs/ | 912 | 1 repo | ~604 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 7 | Deploys SageMaker endpoints with autoscaling, CloudWatch alarms and tags on by default, using scripts for real-time, scale-to-zero and async setups. | huggingface/ | 11k | 1 repo | ~6.9k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 8 | Generate comprehensive issue reports from HyperPod clusters (EKS and Slurm) by collecting diagnostic logs and configurations for troubleshooting and AWS Support cases. | awslabs/ | 912 | — | ~890 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 9 | Diagnose performance issues on Amazon SageMaker HyperPod clusters — uneven NCCL bandwidth across nodes and poor filesystem throughput. | awslabs/ | 912 | — | ~4.1k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 10 | Diagnostic-only skill for Slurm scheduler and node-daemon issues on Amazon SageMaker HyperPod Slurm clusters. | awslabs/ | 912 | — | ~3.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 11 | Entry point for hosting a model on Amazon SageMaker: asks a few questions, picks a deployment pathway and hands off to the specialist skills. | huggingface/ | 11k | 1 repo | ~2.1k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 12 | Remote command execution and file transfer on SageMaker HyperPod cluster nodes via AWS Systems Manager (SSM). | awslabs/ | 912 | — | ~1.3k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 13 | Check and compare software component versions on SageMaker HyperPod cluster nodes - NVIDIA drivers, CUDA toolkit, cuDNN, NCCL, EFA, AWS OFI NCCL, GDRCopy, MPI, Neuron SDK (Trainium/Inferentia)… | awslabs/ | 912 | 1 repo | ~910 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 14 | Selects a base model for the user's use case by querying SageMaker Hub. | awslabs/ | 912 | — | ~844 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 15 | Discover the user's local AWS context (active profile, region, account ID, caller identity) at the start of any AWS task. | huggingface/ | 11k | 2 repos | ~989 | Automated safety check: Warn | Apache-2.0 | 6 days ago |
| 16 | Generates code that fine-tunes a base model using SageMaker serverless training jobs. | awslabs/ | 912 | 1 repo | ~2.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 17 | Generates python code that evaluates SageMaker models. An agent skill from awslabs/agent-plugins. | awslabs/ | 912 | 1 repo | ~1.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 18 | Generates code that deploys fine-tuned models from SageMaker Serverless Model Customization to SageMaker endpoints or Bedrock. | awslabs/ | 912 | 1 repo | ~1.5k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 19 | Validates the user's environment for SageMaker AI operations — checks SDK version, AWS region, and execution role. | awslabs/ | 912 | 1 repo | ~248 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 20 | Diagnose NCCL failures and adjacent training-pod failures on HyperPod GPU clusters (EKS or Slurm) — training hangs, AllReduce / collective-op timeouts, EFA or libfabric errors, rendezvous failures… | awslabs/ | 912 | — | ~3.4k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 21 | Runs SQL analytics on SageMaker Catalog asset metadata tables exported as Apache Iceberg in S3 Tables. | aws/ | 2.8k | — | ~2.6k | Automated safety check: Pass | Apache-2.0 | today |
| 22 | Selects, deploys, and customizes AI models on Amazon SageMaker. | aws/ | 2.8k | — | ~1.7k | Automated safety check: Pass | Apache-2.0 | today |
Community
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 23 | Verify or select a SageMaker execution role before creating models, endpoints, or training jobs. | waybarrios/ | 533 | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 24 | Uses Meta's LlamaGuard moderation model to screen prompts and model replies against six safety categories, with vLLM, FastAPI and NeMo Guardrails setups. | Orchestra-Research/ | 13k | 3 repos | ~2.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 25 | Deep expertise in ML/CV model selection, training pipelines, and inference architecture. | alirezarezvani/ | 117 | — | ~3.1k | Automated safety check: Pass | MIT | 9 mo ago |
| 26 | Implement a production SageMaker endpoint with autoscaling, CloudWatch alarms, and tags. | waybarrios/ | 533 | — | ~4.6k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 27 | Select and verify the current region-specific serving container URI for a SageMaker model deployment. | waybarrios/ | 533 | — | ~4.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 28 | Plan and coordinate a model deployment to Amazon SageMaker, including serving stack and real-time versus async inference. | waybarrios/ | 533 | — | ~1.7k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 29 | A full ML pipeline where an agent team collaborates to perform data preparation, model design, training, evaluation, and deployment readiness. | revfactory/ | 1.3k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
Questions, answered from the data.
What is the best Amazon SageMaker skill?
SageMaker IAM Role Preflight (official) from huggingface/skills ranks first of the 29 Amazon SageMaker skills listed here, with the highest score: its repository has 11k GitHub stars, 1 other GitHub owner carry a copy, its SKILL.md loads about 1.8k tokens and it passes the automated safety check with no findings. Next come SageMaker Serving Image Selection and Dataset Evaluation.
Is there an official Amazon SageMaker skill?
22 of the 29 Amazon SageMaker skills are official, published by the vendor's own GitHub organization: SageMaker IAM Role Preflight, SageMaker Serving Image Selection, Dataset Evaluation, Python Environment Setup for SageMaker, Dataset Transformation and 17 more.
How are these skills ranked?
By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.