Agent skill

Hf Cloud Serving Image Selection

by waybarrios in waybarrios/opencode-power-pack

Select and verify the current region-specific serving container URI for a SageMaker model deployment.

Apache-2.0Auto-check passedAI & LLM Engineering

Install Hf Cloud Serving Image Selection

skills CLI
$ npx skills add waybarrios/opencode-power-pack --skill hf-cloud-serving-image-selection -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install waybarrios/opencode-power-pack hf-cloud-serving-image-selection --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/waybarrios/opencode-power-pack.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/hf-cloud-serving-image-selection .claude/skills/hf-cloud-serving-image-selection && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
hf-cloud-serving-image-selection
GitHub stars
533
Token cost
~4.3k tokens
SKILL.md length
2,059 words
Files
3 (incl. scripts, references)
Skills in repo
32
Repo updated
First seen
Licence
Apache-2.0

At a glance

Select and verify the current region-specific serving container URI for a SageMaker model deployment.

  • Works in 3 steps: Verified incompatibility — the model… → No HuggingFace tag exists in the target… → The HuggingFace image is in…
  • Tasks that involve Model hubs and datasets
  • SKILL.md covers Rule zero: HuggingFace images…, Where image URIs come from, Quick decision and Rerankers: TEI or vLLM?, plus 9 more sections
  • Runs Python scripts from its folder; calls curl, python3 and aws; reaches huggingface.co and api.github.com; needs HUGGING_FACE_HUB_TOKEN and HF_TOKEN

What it does

Hf Cloud Serving Image Selection is an agent skill from waybarrios/opencode-power-pack. Select and verify the current region-specific serving container URI for a SageMaker model deployment. Use after the deployment pathway is chosen and before endpoint code; never infer an image URI from memory.

Its SKILL.md is about 4.3k tokens, which your agent loads only when the skill is triggered. The skill folder holds 4 other files, including scripts and reference files (for example `references/model-to-image.md` and `scripts/mirror_image.py`).

It sits in AI & LLM Engineering, covering Model hubs and datasets, LLM inference and serving and Deployment. It works with Amazon SageMaker, Hugging Face, vLLM and Amazon Web Services. The repository describes itself as: 54 rigorous skills for Codex, OpenCode, and Pi: code review, security audit, feature development, frontend design, MCP tools, Hugging Face ML/training, and more. The licence is Apache-2.0.

When your agent uses it

  • Tasks that involve Model hubs and datasets
  • Tasks that involve LLM inference and serving
  • Tasks that involve Deployment

Example prompts

  • “/hf-cloud-serving-image-selection”

Requirements

  • Python 3
  • A credential in HUGGING_FACE_HUB_TOKEN

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. Verified incompatibility — the model needs an architecture/modality/feature no available HuggingFace tag supports, confirmed against the…
  2. No HuggingFace tag exists in the target region and mirroring is not an option.
  3. The HuggingFace image is in "Known-broken images" below.

What it can do on your machine

Read from SKILL.md and the folder at commit 9dccb6d. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • curl
    • python3
    • aws

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • huggingface.co
    • api.github.com
    • raw.githubusercontent.com

    Also links to:

    • aws.github.io
    • github.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • HUGGING_FACE_HUB_TOKEN
    • HF_TOKEN

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Hf Cloud Serving Image Selection loads about 4.3k tokens when it runs, and up to ~7k if it reads all its reference files. Until then it costs about 60 tokens; SKILL.md has 2,059 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~60
When it runs · the whole SKILL.md, loaded when a task matches
~4.3k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~7k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from waybarrios/opencode-power-pack at commit 9dccb6d, republished under its Apache-2.0 licence (© waybarrios). 2,059 words, ~4,286 tokens.

Download SKILL.mdSave it as .claude/skills/hf-cloud-serving-image-selection/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.
name
hf-cloud-serving-image-selection
description
Select and verify the current region-specific serving container URI for a SageMaker model deployment. Use after the deployment pathway is chosen and before endpoint code; never infer an image URI from memory.
license
Apache-2.0 (modified; see UPSTREAMS.json)

Serving Image Selection

The serving container is the single thing most likely to break a SageMaker deployment that "looked correct on paper". Wrong container, stale tag, or the wrong AMI — all produce the same opaque Failed to pass health check error.

Rule zero: HuggingFace images always win

When both a HuggingFace-curated family (huggingface-vllm, huggingface-vllm-omni, huggingface-sglang, tei, huggingface-pytorch-inference) and a generic family (vllm, vllm-omni, sglang, djl-inference) can serve the model, the HuggingFace one is mandatory, not preferred. The only valid reasons to use a generic image:

  1. Verified incompatibility — the model needs an architecture/modality/feature no available HuggingFace tag supports, confirmed against the catalog (not assumed).
  2. No HuggingFace tag exists in the target region and mirroring is not an option.
  3. The HuggingFace image is in "Known-broken images" below.

A newer version number on the generic repo is not a reason. The AWS vllm repo often publishes a higher vLLM version than huggingface-vllm; an older-but-compatible huggingface-vllm tag still wins. "Latest vLLM" is not a requirement anyone stated — compatibility with the model is. If you fall back, record in the deployment log which of the three reasons applied.

Where image URIs come from

Primary source: AWS's official Deep Learning Containers catalog.

URL: https://aws.github.io/deep-learning-containers/reference/available_images/

This page is AWS-maintained and lists every image family with example URIs, tags, CUDA versions, Python versions, and platform (SageMaker vs EC2/ECS/EKS). When picking a URI for a deployment, read it from this page directly — copy the example URL, substitute <region> with the user's region, and pass it to deploy.py --image-uri.

The example URLs use 763104351884 as the account ID for most regions. A few regions use different accounts (e.g. eu-south-1 uses 692866216735). Check the Region Availability page when in doubt.

Exception: none currently. Every image family used by this workflow is now on the AWS catalog page (TEI was added in late 2026). If you encounter a new family that isn't there, mirror it via mirror_image.py and pass the resulting URI directly.

Quick decision

ModelContainer familyHow to get the URI
HuggingFace text-generation LLM (Llama, Qwen, Mistral, etc.)HuggingFace vLLMAWS catalog → "HuggingFace vLLM Inference" (ECR repo huggingface-vllm)
Same as above, multimodalHuggingFace vLLM-OmniAWS catalog → "HuggingFace vLLM-Omni Inference" (ECR repo huggingface-vllm-omni)
HuggingFace embeddingsTEIAWS catalog → "HuggingFace Text Embeddings Inference"
Encoder / cross-encoder rerankers (BERT-family *ForSequenceClassification)TEISame as embeddings
Generative rerankers (causal-LM, e.g. Qwen3-Reranker)HuggingFace vLLMSame as text-generation LLMs — not TEI, see "Rerankers: TEI or vLLM?"
Text-to-image / diffusion (Stable Diffusion, FLUX)DJL InferenceAWS catalog → "DJL Inference" — not HF Inference Toolkit, see "Known-broken images"
HuggingFace classifiers, NER, QA, summarizationHF Inference Toolkit (CPU)AWS catalog → "HuggingFace PyTorch Inference"; GPU tags currently broken — see "Known-broken images"
User specifically wants SGLangHuggingFace SGLangAWS catalog → "HuggingFace SGLang Inference"
No compatible huggingface-vllm tag (verified incompatibility or region gap — see "Rule zero")vLLM (AWS)AWS catalog → "vLLM" section — fallback only, never for version freshness
User specifically wants DJL-LMIDJL InferenceAWS catalog → "DJL Inference"
Amazon NovaSageMaker JumpStartUse JumpStart, not raw endpoint creation
Custom inference codeBYOCUser provides URI

HuggingFace-curated DLCs are mandatory when one is compatible (see "Rule zero"). huggingface-vllm is layered directly on the AWS vLLM DLC — identical SM_VLLM_* env contract and the same cu130 AMI rule — and adds current transformers, current huggingface_hub + hf_xet (avoids the XET-CDN 403 download failures older images hit), and HF performance defaults. It is also what SageMaker SDK v3 auto-routes to. The AWS vllm image is a compatibility escape hatch only; it usually shows a higher vLLM version than huggingface-vllm, and that is not a reason to pick it.

Do not use TGI. Text Generation Inference is archived. Models released after the archive (Qwen3 most famously) fail ping health checks on TGI. Use vLLM instead. (The SageMaker SDK v3 agrees: since PR #5960, June 2026, its ModelBuilder auto-routes text-generation to the HuggingFace vLLM DLC and multimodal tasks to HuggingFace vLLM-Omni.)

Full reasoning for each family in references/model-to-image.md.

Rerankers: TEI or vLLM?

"Reranker" covers two very different architectures, and picking wrong wastes a full endpoint-creation cycle (~20 min) before TEI rejects the model:

  • Encoder cross-encoders (BAAI/bge-reranker-*, mixedbread, most sentence-transformers rerankers) — BERT-family models with a classification head. config.json has architectures: [..ForSequenceClassification] on a TEI-supported encoder type. → TEI.
  • Generative rerankers (Qwen/Qwen3-Reranker-*, and similar causal-LM judges) — decoder LLMs that score relevance via the logprob of a yes/no token. config.json has architectures: [..ForCausalLM]. → HuggingFace vLLM, deployed exactly like a text-generation LLM. TEI will load the architecture then reject the classifier model type (Qwen3 support in TEI is embeddings-only). Invocation pattern (raw completions API, max_tokens=1, logprobs scoring) is in hf-cloud-sagemaker-production-defaults.

Preflight before creating any resources — one HTTP GET settles it:

bash
curl -s https://huggingface.co/<model-id>/raw/main/config.json
# "architectures": ["Qwen3ForCausalLM"]              → vLLM
# "architectures": ["XLMRobertaForSequenceClassification"] → TEI

For TEI also confirm the (architecture, task) pair: an architecture appearing in TEI's supported list means embeddings support, not necessarily classification/reranking support.

Heads-up: SageMaker SDK v3 (PR #5960) routes the text-ranking task to TEI unconditionally — correct for cross-encoders, wrong for generative rerankers. Don't treat the SDK's routing as evidence that TEI can serve a given reranker.

Workflow

For every family: read the URI from the AWS catalog page.

  1. Open https://aws.github.io/deep-learning-containers/reference/available_images/
  2. Find the section for the right family (e.g. "HuggingFace vLLM Inference" for HuggingFace LLMs, "HuggingFace Text Embeddings Inference" for embeddings)
  3. Pick the newest row marked SageMaker for the platform column — newest within that family. Do not switch to another family's section because it lists a higher engine version (see "Rule zero")
  4. Substitute <region> with the user's region (from hf-cloud-aws-context-discovery)
  5. For vLLM: also check the AMI requirement (see "vLLM AMI requirement" below)
  6. Pass the URI to deploy.py --image-uri (real-time) or deploy_async.py --image-uri (async)
TEI: pick the right variant

The TEI catalog row lists two URIs — GPU (tei repo) and CPU (tei-cpu repo). Pick based on the instance type:

  • ml.g*, ml.p*, ml.inf* → GPU variant
  • ml.c*, ml.m*, ml.t* → CPU variant

Mixing them fails: CPU image on a GPU instance wastes hardware, GPU image on a CPU instance fails to start.

Note on the TEI account ID: the catalog page shows 683313688378 as the example account, but TEI is published from a different account namespace than the main AWS DLCs and the per-region account IDs vary. If 683313688378.dkr.ecr.<region>.amazonaws.com/tei:... returns an ECR pull error for a region other than us-east-1, check the Region Availability page for the correct account ID for that region.

vLLM AMI requirement

vLLM DLC images with CUDA 13 or higher (current default: cu130) require setting InferenceAmiVersion=al2-ami-sagemaker-inference-gpu-3-1 on the ProductionVariant. This applies equally to huggingface-vllm and huggingface-vllm-omni (layered on the same cu130 base) and to the AWS vllm repo. Without it the container dies on startup with no CloudWatch logs ever created. The failure looks identical to many other things (account-level issues, quota, networking) and routinely sends people down wrong diagnostic paths.

Lookup table:

Tag containsInferenceAmiVersion to pass
cu130 (or higher)al2-ami-sagemaker-inference-gpu-3-1
cu129 or lower(omit the flag; default AMI works)

Rule of thumb: if the vLLM tag you picked contains cu130 or later, pass --inference-ami-version al2-ami-sagemaker-inference-gpu-3-1 to deploy.py. If a future CUDA version (cu140+) needs a different AMI, add a row to the table when AWS publishes the new image.

This is a vLLM-specific concern. TEI and HF Inference Toolkit images don't need an AMI override.

Configuring the vLLM DLCs (HuggingFace vLLM and AWS vLLM)

Both images share the same contract: configuration as environment variables on the SageMaker model definition, SM_VLLM_* mapped to vLLM CLI flags. The huggingface-vllm entrypoint additionally auto-detects the model when SM_VLLM_MODEL is unset — from /opt/ml/model if artifacts are mounted, else from HF_MODEL_ID — but setting SM_VLLM_MODEL explicitly works on both and is what our examples use.

Show full SKILL.md (831 more words)Show less
Required for every HuggingFace LLM deployment
Env varPurposeNotes
SM_VLLM_MODELHF model ID (e.g. Qwen/Qwen3-0.6B) or /opt/ml/model if loading from S3—
SM_VLLM_HOSTMust be 0.0.0.0Otherwise vLLM binds localhost only, ping fails, container dies before logs. Top cause of mystery failures with this image.
SM_VLLM_TRUST_REMOTE_CODEtrue for Qwen and several recent architecturesSet unconditionally — downside negligible, upside is the model loads.
HUGGING_FACE_HUB_TOKENHF tokenRequired for gated models.
Tuning (optional)
Env varPurpose
SM_VLLM_MAX_MODEL_LENMax sequence length — set this; defaults can be wrong for fine-tunes
SM_VLLM_GPU_MEMORY_UTILIZATIONFloat 0.0–1.0, ~0.9 reasonable
SM_VLLM_TENSOR_PARALLEL_SIZEGPU count for multi-GPU instances
SM_VLLM_DTYPEauto, bfloat16, float16

Any vLLM CLI flag works — uppercase, replace dashes with underscores, prepend SM_VLLM_.

Configuring TEI

Simpler env contract than vLLM:

Env varPurposeRequired
HF_MODEL_IDHF model ID (e.g. BAAI/bge-large-en-v1.5) or /opt/ml/modelYes
HF_TOKENHF auth tokenOnly for gated models
MAX_BATCH_TOKENSMax tokens per batch (default 16384)No
MAX_CLIENT_BATCH_SIZEMax requests per client batch (default 32)No

No host-binding to configure, no trust-remote-code flag. The architectures TEI supports (BERT, CamemBERT, RoBERTa, XLM-RoBERTa, NomicBert, JinaBert, JinaCodeBert, Mistral, Qwen2/3, Gemma2/3, ModernBert) are baked into the image.

CUDA / instance compatibility

Critical and easy to get wrong:

CUDA in image tagDefault AMIWith al2-ami-sagemaker-inference-gpu-3-1
cu124 / cu128g5, g6, p5 all work(not needed)
cu129g6, p5; g5 fails (driver mismatch → CannotStartContainerError)expected to fix g5 (unverified)
cu130+fails everywhere — AMI flag is mandatoryg5, g6, p5 all work (cu130-on-g5 verified June 2026)

The driver comes from the host AMI, not the instance family — so passing the gpu-3-1 AMI (which vLLM cu130 images require anyway) also makes ml.g5.* viable for cu129+ images.

VPC / NAT gateway problem

SageMaker endpoints inside a VPC without a NAT gateway can't pull from public.ecr.aws. The deployment fails with an image-pull error that doesn't mention "VPC" or "egress".

For images on AWS's regional ECR (everything in the catalog): SageMaker reaches them through built-in routing, no NAT needed. Use the regional URI pattern (<account>.dkr.ecr.<region>.amazonaws.com/...), not the public.ecr.aws/... pattern.

For images requiring public.ecr.aws access (less common): mirror to a private ECR repo in your account with scripts/mirror_image.py (cross-platform; needs Docker + the aws CLI). Run it from the shell where the AWS CLI works.

bash
# macOS / Linux
PRIVATE_URI=$(python3 scripts/mirror_image.py \
    public.ecr.aws/deep-learning-containers/vllm:<tag> \
    vllm-mirror)
powershell
# Windows (PowerShell) — capture stdout into a variable
$PRIVATE_URI = python scripts\mirror_image.py `
    public.ecr.aws/deep-learning-containers/vllm:<tag> vllm-mirror

When the catalog page won't render, is stale, or is wrong

The page won't render / fetch returns junk: the catalog page is JavaScript-heavy and some fetch tools get an empty shell. Fallbacks, in order:

  1. The catalog's source data on GitHub — the page is generated from one YAML file per version, listing exact tags, CUDA, and Python versions. List a family's files, then fetch the newest:
    bash
    curl -s https://api.github.com/repos/aws/deep-learning-containers/contents/docs/src/data/huggingface-vllm
    curl -s https://raw.githubusercontent.com/aws/deep-learning-containers/main/docs/src/data/huggingface-vllm/0.21.0-gpu-sagemaker.yml
    Directory names match ECR repos (huggingface-vllm, huggingface-vllm-omni, huggingface-tei, vllm, djl-inference, ...).
  2. Query ECR directly for current tags in the target region (works with credentials that can read the DLC registry; if it returns AccessDenied, use the YAML files):
    bash
    aws ecr describe-images --registry-id 763104351884 --repository-name huggingface-vllm \
        --region <region> --query 'sort_by(imageDetails,&imagePushedAt)[-5:].imageTags' --output json
  3. Release notes on the DLC GitHub repo.

A tag was just released and isn't on the page yet: rare; AWS updates the page on each release. Check the release notes above.

An architecture you need isn't supported by the listed image yet: for TEI specifically, you can mirror the upstream image from GHCR (ghcr.io/huggingface/text-embeddings-inference:<version>) into private ECR and pass the resulting URI directly to deploy.py --image-uri. Same mirror_image.py script.

Known-broken images (last checked July 2026)

ImageDefectUse instead
huggingface-pytorch-inference GPU tags — all recent ones tested (PT 2.3–2.6, cu121/cu124, transformers 4.48–5.5.3)ImportError: libtorch_cuda.so: undefined symbol: ncclCommResume at import torch. The NCCL bundled in the image is older than what torch links against — a packaging defect inside the container, on g5 and g6, regardless of AMI, model, or inference code. The MMS Java front-end keeps answering /ping, so the endpoint can reach InService while the Python worker crash-loops and serves nothing.DJL Inference (bundles its own complete CUDA/NCCL stack) or BYOC. CPU tags are unaffected.

Re-check when AWS publishes new huggingface-pytorch-inference GPU tags — remove the row once a fixed image is confirmed.

General fallback rule: when an HF DLC fails with CUDA/NCCL linker errors, switch to DJL Inference rather than iterating over sibling tags — the defect class is per-repo, not per-tag (three different tags were tried for the case above; all broken).

Related HF Hub gotcha: older DLCs can fail model download with 403 Forbidden from HF's XET CDN (their bundled huggingface_hub predates XET auth). Set HF_HUB_ENABLE_HF_TRANSFER=0 to force the standard download path, or pre-stage weights in S3.

Hub download time at first boot

Loading the model from HF Hub happens inside the container after the endpoint starts — expect 5–15+ minutes before InService even for small models, longer for multi-GB ones. A slow first boot is not a failure; don't tear down or re-diagnose before the deploy script's 30-minute wait expires.

For production or repeated deployments, pre-stage the weights in S3 and pass --model-s3-uri to deploy.py (the model then loads from /opt/ml/model) — faster, immune to Hub rate limits/outages, and no HUGGING_FACE_HUB_TOKEN needed at runtime.

© waybarrios, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 2 other files (scripts, references) in skills/hf-cloud-serving-image-selection of waybarrios/opencode-power-pack.

  • SKILL.md
  • references/model-to-image.md
  • scripts/mirror_image.py

Open the folder on GitHubat commit 9dccb6d

Compare with similar skills

Hf Cloud Serving Image Selection next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Hf Cloud Serving Image Selection compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Hf Cloud Serving Image Selection this skillwaybarrios/opencode-power-pack533—~4.3kAutomated safety check: PassApache-2.0
SageMaker Serving Image Selectionhuggingface/skills11k1 repos~4.6kAutomated safety check: PassApache-2.0
Rtvi Byom PortingNVIDIA-AI-Blueprints/video-search-and-summarization1.9k—~1.4kAutomated safety check: PassApache-2.0
SageMaker Deployment Plannerhuggingface/skills11k1 repos~2.1kAutomated safety check: PassApache-2.0
SageMaker Production Defaultshuggingface/skills11k1 repos~6.9kAutomated safety check: PassApache-2.0
Python Environment Setup for SageMakerhuggingface/skills11k2 repos~1.7kAutomated safety check: PassApache-2.0

Similar skills

  • Official

    Chooses the right serving container and current image URI for deploying a Hugging Face model to a SageMaker endpoint, preferring Hugging Face images over generic ones.

    11k GitHub starsUsed in 1 repo~4.6k tokens
    AI & LLM EngineeringAuto-check passed
  • Rtvi Byom Porting

    NVIDIA-AI-Blueprints/video-search-and-summarization

    A skill your agent uses when adding, debugging, or validating a bring-your-own VLM in VSS RT-VLM, including custom Hugging Face or NGC checkpoints, vLLM adapters or plugins, model shims, and…

    1.9k GitHub stars~1.4k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Official

    Entry point for hosting a model on Amazon SageMaker: asks a few questions, picks a deployment pathway and hands off to the specialist skills.

    11k GitHub starsUsed in 1 repo~2.1k tokens
    AI & LLM EngineeringAuto-check passed
  • Official

    Deploys SageMaker endpoints with autoscaling, CloudWatch alarms and tags on by default, using scripts for real-time, scale-to-zero and async setups.

    11k GitHub starsUsed in 1 repo~6.9k tokens
    DevOps & CloudAuto-check passed
  • Official

    Sets up an isolated Python environment with a supported interpreter and current boto3 before any SageMaker deployment, training or AWS automation code runs.

    11k GitHub starsUsed in 2 repos~1.7k tokens
    AI & LLM EngineeringAuto-check passed
  • Official

    Runs evaluations of Hugging Face Hub models on local hardware with inspect-ai or lighteval, and helps choose between vLLM, Transformers and accelerate backends.

    11k GitHub starsUsed in 2 repos~1.6k tokens
    AI & LLM EngineeringAuto-check passed

More from waybarrios/opencode-power-pack

All 32 skills in this repo
  • Hf Cloud Sagemaker Iam Preflight

    waybarrios/opencode-power-pack

    Verify or select a SageMaker execution role before creating models, endpoints, or training jobs.

    533 GitHub stars~1.6k tokensUpdated 2 days ago
    Auto-check passed
  • Huggingface LLM Trainer

    waybarrios/opencode-power-pack

    Train or fine-tune language models with TRL or Unsloth on Hugging Face Jobs, including SFT, DPO, GRPO, reward models, and GGUF conversion.

    533 GitHub stars~3k tokensUpdated 2 days ago
    Auto-check passed
  • Huggingface Vision Trainer

    waybarrios/opencode-power-pack

    Train object-detection, image-classification, or SAM segmentation models on Hugging Face Jobs.

    533 GitHub stars~2.7k tokensUpdated 2 days ago
    Auto-check passed
  • Codeql

    waybarrios/opencode-power-pack

    Run CodeQL database creation and security queries, add data-extension models, or process CodeQL SARIF.

    533 GitHub starsUsed in 2 repos~3.7k tokens
    Auto-check passed
  • Semgrep

    waybarrios/opencode-power-pack

    Run Semgrep static analysis across a codebase, optionally using Semgrep Pro for cross-file taint analysis.

    533 GitHub stars~2.4k tokensUpdated 2 days ago
    Auto-check passed
  • Insecure Defaults

    waybarrios/opencode-power-pack

    Detects fail-open insecure defaults (hardcoded secrets, weak auth, permissive security) that allow apps to run insecurely in production.

    533 GitHub starsUsed in 2 repos~1.3k tokens
    Auto-check passed

Questions about Hf Cloud Serving Image Selection

What does Hf Cloud Serving Image Selection do?

Select and verify the current region-specific serving container URI for a SageMaker model deployment. Hf Cloud Serving Image Selection is an agent skill from waybarrios/opencode-power-pack. Select and verify the current region-specific serving container URI for a SageMaker model deployment.

When should I use Hf Cloud Serving Image Selection?

Hf Cloud Serving Image Selection fits situations like: tasks that involve Model hubs and datasets; tasks that involve LLM inference and serving; tasks that involve Deployment.

How do I install Hf Cloud Serving Image Selection in Claude Code?

Run `npx skills add waybarrios/opencode-power-pack --skill hf-cloud-serving-image-selection -a claude-code`. Or copy the skill folder (skills/hf-cloud-serving-image-selection in waybarrios/opencode-power-pack) into .claude/skills/hf-cloud-serving-image-selection in your project. Claude Code loads it when a task matches its description.

How do I install Hf Cloud Serving Image Selection in Codex?

Run `npx skills add waybarrios/opencode-power-pack --skill hf-cloud-serving-image-selection -a codex`. Or copy the skill folder (skills/hf-cloud-serving-image-selection in waybarrios/opencode-power-pack) into .agents/skills/hf-cloud-serving-image-selection in your project. Codex loads it when a task matches its description.

Can I use Hf Cloud Serving Image Selection in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add waybarrios/opencode-power-pack --skill hf-cloud-serving-image-selection -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/hf-cloud-serving-image-selection, .gemini/skills/hf-cloud-serving-image-selection, .github/skills/hf-cloud-serving-image-selection and .opencode/skills/hf-cloud-serving-image-selection in your project.

What does Hf Cloud Serving Image Selection need to run?

Going by SKILL.md and its folder, Hf Cloud Serving Image Selection needs Python for the scripts in its folder, the command-line tools its instructions call (curl, python3 and aws) and credentials named HUGGING_FACE_HUB_TOKEN and HF_TOKEN. Our summary lists: Python 3; A credential in HUGGING_FACE_HUB_TOKEN.

Does Hf Cloud Serving Image Selection access the network?

SKILL.md names 5 domains. In commands or code: huggingface.co, api.github.com and raw.githubusercontent.com; the agent is likely to contact these when it follows the instructions. As links in the text: aws.github.io and github.com. This is read from the text; nothing was executed.

Is Hf Cloud Serving Image Selection safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Hf Cloud Serving Image Selection use?

Hf Cloud Serving Image Selection is published under the Apache-2.0 licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Hf Cloud Serving Image Selection use?

About 4.3k tokens (SKILL.md is roughly 17k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 2.7k tokens, read only when the agent opens those files.

What are the alternatives to Hf Cloud Serving Image Selection?

Skills that share tags, products or a category with Hf Cloud Serving Image Selection: SageMaker Serving Image Selection (huggingface/skills, 11k stars), Rtvi Byom Porting (NVIDIA-AI-Blueprints/video-search-and-summarization, 1.9k stars), SageMaker Deployment Planner (huggingface/skills, 11k stars) and SageMaker Production Defaults (huggingface/skills, 11k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Hf Cloud Serving Image Selection?

waybarrios (a GitHub user) maintains it in waybarrios/opencode-power-pack, which has 533 GitHub stars. The repository holds 32 skills in this directory. The repository was last updated on October 6, 2026.

Source: waybarrios/opencode-power-pack on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.