Querying AWS Sagemaker Catalog
aws/agent-toolkit-for-aws
Runs SQL analytics on SageMaker Catalog asset metadata tables exported as Apache Iceberg in S3 Tables.
A skill your agent uses to investigate and troubleshoot AWS Glue problems by analyzing ETL jobs, crawlers, connections, Data Catalog, DPU utilization, Spark execution, and job bookmarks following…
$ npx skills add Kilo-Org/kilo-marketplace --skill glue-diagnostics -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install Kilo-Org/kilo-marketplace glue-diagnostics --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/Kilo-Org/kilo-marketplace.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/glue-diagnostics .claude/skills/glue-diagnostics && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "glue-diagnostics" agent skill from https://github.com/Kilo-Org/kilo-marketplace/tree/main/skills/glue-diagnostics into .claude/skills/glue-diagnostics/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "glue-diagnostics", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/Kilo-Org/kilo-marketplace/tree/main/skills/glue-diagnosticsType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add Kilo-Org/kilo-marketplace --skill glue-diagnostics -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install Kilo-Org/kilo-marketplace glue-diagnostics --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Kilo-Org/kilo-marketplace.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/glue-diagnostics .agents/skills/glue-diagnostics && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "glue-diagnostics" agent skill from https://github.com/Kilo-Org/kilo-marketplace/tree/main/skills/glue-diagnostics into .agents/skills/glue-diagnostics/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "glue-diagnostics", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add Kilo-Org/kilo-marketplace --skill glue-diagnostics -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install Kilo-Org/kilo-marketplace glue-diagnostics --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Kilo-Org/kilo-marketplace.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/glue-diagnostics .cursor/skills/glue-diagnostics && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "glue-diagnostics" agent skill from https://github.com/Kilo-Org/kilo-marketplace/tree/main/skills/glue-diagnostics into .cursor/skills/glue-diagnostics/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "glue-diagnostics", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/Kilo-Org/kilo-marketplace.git --path skills/glue-diagnostics--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add Kilo-Org/kilo-marketplace --skill glue-diagnostics -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install Kilo-Org/kilo-marketplace glue-diagnostics --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Kilo-Org/kilo-marketplace.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/glue-diagnostics .gemini/skills/glue-diagnostics && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "glue-diagnostics" agent skill from https://github.com/Kilo-Org/kilo-marketplace/tree/main/skills/glue-diagnostics into .gemini/skills/glue-diagnostics/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "glue-diagnostics", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install Kilo-Org/kilo-marketplace glue-diagnosticsInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add Kilo-Org/kilo-marketplace --skill glue-diagnostics -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/Kilo-Org/kilo-marketplace.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/glue-diagnostics .github/skills/glue-diagnostics && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "glue-diagnostics" agent skill from https://github.com/Kilo-Org/kilo-marketplace/tree/main/skills/glue-diagnostics into .github/skills/glue-diagnostics/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "glue-diagnostics", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add Kilo-Org/kilo-marketplace --skill glue-diagnostics -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install Kilo-Org/kilo-marketplace glue-diagnostics --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Kilo-Org/kilo-marketplace.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/glue-diagnostics .opencode/skills/glue-diagnostics && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "glue-diagnostics" agent skill from https://github.com/Kilo-Org/kilo-marketplace/tree/main/skills/glue-diagnostics into .opencode/skills/glue-diagnostics/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "glue-diagnostics", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
glue-diagnosticsA skill your agent uses to investigate and troubleshoot AWS Glue problems by analyzing ETL jobs, crawlers, connections, Data Catalog, DPU utilization, Spark execution, and job bookmarks following…
Glue Diagnostics is an agent skill from Kilo-Org/kilo-marketplace. Use this skill to investigate and troubleshoot AWS Glue problems by analyzing ETL jobs, crawlers, connections, Data Catalog, DPU utilization, Spark execution, and job bookmarks following structured runbooks. Activate when: job failures, job timeouts, OOM errors, Spark executor or driver crashes, crawler failures, schema detection issues, partition problems, JDBC connection failures, VPC/subnet connectivity, S3 endpoint access, Data Catalog sync issues, schema evolution conflicts, DPU sizing problems, shuffle…
Its SKILL.md is about 2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 28 other files, including reference files (for example `README.md`, `references/A1-job-failures.md` and `references/A2-job-timeout.md`). Compatibility notes: Requires AWS CLI or SDK access with Glue, S3, CloudWatch Logs, IAM, EC2 (for VPC/connections), and optionally KMS permissions.
It sits in Data & Analytics, covering Web scraping, Data governance and Data pipelines and ETL. It works with Amazon Web Services. The repository describes itself as: Kilo Marketplace - A curated collection of Skills, MCP Servers, and Modes for enhancing AI agent capabilities across the Kilo ecosystem—including Kilo Code (VS Code extension)… The licence is MIT.
2 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit ff51758. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
awsFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use aws, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Requires AWS CLI or SDK access with Glue, S3, CloudWatch Logs, IAM, EC2 (for VPC/connections), and optionally KMS permissions.
From compatibility in the SKILL.md frontmatter.
Glue Diagnostics loads about 2k tokens when it runs, and up to ~38k if it reads all its reference files. Until then it costs about 200 tokens; SKILL.md has 753 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from Kilo-Org/kilo-marketplace at commit ff51758, republished under its MIT licence (© Kilo-Org). 753 words, ~1,985 tokens.
.claude/skills/glue-diagnostics/SKILL.md (or your agent's skills folder). This skill also uses 27 other files; get the full folder from GitHub.Any AWS Glue investigation where the console alone is insufficient — job failures, OOM errors, Spark crashes, crawler schema misdetection, connection timeouts, Data Catalog drift, DPU under/over-provisioning, data skew, bookmark corruption, or Glue Studio generation errors.
aws glue get-job --name <job-name>
aws glue get-job-run --job-name <job-name> --run-id <run-id>
aws glue batch-get-jobs --job-names <job1> <job2>
aws glue get-crawler --name <crawler-name>
aws glue get-connection --name <connection-name>
aws logs filter-log-events --log-group-name /aws-glue/jobs/logs-v2 --log-stream-name-prefix <run-id>aws glue get-job-runs --job-name <job-name> --max-results 10
aws glue get-crawler-metrics --crawler-name-list <crawler-name>
aws glue get-databases
aws glue get-tables --database-name <db-name>
aws glue get-partitions --database-name <db-name> --table-name <table-name>
aws glue get-job-bookmark --job-name <job-name>
aws cloudwatch get-metric-statistics --namespace Glue --metric-name glue.driver.aggregate.bytesRead --dimensions Name=JobName,Value=<job-name> --start-time <iso> --end-time <iso> --period 300 --statistics SumRead references/glue-guardrails.md before concluding on any Glue issue.
| Tool / API | When to use |
|---|---|
glue get-job | Job configuration, Glue version, DPU, worker type |
glue get-job-run | Specific run status, error message, execution time |
glue batch-get-jobs | Retrieve multiple job configs at once |
glue get-job-runs | Job run history, failure patterns |
glue get-crawler | Crawler config, targets, schedule, schema change policy |
glue get-crawler-metrics | Crawler runtime stats, tables created/updated |
glue get-connection | JDBC/network connection config, VPC, subnet |
glue get-databases / get-tables | Data Catalog metadata, schema definitions |
glue get-partitions | Partition metadata, partition keys, storage location |
glue get-job-bookmark | Bookmark state for incremental processing |
logs filter-log-events | Glue job CloudWatch logs for Spark errors |
cloudwatch get-metric-statistics | Glue job metrics (bytes read/written, DPU usage) |
| Worker Type | DPU | Memory | vCPU | Use Case |
|---|---|---|---|---|
| G.1X | 1 | 16 GB | 4 | Standard ETL, small-medium datasets |
| G.2X | 2 | 32 GB | 8 | Memory-intensive transforms, large joins |
| G.4X | 4 | 64 GB | 16 | ML transforms, very large datasets |
| G.8X | 8 | 128 GB | 32 | Massive datasets, complex aggregations |
| G.025X | 0.25 | 2 GB | 2 | Python shell jobs only |
| Z.2X | 2 | 32 GB | 8 | Ray jobs (Glue 4.0+) |
| Version | Spark | Python | Key Features |
|---|---|---|---|
| Glue 2.0 | 2.4 | 3.7 | Spark UI, no startup overhead |
| Glue 3.0 | 3.1 | 3.7 | Optimized shuffle, auto-scaling |
| Glue 4.0 | 3.3 | 3.10 | Ray support, Python 3.10, improved performance |
| Category | IDs | Covers |
|---|---|---|
| A — Jobs | A1–A4 | Job failures, timeout, OOM, Spark errors |
| B — Crawlers | B1–B3 | Crawler failures, schema detection, partition issues |
| C — Connections | C1–C3 | JDBC connection failures, VPC/subnet, S3 endpoint |
| D — Data Catalog | D1–D2 | Catalog sync issues, schema evolution |
| E — Performance | E1–E3 | DPU sizing, shuffle issues, data skew |
| F — ETL | F1–F3 | Transformation errors, bookmark issues, data quality |
| G — Security | G1–G2 | IAM permissions, encryption |
| H — Glue Studio | H1–H2 | Visual editor errors, job generation |
| Z — Catch-All | Z1 | General troubleshooting |
© Kilo-Org, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 27 other files (references) in skills/glue-diagnostics of Kilo-Org/kilo-marketplace.
Open the folder on GitHubat commit ff51758
Glue Diagnostics next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Glue Diagnostics this skillKilo-Org/kilo-marketplace | 189 | — | ~2k | Automated safety check: Pass | MIT | |
| Querying AWS Sagemaker Catalogaws/agent-toolkit-for-aws | 2.8k | — | ~2.6k | Automated safety check: Pass | Apache-2.0 | |
| Authoritative Data Harvesteryushui2022/MathModel-Skill | 452 | 1 repos | ~1.1k | Automated safety check: Pass | MIT | |
| Data Quality Frameworkswshobson/agents | 40k | 10 repos | ~1.1k | Automated safety check: Pass | MIT | |
| Ingesting Dataancoleman/ai-design-components | 526 | — | ~1.9k | Automated safety check: Pass | MIT | |
| Exploring Data Catalogaws/agent-toolkit-for-aws | 2.8k | 1 repos | ~2.6k | Automated safety check: Pass | Apache-2.0 |
aws/agent-toolkit-for-aws
Runs SQL analytics on SageMaker Catalog asset metadata tables exported as Apache Iceberg in S3 Tables.
yushui2022/MathModel-Skill
Finds authoritative public data sources for modeling tasks, prefers official APIs and bulk downloads, and outputs a reproducible fetch and cleaning plan with citations.
wshobson/agents
Sets up data quality checks with Great Expectations, dbt tests and data contracts, with checkpoints and pass-fail reports for pipelines.
ancoleman/ai-design-components
Data ingestion patterns for loading data from cloud storage, APIs, files, and streaming sources into databases.
aws/agent-toolkit-for-aws
Full inventory and audit of AWS Glue Data Catalog assets across S3 Tables, Redshift-federated, and remote Iceberg catalogs.
aws/agent-toolkit-for-aws
Authors and deploys MWAA workflow artifacts: Python Airflow DAGs for provisioned environments or YAML workflow files for Serverless.
Kilo-Org/kilo-marketplace
Sets up and maintains AzureML-ready Python projects as uv workspaces with devcontainers, a Makefile and job YAML, so local runs match cloud jobs and experiments stay reproducible.
Kilo-Org/kilo-marketplace
Creates, inspects, edits and runs Jupyter notebooks, scaffolding experiment or tutorial notebooks from templates and preferring a Jupyter MCP server over raw JSON edits.
Kilo-Org/kilo-marketplace
Takes a plain-language dashboard request through brand setup, data exploration, planning, an interactive HTML mock and a Tableau implementation spec.
Kilo-Org/kilo-marketplace
Ingest and transform data files (CSV/JSON/Parquet/Arrow IPC) into Elasticsearch with stream processing and custom transforms.
Kilo-Org/kilo-marketplace
A skill your agent uses when arranging Apache NiFi processors, process groups, ports, comments, numbering, crossing connections, dense fan-in/fan-out, or reusable readable canvas layouts.
Kilo-Org/kilo-marketplace
Render Cisco Data Fabric ingest-time routing workflows and Splunk Cloud Platform Ingest Processor setup plans with SPL2 pipelines, source types, destinations, lifecycle handoffs, queue and…
Works with
Categories
A skill your agent uses to investigate and troubleshoot AWS Glue problems by analyzing ETL jobs, crawlers, connections, Data Catalog, DPU utilization, Spark execution, and job bookmarks following…. Glue Diagnostics is an agent skill from Kilo-Org/kilo-marketplace. Use this skill to investigate and troubleshoot AWS Glue problems by analyzing ETL jobs, crawlers, connections, Data Catalog, DPU utilization, Spark execution, and job bookmarks following structured runbooks.
Glue Diagnostics fits situations like: investigate and troubleshoot AWS Glue problems by analyzing ETL jobs; DPU utilization; spark execution; job bookmarks following structured runbooks.
Run `npx skills add Kilo-Org/kilo-marketplace --skill glue-diagnostics -a claude-code`. Or copy the skill folder (skills/glue-diagnostics in Kilo-Org/kilo-marketplace) into .claude/skills/glue-diagnostics in your project. Claude Code loads it when a task matches its description.
Run `npx skills add Kilo-Org/kilo-marketplace --skill glue-diagnostics -a codex`. Or copy the skill folder (skills/glue-diagnostics in Kilo-Org/kilo-marketplace) into .agents/skills/glue-diagnostics in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Kilo-Org/kilo-marketplace --skill glue-diagnostics -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/glue-diagnostics, .gemini/skills/glue-diagnostics, .github/skills/glue-diagnostics and .opencode/skills/glue-diagnostics in your project.
Going by SKILL.md and its folder, Glue Diagnostics needs the command-line tools its instructions call (aws). Our summary lists: Python 3. Compatibility (from SKILL.md): Requires AWS CLI or SDK access with Glue, S3, CloudWatch Logs, IAM, EC2 (for VPC/connections), and optionally KMS permissions. .
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Glue Diagnostics is published under the MIT licence (from the LICENSE file in the skill folder). It allows redistribution, so the full SKILL.md is shown on this page.
About 2k tokens (SKILL.md is roughly 7.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 36k tokens, read only when the agent opens those files.
Skills that share tags, products or a category with Glue Diagnostics: Querying AWS Sagemaker Catalog (aws/agent-toolkit-for-aws, 2.8k stars), Authoritative Data Harvester (yushui2022/MathModel-Skill, 452 stars), Data Quality Frameworks (wshobson/agents, 40k stars) and Ingesting Data (ancoleman/ai-design-components, 526 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
Kilo-Org (a GitHub organization) maintains it in Kilo-Org/kilo-marketplace, which has 189 GitHub stars. The repository holds 87 skills in this directory. The repository was last updated on September 28, 2026.
Source: Kilo-Org/kilo-marketplace on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.