Agent skill

Collecting Infrastructure Metrics

by jeremylongshore in jeremylongshore/tons-of-skills-marketplace

Collect comprehensive infrastructure performance metrics across compute, storage, network, containers, load balancers, and databases.

MITAuto-check passedDevOps & Cloud

Install Collecting Infrastructure Metrics

skills CLI
$ npx skills add jeremylongshore/tons-of-skills-marketplace --skill collecting-infrastructure-metrics -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install jeremylongshore/tons-of-skills-marketplace collecting-infrastructure-metrics --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/jeremylongshore/tons-of-skills-marketplace.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/.curated/collecting-infrastructure-metrics .claude/skills/collecting-infrastructure-metrics && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
collecting-infrastructure-metrics
GitHub stars
2.8k
Token cost
~1.2k tokens
SKILL.md length
499 words
Files
4 (incl. scripts, references, assets)
Skills in repo
3,342
Repo updated
First seen
Licence
MIT

At a glance

Collect comprehensive infrastructure performance metrics across compute, storage, network, containers, load balancers, and databases.

  • Works in 4 steps: Identify Infrastructure Layers:… → Configure Metrics Collection: Sets up… → Aggregate Metrics: Configures central… → …
  • Monitoring system performance
  • SKILL.md covers Overview, How It Works, When to Use This Skill and Examples, plus 7 more sections
  • Troubleshooting infrastructure issues

What it does

Collecting Infrastructure Metrics is an agent skill from jeremylongshore/tons-of-skills-marketplace. Collect comprehensive infrastructure performance metrics across compute, storage, network, containers, load balancers, and databases. Use when monitoring system performance or troubleshooting infrastructure issues. Trigger with phrases like "collect infrastructure metrics", "monitor server performance", or "track system resources".

Its SKILL.md is about 1.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 6 other files, including scripts, reference files and assets (for example `assets/README.md`, `references/README.md` and `scripts/README.md`). Compatibility notes: Designed for Claude Code

It sits in DevOps & Cloud, covering Cloud networking and OKRs and executive reporting. It works with Prometheus and Datadog. The repository describes itself as: Model-agnostic agent-skills platform with a harness-free canonical layer, verified adapters, and the ccpi package manager. Explore at tonsofskills.com. The licence is MIT.

When your agent uses it

  • Monitoring system performance
  • Troubleshooting infrastructure issues
  • With phrases like collect infrastructure metrics
  • Monitor server performance

Example prompts

  • “collect infrastructure metrics”
  • “monitor server performance”
  • “track system resources”
  • “/collecting-infrastructure-metrics”

Requirements

  • Compatibility (from SKILL.md): Designed for Claude Code
  • Pre-approved tools (allowed-tools): Read, Write, Edit, Grep, Glob, Bash(metrics:*), Bash(monitoring:*), Bash(system:*)

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Identify Infrastructure Layers: Determines the infrastructure layers to monitor (compute, storage, network, containers, load balancers…
  2. Configure Metrics Collection: Sets up agents (Prometheus, Datadog, CloudWatch) to collect metrics from the identified layers.
  3. Aggregate Metrics: Configures central aggregation of the collected metrics for analysis and visualization.
  4. Create Dashboards: Generates infrastructure dashboards for health monitoring, performance analysis, and capacity tracking.

What it can do on your machine

Read from SKILL.md and the folder at commit cfae287. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Read
    • Write
    • Edit
    • Grep
    • Glob
    • Bash(metrics:*)
    • Bash(monitoring:*)
    • Bash(system:*)

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/, which the agent can run.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

  • Compatibility

    Designed for Claude Code

    From compatibility in the SKILL.md frontmatter.

Context cost

Collecting Infrastructure Metrics loads about 1.2k tokens when it runs, and up to ~1.2k if it reads all its reference files. Until then it costs about 92 tokens; SKILL.md has 499 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~92
When it runs · the whole SKILL.md, loaded when a task matches
~1.2k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~1.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from jeremylongshore/tons-of-skills-marketplace at commit cfae287, republished under its MIT licence (© jeremylongshore). 499 words, ~1,170 tokens.

Download SKILL.mdSave it as .claude/skills/collecting-infrastructure-metrics/SKILL.md (or your agent's skills folder). This skill also uses 3 other files; get the full folder from GitHub.
name
collecting-infrastructure-metrics
description
Collect comprehensive infrastructure performance metrics across compute, storage, network, containers, load balancers, and databases. Use when monitoring system performance or troubleshooting infrastructure issues. Trigger with phrases like "collect infrastructure metrics", "monitor server performance", or "track system resources".
allowed-tools
Read, Write, Edit, Grep, Glob, Bash(metrics:*), Bash(monitoring:*), Bash(system:*)
compatibility
Designed for Claude Code
version
1.20.0
license
MIT
author
Jeremy Longshore <jeremy@intentsolutions.io>
tags
performance, database, monitoring

Infrastructure Metrics Collector

Collect and centralize infrastructure metrics across compute, storage, network, containers, load balancers, and databases using Prometheus, Datadog, or CloudWatch.

Overview

This skill automates the process of setting up infrastructure metrics collection. It identifies key performance indicators (KPIs) across various infrastructure layers, configures agents to collect these metrics, and assists in setting up central aggregation and visualization.

How It Works

  1. Identify Infrastructure Layers: Determines the infrastructure layers to monitor (compute, storage, network, containers, load balancers, databases).
  2. Configure Metrics Collection: Sets up agents (Prometheus, Datadog, CloudWatch) to collect metrics from the identified layers.
  3. Aggregate Metrics: Configures central aggregation of the collected metrics for analysis and visualization.
  4. Create Dashboards: Generates infrastructure dashboards for health monitoring, performance analysis, and capacity tracking.

When to Use This Skill

This skill activates when you need to:

  • Monitor the performance of your infrastructure.
  • Identify bottlenecks in your system.
  • Set up dashboards for real-time monitoring.

Examples

Example 1: Setting up basic monitoring

User request: "Collect infrastructure metrics for my web server."

The skill will:

  1. Identify compute, storage, and network layers relevant to the web server.
  2. Configure Prometheus to collect CPU, memory, disk I/O, and network bandwidth metrics.
Example 2: Troubleshooting database performance

User request: "I'm seeing slow database queries. Can you help me monitor the database performance?"

The skill will:

  1. Identify the database layer and relevant metrics such as connection pool usage, replication lag, and cache hit rates.
  2. Configure Datadog to collect these metrics and create a dashboard to visualize performance trends.

Best Practices

  • Agent Selection: Choose the appropriate agent (Prometheus, Datadog, CloudWatch) based on your existing infrastructure and monitoring tools.
  • Metric Granularity: Balance the granularity of metrics collection with the storage and processing overhead. Collect only the essential metrics for your use case.
  • Alerting: Configure alerts based on thresholds for key metrics to proactively identify and address performance issues.
Show full SKILL.md (187 more words)Show less

Integration

This skill can be integrated with other plugins for deployment, configuration management, and alerting to provide a comprehensive infrastructure management solution. For example, it can be used with a deployment plugin to automatically configure metrics collection after deploying new infrastructure.

Prerequisites

  • Access to infrastructure monitoring systems (Prometheus, Datadog, CloudWatch)
  • System permissions for metrics agent installation
  • Network access to monitored infrastructure components
  • Storage for metrics data in ${CLAUDE_SKILL_DIR}/metrics/

Instructions

  1. Identify infrastructure layers to monitor (compute, storage, network, databases)
  2. Select appropriate metrics collection agent based on environment
  3. Configure agent with target endpoints and metric types
  4. Set up central aggregation for collected metrics
  5. Create dashboards for visualization
  6. Configure alerts for critical metrics thresholds

Output

  • Metrics collection configuration files
  • Agent installation and setup scripts
  • Dashboard definitions for infrastructure monitoring
  • Metric export configurations
  • Alert rules for critical thresholds

Error Handling

If metrics collection fails:

  • Verify agent installation and permissions
  • Check network connectivity to targets
  • Validate authentication credentials
  • Review firewall and security group rules
  • Confirm metric endpoint availability

Resources

  • Prometheus documentation for metric collection
  • Datadog agent configuration guides
  • AWS CloudWatch metrics reference
  • Infrastructure monitoring best practices

© jeremylongshore, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 3 other files (scripts, references, assets) in skills/.curated/collecting-infrastructure-metrics of jeremylongshore/tons-of-skills-marketplace.

  • SKILL.md
  • assets/README.md
  • references/README.md
  • scripts/README.md

Open the folder on GitHubat commit cfae287

Compare with similar skills

Collecting Infrastructure Metrics next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Collecting Infrastructure Metrics compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Collecting Infrastructure Metrics this skilljeremylongshore/tons-of-skills-marketplace2.8k—~1.2kAutomated safety check: PassMIT
Redis Observabilityredis/agent-skills1662 repos~911Automated safety check: PassMIT
KubeSphere Gateway Managementkubesphere/kubesphere17k—~2.9kAutomated safety check: PassCustom licence
Frontmcp Observabilityagentfront/frontmcp146—~4.6kAutomated safety check: PassApache-2.0
Monitoring Observabilityahmedasmar/devops-claude-skills203—~3.9kAutomated safety check: PassNone
Observability Architecturemajiayu000/litellm-rs118—~1.3kAutomated safety check: PassMIT

Similar skills

  • Redis Observability

    redis/agent-skills

    Official

    Redis observability guidance — which metrics to monitor (memory, connections, hit ratio, ops/sec, rejected connections), which built-in commands to reach for during incident triage (SLOWLOG, INFO…

    166 GitHub starsUsed in 2 repos~911 tokens
    DevOps & CloudAuto-check passed
  • KubeSphere Gateway Management

    kubesphere/kubesphere

    Installs, uninstalls, checks and troubleshoots the KubeSphere Gateway extension built on ingress-nginx, including gateways stuck in bad states and Helm or pod failures.

    17k GitHub stars~2.9k tokensUpdated 2 mo ago
    DevOps & CloudAuto-check passed
  • Frontmcp Observability

    agentfront/frontmcp

    A skill your agent uses when adding tracing, structured logging, metrics, or monitoring to a FrontMCP server.

    146 GitHub stars~4.6k tokensUpdated today
    DevOps & CloudAuto-check passed
  • Monitoring Observability

    ahmedasmar/devops-claude-skills

    Monitoring and observability strategy, implementation, and troubleshooting.

    203 GitHub stars~3.9k tokensUpdated 6 mo ago
    DevOps & CloudAuto-check passed
  • Observability Architecture

    majiayu000/litellm-rs

    LiteLLM-RS Observability Architecture. An agent skill from majiayu000/litellm-rs.

    118 GitHub stars~1.3k tokensUpdated today
    DevOps & CloudAuto-check passed
  • Cost Export

    ruvnet/ruflo

    Export cost-tracking telemetry in Prometheus textfile or webhook JSON formats — for external observability (Grafana, Datadog, custom dashboards)

    74k GitHub stars~687 tokensUpdated yesterday
    DevOps & CloudAuto-check: notes

More from jeremylongshore/tons-of-skills-marketplace

All 3,342 skills in this repo
  • Performing Security Code Review

    jeremylongshore/tons-of-skills-marketplace

    Execute this skill enables AI assistant to conduct a security-focused code review using the security-agent plugin.

    2.8k GitHub starsUsed in 2 repos~1.3k tokens
    Auto-check: notes
  • Adapting Transfer Learning Models

    jeremylongshore/tons-of-skills-marketplace

    Build this skill automates the adaptation of pre-trained machine learning models using transfer learning techniques.

    2.8k GitHub stars~1.1k tokensUpdated today
    Auto-check passed
  • Agent Context Loader

    jeremylongshore/tons-of-skills-marketplace

    Execute proactive auto-loading: automatically detects and loads agents.md files.

    2.8k GitHub stars~1.1k tokensUpdated today
    Auto-check passed
  • Aggregating Performance Metrics

    jeremylongshore/tons-of-skills-marketplace

    Aggregate and centralize performance metrics from applications, systems, databases, caches, and services.

    2.8k GitHub stars~1.2k tokensUpdated today
    Auto-check passed
  • Analyzing Capacity Planning

    jeremylongshore/tons-of-skills-marketplace

    Execute this skill enables AI assistant to analyze capacity requirements and plan for future growth.

    2.8k GitHub stars~947 tokensUpdated today
    Auto-check passed
  • Analyzing Database Indexes

    jeremylongshore/tons-of-skills-marketplace

    Process use when you need to work with database indexing. An agent skill from jeremylongshore/tons-of-skills-marketplace.

    2.8k GitHub stars~2k tokensUpdated today
    Auto-check passed

Categories

Questions about Collecting Infrastructure Metrics

What does Collecting Infrastructure Metrics do?

Collect comprehensive infrastructure performance metrics across compute, storage, network, containers, load balancers, and databases. Collecting Infrastructure Metrics is an agent skill from jeremylongshore/tons-of-skills-marketplace. Collect comprehensive infrastructure performance metrics across compute, storage, network, containers, load balancers, and databases.

When should I use Collecting Infrastructure Metrics?

Collecting Infrastructure Metrics fits situations like: monitoring system performance; troubleshooting infrastructure issues; with phrases like collect infrastructure metrics; monitor server performance.

How do I install Collecting Infrastructure Metrics in Claude Code?

Run `npx skills add jeremylongshore/tons-of-skills-marketplace --skill collecting-infrastructure-metrics -a claude-code`. Or copy the skill folder (skills/.curated/collecting-infrastructure-metrics in jeremylongshore/tons-of-skills-marketplace) into .claude/skills/collecting-infrastructure-metrics in your project. Claude Code loads it when a task matches its description.

How do I install Collecting Infrastructure Metrics in Codex?

Run `npx skills add jeremylongshore/tons-of-skills-marketplace --skill collecting-infrastructure-metrics -a codex`. Or copy the skill folder (skills/.curated/collecting-infrastructure-metrics in jeremylongshore/tons-of-skills-marketplace) into .agents/skills/collecting-infrastructure-metrics in your project. Codex loads it when a task matches its description.

Can I use Collecting Infrastructure Metrics in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add jeremylongshore/tons-of-skills-marketplace --skill collecting-infrastructure-metrics -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/collecting-infrastructure-metrics, .gemini/skills/collecting-infrastructure-metrics, .github/skills/collecting-infrastructure-metrics and .opencode/skills/collecting-infrastructure-metrics in your project.

What does Collecting Infrastructure Metrics need to run?

SKILL.md names no scripts, command-line tools or credentials: Collecting Infrastructure Metrics is instructions for the agent only. Its frontmatter pre-approves these tools: Read, Write, Edit, Grep, Glob, Bash(metrics:*), Bash(monitoring:*), Bash(system:*). Compatibility (from SKILL.md): Designed for Claude Code.

Does Collecting Infrastructure Metrics access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Collecting Infrastructure Metrics safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Collecting Infrastructure Metrics use?

Collecting Infrastructure Metrics is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Collecting Infrastructure Metrics use?

About 1.2k tokens (SKILL.md is roughly 4.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 19 tokens, read only when the agent opens those files.

What are the alternatives to Collecting Infrastructure Metrics?

Skills that share tags, products or a category with Collecting Infrastructure Metrics: Redis Observability (redis/agent-skills, 166 stars), KubeSphere Gateway Management (kubesphere/kubesphere, 17k stars), Frontmcp Observability (agentfront/frontmcp, 146 stars) and Monitoring Observability (ahmedasmar/devops-claude-skills, 203 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Collecting Infrastructure Metrics?

jeremylongshore (a GitHub user) maintains it in jeremylongshore/tons-of-skills-marketplace, which has 2,827 GitHub stars. The repository holds 3,342 skills in this directory. The repository was last updated on October 10, 2026.

Source: jeremylongshore/tons-of-skills-marketplace on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.