Search

Prometheus · For developers

108 skills found, page 2.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
49

Monitors and troubleshoots GKE TPU workloads, nodes, and node pools using GKE system metrics and PromQL.

google/skills21k—~1.9kAutomated safety check: PassApache-2.0yesterday
50

Initialize Navigator documentation structure in a project. An agent skill from qf-studio/navigator.

qf-studio/navigator355—~3kAutomated safety check: NotesMIT2 days ago
51

Monitoring, logging, and tracing implementation using OpenTelemetry as the unified standard.

ancoleman/ai-design-components525—~3kAutomated safety check: PassMIT10 mo ago
52

Generates Cloud Monitoring Server-Driven UI (SDUI) Widget and XyChart Protocol Buffer textprotos on Google Cloud from resolved PromQL or ListTimeSeries queries.

google/skills21k—~2.7kAutomated safety check: PassApache-2.0yesterday
53

Configures alerting policies in Terraform for Google Kubernetes Engine (GKE) clusters, workloads, and services using PromQL and Google Cloud Managed Service for Prometheus.

google/skills21k—~5.3kAutomated safety check: PassApache-2.0yesterday
54

Auto-scale LLM inference clusters on Kubernetes using KEDA, custom GPU metrics, and horizontal pod autoscaling.

sickn33/agentic-awesome-skills47k1 repo~2.1kAutomated safety check: PassMIT2 days ago
55

Set up metrics collection and visualization with Prometheus and Grafana.

sickn33/agentic-awesome-skills47k1 repo~2.7kAutomated safety check: PassMIT2 days ago
56

Configures Cloud Monitoring PromQL-based Service Level Objective (SLO) alerting policies on Google Cloud for resources registered in App Hub or individually specified.

google/skills21k—~3.1kAutomated safety check: PassApache-2.0yesterday
57

Guides Qdrant monitoring and observability setup. An agent skill from github/awesome-copilot.

github/awesome-copilot40k1 repo~276Automated safety check: PassMIT2 days ago
58

Analyze the experiment precompute result-consistency canary across prod-US and prod-EU, deep-dive any issues, and produce an actionable report.

PostHog/posthog40k—~3.5kAutomated safety check: PassUnknownyesterday
59

Investigates server/infrastructure metric anomalies in PostHog Metrics — from "this metric is rising/dropping/spiking" or a fired alert to a probable cause with evidence.

PostHog/posthog40k—~1.5kAutomated safety check: PassUnknownyesterday
60

Manages scaling for GKE workloads using HPA and VPA. An agent skill from google/skills.

google/skills21k—~1.3kAutomated safety check: PassApache-2.0yesterday
61

Diagnoses, predicts, and mitigates node disruptions during Compute Engine host maintenance and hardware or software maintenance events for GPU and TPU workloads on GKE.

google/skills21k—~1.6kAutomated safety check: PassApache-2.0yesterday
62

Configures GKE observability, including Cloud Logging, Cloud Monitoring, and managed Prometheus.

google/skills21k—~4.4kAutomated safety check: PassApache-2.0yesterday
63

Kubernetes deployment workflow for container orchestration, Helm charts, service mesh, and production-ready K8s configurations.

aiskillstore/marketplace4335 repos~839Automated safety check: PassNo licenceyesterday
64
64.BeylaOfficial

Auto-instrument an application's HTTP / gRPC / DB traffic with Grafana Beyla eBPF — no code changes, no SDK, no restart.

grafana/skills282—~1.1kAutomated safety check: PassApache-2.02 days ago
65
65.Fleet ManagementOfficial

Manage a fleet of Grafana Alloy collectors with Fleet Management — author Alloy pipelines once, target them via attribute matchers (env="production", regex region=~"us-."), push remotely via OpAMP…

grafana/skills282—~1.3kAutomated safety check: PassApache-2.02 days ago
66
66.Grafana OssOfficial

Configure Grafana OSS — provisions dashboards from YAML, sets up data sources (Prometheus / Loki / Tempo / Pyroscope), writes dashboard JSON with template variables, builds panel queries, assigns…

grafana/skills282—~1.5kAutomated safety check: PassApache-2.02 days ago
67
67.MimirOfficial

Stand up Grafana Mimir for horizontally scalable, multi-tenant, long-term Prometheus + OTLP metrics storage.

grafana/skills282—~1.2kAutomated safety check: PassApache-2.02 days ago
68
68.ML AIOfficial

Turn on AI + ML features in Grafana Cloud — Grafana Assistant (NL → PromQL/LogQL/TraceQL, dashboard build, incident investigation, MCP integration), Dynamic Alerting (Prophet forecasting + DBSCAN…

grafana/skills282—~1.3kAutomated safety check: PassApache-2.02 days ago
69

Canvas/A2UI inline network visualizations — topology maps, health dashboards, alert cards, change timelines, config diffs, path traces, and health scorecards rendered directly in the OpenClaw chat…

automateyournetwork/netclaw676—~941Automated safety check: PassApache-2.0yesterday
70

Reusable investigation patterns for AWS CloudWatch: Logs Insights query templates, alarm-to-deployment correlation, blast-radius narrowing decision tree, and PromQL-style metric query patterns for…

github/awesome-copilot40k—~2.6kAutomated safety check: PassMIT2 days ago
71

Monitor use when deploying monitoring stacks including Prometheus, Grafana, and Datadog.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.2kAutomated safety check: PassMITyesterday
72
72.Send DataOfficial

Sending telemetry data to Grafana Cloud — metrics via Prometheus remote write or OTLP, logs via Loki push or Alloy, traces via OTLP to Tempo, profiles via Pyroscope.

grafana/skills282—~1.4kAutomated safety check: PassApache-2.02 days ago
73

A skill your agent uses to deploy and operate a CollectX (clx) based DOCA telemetry collector on a host or BlueField — wiring providers / counters into the collector, running the collection daemon…

NVIDIA/skills3.6k—~2.8kAutomated safety check: PassApache-2.0yesterday
74

Multi-agent orchestration plugin for OpenCode. An agent skill from LeoYeAI/openclaw-master-skills.

LeoYeAI/openclaw-master-skills2.2k—~5.1kAutomated safety check: NotesMIT2 mo ago
75

Design, audit, and troubleshoot production monitoring and observability using user-impact checks, layered telemetry, USE/RED, SLI/SLO/SLA, error budgets, cardinality controls, actionable alerting…

AnastasiyaW/codex-claude-code-config154—~4.1kAutomated safety check: PassMITyesterday
76

Builds, configures, debugs, and optimizes AWS observability - operator-symptom questions and detecting Omni vs classic CloudWatch.

aws/agent-toolkit-for-aws2.8k—~7.2kAutomated safety check: PassApache-2.0yesterday
77

Probe a target for accidentally-public admin / debug / introspection endpoints — Spring Boot Actuator, Apache server-status, Prometheus metrics, GraphQL playground, Swagger UI, phpMyAdmin…

jeremylongshore/tons-of-skills-marketplace2.8k—~2kAutomated safety check: PassMITyesterday
78

A skill your agent uses when you need to implement or improve Java metrics observability with Micrometer — including meter design, naming/tag conventions, cardinality control…

jabrena/plinth447—~868Automated safety check: PassApache-2.03 days ago
79

Observability patterns for Python applications. An agent skill from aiskillstore/marketplace.

aiskillstore/marketplace4331 repo~1.3kAutomated safety check: PassNo licenceyesterday
80

This skill provides AWS cost optimization, monitoring, and operational best practices with integrated MCP servers for billing analysis, cost estimation, observability, and security assessment.

Microck/ordinary-claude-skills4041 repo~2.5kAutomated safety check: PassUnknown1 mo ago
81

Monitoring and observability patterns for Prometheus metrics, Grafana dashboards, Langfuse v4 LLM tracing (astype, scorecurrentspan, shouldexportspan, LangfuseMedia), and drift detection.

yonatangross/orchestkit292—~2.2kAutomated safety check: PassMITyesterday
82

Golang benchmarking, profiling, and performance measurement.

aiskillstore/marketplace4331 repo~3.4kAutomated safety check: PassMITyesterday
83

GORM v2 最佳实践与性能优化。适用于:代码审查、慢查询优化、N+1、连接池、 事务管理、分库分表、Prometheus/OTel监控、Session安全、Clause/Upsert、 缓存集成、BaseModel脚手架、SQL→struct生成、多租户隔离。

LeoYeAI/openclaw-master-skills2.2k—~3.4kAutomated safety check: PassMIT2 mo ago
84

OpenTelemetry, distributed tracing, structured logging, metrics (Prometheus, Grafana, Datadog).

TheBeardedBearSAS/claude-craft107—~547Automated safety check: PassMIT27 days ago
85

Monitoring and observability with OpenTelemetry, Prometheus, Grafana dashboards, and structured logging

rohitg00/awesome-claude-code-toolkit2.7k—~1.4kAutomated safety check: PassApache-2.05 mo ago
86

Observability: structured logs, metrics (RED/USE), tracing, SLO/SLI.

softspark/ai-toolkit179—~2.2kAutomated safety check: PassApache-2.03 days ago
87

Set up Apollo.io monitoring and observability. An agent skill from jeremylongshore/tons-of-skills-marketplace.

jeremylongshore/tons-of-skills-marketplace2.8k—~2.3kAutomated safety check: PassMITyesterday
88

Monitor ClickHouse with Prometheus metrics, Grafana dashboards, system table queries, and alerting for query performance, merge health, and resource usage.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.4kAutomated safety check: PassMITyesterday
89

Set up comprehensive observability for Deepgram integrations.

jeremylongshore/tons-of-skills-marketplace2.8k—~3kAutomated safety check: PassMITyesterday
90

Execute Deepgram production deployment checklist. An agent skill from jeremylongshore/tons-of-skills-marketplace.

jeremylongshore/tons-of-skills-marketplace2.8k—~2.3kAutomated safety check: PassMITyesterday
91

A skill your agent uses when you need production monitoring for an Intercom integration — instrumenting API calls with metrics and traces, standing up dashboards, or wiring alerts for error rate…

jeremylongshore/tons-of-skills-marketplace2.8k—~1.7kAutomated safety check: PassMITyesterday
92

Set up observability for Klaviyo integrations with metrics, traces, and alerts.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.5kAutomated safety check: PassMITyesterday
93

Set up comprehensive observability for Langfuse with metrics, dashboards, and alerts.

jeremylongshore/tons-of-skills-marketplace2.8k—~2.2kAutomated safety check: PassMITyesterday
94

Set up observability for Shopify app integrations with query cost tracking, rate limit monitoring, webhook delivery metrics, and structured logging.

jeremylongshore/tons-of-skills-marketplace2.8k—~1kAutomated safety check: PassMITyesterday
95

Cost guardrail for AWS DevOps Agent that covers ALL AWS services and native agent tools.

aws/tools-for-devops-agent103—~4.5kAutomated safety check: PassApache-2.0yesterday
96

Generate a 24-hour operational health report for a PostHog service by querying Grafana dashboards and Prometheus metrics.

haacked/dotfiles134—~8kAutomated safety check: NotesNo licenceyesterday