Search
Prometheus · For developers
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 49 | Monitors and troubleshoots GKE TPU workloads, nodes, and node pools using GKE system metrics and PromQL. | google/ | 21k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 50 | 50.Nav Init Initialize Navigator documentation structure in a project. An agent skill from qf-studio/navigator. | qf-studio/ | 355 | — | ~3k | Automated safety check: Notes | MIT | 2 days ago |
| 51 | Monitoring, logging, and tracing implementation using OpenTelemetry as the unified standard. | ancoleman/ | 525 | — | ~3k | Automated safety check: Pass | MIT | 10 mo ago |
| 52 | Generates Cloud Monitoring Server-Driven UI (SDUI) Widget and XyChart Protocol Buffer textprotos on Google Cloud from resolved PromQL or ListTimeSeries queries. | google/ | 21k | — | ~2.7k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 53 | Configures alerting policies in Terraform for Google Kubernetes Engine (GKE) clusters, workloads, and services using PromQL and Google Cloud Managed Service for Prometheus. | google/ | 21k | — | ~5.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 54 | Auto-scale LLM inference clusters on Kubernetes using KEDA, custom GPU metrics, and horizontal pod autoscaling. | sickn33/ | 47k | 1 repo | ~2.1k | Automated safety check: Pass | MIT | 2 days ago |
| 55 | Set up metrics collection and visualization with Prometheus and Grafana. | sickn33/ | 47k | 1 repo | ~2.7k | Automated safety check: Pass | MIT | 2 days ago |
| 56 | Configures Cloud Monitoring PromQL-based Service Level Objective (SLO) alerting policies on Google Cloud for resources registered in App Hub or individually specified. | google/ | 21k | — | ~3.1k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 57 | Guides Qdrant monitoring and observability setup. An agent skill from github/awesome-copilot. | github/ | 40k | 1 repo | ~276 | Automated safety check: Pass | MIT | 2 days ago |
| 58 | Analyze the experiment precompute result-consistency canary across prod-US and prod-EU, deep-dive any issues, and produce an actionable report. | PostHog/ | 40k | — | ~3.5k | Automated safety check: Pass | Unknown | yesterday |
| 59 | Investigates server/infrastructure metric anomalies in PostHog Metrics — from "this metric is rising/dropping/spiking" or a fired alert to a probable cause with evidence. | PostHog/ | 40k | — | ~1.5k | Automated safety check: Pass | Unknown | yesterday |
| 60 | Manages scaling for GKE workloads using HPA and VPA. An agent skill from google/skills. | google/ | 21k | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 61 | Diagnoses, predicts, and mitigates node disruptions during Compute Engine host maintenance and hardware or software maintenance events for GPU and TPU workloads on GKE. | google/ | 21k | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 62 | Configures GKE observability, including Cloud Logging, Cloud Monitoring, and managed Prometheus. | google/ | 21k | — | ~4.4k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 63 | Kubernetes deployment workflow for container orchestration, Helm charts, service mesh, and production-ready K8s configurations. | aiskillstore/ | 433 | 5 repos | ~839 | Automated safety check: Pass | No licence | yesterday |
| 64 | Auto-instrument an application's HTTP / gRPC / DB traffic with Grafana Beyla eBPF — no code changes, no SDK, no restart. | grafana/ | 282 | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 65 | Manage a fleet of Grafana Alloy collectors with Fleet Management — author Alloy pipelines once, target them via attribute matchers (env="production", regex region=~"us-."), push remotely via OpAMP… | grafana/ | 282 | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 66 | Configure Grafana OSS — provisions dashboards from YAML, sets up data sources (Prometheus / Loki / Tempo / Pyroscope), writes dashboard JSON with template variables, builds panel queries, assigns… | grafana/ | 282 | — | ~1.5k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 67 | Stand up Grafana Mimir for horizontally scalable, multi-tenant, long-term Prometheus + OTLP metrics storage. | grafana/ | 282 | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 68 | Turn on AI + ML features in Grafana Cloud — Grafana Assistant (NL → PromQL/LogQL/TraceQL, dashboard build, incident investigation, MCP integration), Dynamic Alerting (Prophet forecasting + DBSCAN… | grafana/ | 282 | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 69 | Canvas/A2UI inline network visualizations — topology maps, health dashboards, alert cards, change timelines, config diffs, path traces, and health scorecards rendered directly in the OpenClaw chat… | automateyournetwork/ | 676 | — | ~941 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 70 | Reusable investigation patterns for AWS CloudWatch: Logs Insights query templates, alarm-to-deployment correlation, blast-radius narrowing decision tree, and PromQL-style metric query patterns for… | github/ | 40k | — | ~2.6k | Automated safety check: Pass | MIT | 2 days ago |
| 71 | Monitor use when deploying monitoring stacks including Prometheus, Grafana, and Datadog. | jeremylongshore/ | 2.8k | — | ~1.2k | Automated safety check: Pass | MIT | yesterday |
| 72 | Sending telemetry data to Grafana Cloud — metrics via Prometheus remote write or OTLP, logs via Loki push or Alloy, traces via OTLP to Tempo, profiles via Pyroscope. | grafana/ | 282 | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 73 | A skill your agent uses to deploy and operate a CollectX (clx) based DOCA telemetry collector on a host or BlueField — wiring providers / counters into the collector, running the collection daemon… | NVIDIA/ | 3.6k | — | ~2.8k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 74 | Multi-agent orchestration plugin for OpenCode. An agent skill from LeoYeAI/openclaw-master-skills. | LeoYeAI/ | 2.2k | — | ~5.1k | Automated safety check: Notes | MIT | 2 mo ago |
| 75 | Design, audit, and troubleshoot production monitoring and observability using user-impact checks, layered telemetry, USE/RED, SLI/SLO/SLA, error budgets, cardinality controls, actionable alerting… | AnastasiyaW/ | 154 | — | ~4.1k | Automated safety check: Pass | MIT | yesterday |
| 76 | Builds, configures, debugs, and optimizes AWS observability - operator-symptom questions and detecting Omni vs classic CloudWatch. | aws/ | 2.8k | — | ~7.2k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 77 | Probe a target for accidentally-public admin / debug / introspection endpoints — Spring Boot Actuator, Apache server-status, Prometheus metrics, GraphQL playground, Swagger UI, phpMyAdmin… | jeremylongshore/ | 2.8k | — | ~2k | Automated safety check: Pass | MIT | yesterday |
| 78 | A skill your agent uses when you need to implement or improve Java metrics observability with Micrometer — including meter design, naming/tag conventions, cardinality control… | jabrena/ | 447 | — | ~868 | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 79 | Observability patterns for Python applications. An agent skill from aiskillstore/marketplace. | aiskillstore/ | 433 | 1 repo | ~1.3k | Automated safety check: Pass | No licence | yesterday |
| 80 | This skill provides AWS cost optimization, monitoring, and operational best practices with integrated MCP servers for billing analysis, cost estimation, observability, and security assessment. | Microck/ | 404 | 1 repo | ~2.5k | Automated safety check: Pass | Unknown | 1 mo ago |
| 81 | Monitoring and observability patterns for Prometheus metrics, Grafana dashboards, Langfuse v4 LLM tracing (astype, scorecurrentspan, shouldexportspan, LangfuseMedia), and drift detection. | yonatangross/ | 292 | — | ~2.2k | Automated safety check: Pass | MIT | yesterday |
| 82 | Golang benchmarking, profiling, and performance measurement. | aiskillstore/ | 433 | 1 repo | ~3.4k | Automated safety check: Pass | MIT | yesterday |
| 83 | 83.Gorm Expert GORM v2 最佳实践与性能优化。适用于:代码审查、慢查询优化、N+1、连接池、 事务管理、分库分表、Prometheus/OTel监控、Session安全、Clause/Upsert、 缓存集成、BaseModel脚手架、SQL→struct生成、多租户隔离。 | LeoYeAI/ | 2.2k | — | ~3.4k | Automated safety check: Pass | MIT | 2 mo ago |
| 84 | OpenTelemetry, distributed tracing, structured logging, metrics (Prometheus, Grafana, Datadog). | TheBeardedBearSAS/ | 107 | — | ~547 | Automated safety check: Pass | MIT | 27 days ago |
| 85 | Monitoring and observability with OpenTelemetry, Prometheus, Grafana dashboards, and structured logging | rohitg00/ | 2.7k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | 5 mo ago |
| 86 | Observability: structured logs, metrics (RED/USE), tracing, SLO/SLI. | softspark/ | 179 | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 87 | Set up Apollo.io monitoring and observability. An agent skill from jeremylongshore/tons-of-skills-marketplace. | jeremylongshore/ | 2.8k | — | ~2.3k | Automated safety check: Pass | MIT | yesterday |
| 88 | Monitor ClickHouse with Prometheus metrics, Grafana dashboards, system table queries, and alerting for query performance, merge health, and resource usage. | jeremylongshore/ | 2.8k | — | ~1.4k | Automated safety check: Pass | MIT | yesterday |
| 89 | Set up comprehensive observability for Deepgram integrations. | jeremylongshore/ | 2.8k | — | ~3k | Automated safety check: Pass | MIT | yesterday |
| 90 | Execute Deepgram production deployment checklist. An agent skill from jeremylongshore/tons-of-skills-marketplace. | jeremylongshore/ | 2.8k | — | ~2.3k | Automated safety check: Pass | MIT | yesterday |
| 91 | A skill your agent uses when you need production monitoring for an Intercom integration — instrumenting API calls with metrics and traces, standing up dashboards, or wiring alerts for error rate… | jeremylongshore/ | 2.8k | — | ~1.7k | Automated safety check: Pass | MIT | yesterday |
| 92 | Set up observability for Klaviyo integrations with metrics, traces, and alerts. | jeremylongshore/ | 2.8k | — | ~1.5k | Automated safety check: Pass | MIT | yesterday |
| 93 | Set up comprehensive observability for Langfuse with metrics, dashboards, and alerts. | jeremylongshore/ | 2.8k | — | ~2.2k | Automated safety check: Pass | MIT | yesterday |
| 94 | Set up observability for Shopify app integrations with query cost tracking, rate limit monitoring, webhook delivery metrics, and structured logging. | jeremylongshore/ | 2.8k | — | ~1k | Automated safety check: Pass | MIT | yesterday |
| 95 | Cost guardrail for AWS DevOps Agent that covers ALL AWS services and native agent tools. | aws/ | 103 | — | ~4.5k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 96 | 96.Ops Report Generate a 24-hour operational health report for a PostHog service by querying Grafana dashboards and Prometheus metrics. | haacked/ | 134 | — | ~8k | Automated safety check: Notes | No licence | yesterday |