Developer tool
Prometheus agent skills, page 3
Prometheus skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 97 | Build and deploy a Coralogix dashboard for a given service from its logs, spans, metrics, and service specs. | coralogix/ | 121 | — | ~4.7k | Automated safety check: Warn | Apache-2.0 | 2 days ago |
| 98 | 98.Prometheus Prometheus monitoring expert for PromQL, alerting rules, Grafana dashboards, and observability | RightNow-AI/ | 18k | — | ~738 | Automated safety check: Pass | Apache-2.0 | 3 mo ago |
| 99 | Go observability — always-on production signals: slog logging, Prometheus metrics, OpenTelemetry tracing, pprof profiling, alerting, Grafana. | context-labs/ | 1.1k | 1 repo | ~3.3k | Automated safety check: Pass | MIT | 2 days ago |
| 100 | Read what FastLLM has been doing — usage records, time-series aggregates, the configuration audit trail, Prometheus metrics, control-plane health, and per-replica fleet status. | azrtydxb/ | 108 | — | ~916 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 101 | 101.Analyze Runtime Find out what a running Mendix app actually does — logs, Prometheus metrics, OpenTelemetry traces and the model catalog, joined across sources. | mendixlabs/ | 128 | — | ~2.8k | Automated safety check: Pass | Apache-2.0 | today |
| 102 | 当需要为 funboost 任务添加监控、链路追踪或告警时使用。触发场景:Prometheus 指标、OpenTelemetry 链路追踪、异常告警通知、周期额度限制、函数结果持久化。关键词:Prometheus, OpenTelemetry, OTel, 告警, 监控, metrics, tracing, AlertNotifier, PeriodicQuota。 | ydf0509/ | 892 | — | ~4.9k | Automated safety check: Pass | No licence | 1 mo ago |
| 103 | Set up and manage NVIDIA GPU servers for AI workloads. An agent skill from sickn33/agentic-awesome-skills. | sickn33/ | 47k | 2 repos | ~2k | Automated safety check: Notes | MIT | yesterday |
| 104 | 104.Alerting Oncall Set up alerting rules, configure on-call rotations, and manage incident response workflows. | sickn33/ | 47k | 1 repo | ~2.8k | Automated safety check: Pass | MIT | yesterday |
| 105 | 105.Nav Init Initialize Navigator documentation structure in a project. An agent skill from qf-studio/navigator. | qf-studio/ | 354 | — | ~3k | Automated safety check: Notes | MIT | yesterday |
| 106 | Monitoring, logging, and tracing implementation using OpenTelemetry as the unified standard. | ancoleman/ | 526 | — | ~3k | Automated safety check: Pass | MIT | 10 mo ago |
| 107 | Auto-scale LLM inference clusters on Kubernetes using KEDA, custom GPU metrics, and horizontal pod autoscaling. | sickn33/ | 47k | 1 repo | ~2.1k | Automated safety check: Pass | MIT | yesterday |
| 108 | Set up metrics collection and visualization with Prometheus and Grafana. | sickn33/ | 47k | 1 repo | ~2.7k | Automated safety check: Pass | MIT | yesterday |
| 109 | 109.Cost Export Export cost-tracking telemetry in Prometheus textfile or webhook JSON formats — for external observability (Grafana, Datadog, custom dashboards) | ruvnet/ | 74k | — | ~687 | Automated safety check: Notes | MIT | today |
| 110 | Kubernetes deployment workflow for container orchestration, Helm charts, service mesh, and production-ready K8s configurations. | aiskillstore/ | 430 | 5 repos | ~839 | Automated safety check: Pass | No licence | today |
| 111 | 111.Promql Validator Validate, lint, audit, or fix PromQL queries and alerting rules; detects anti-patterns. | akin-ozer/ | 319 | — | ~4k | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 112 | Canvas/A2UI inline network visualizations — topology maps, health dashboards, alert cards, change timelines, config diffs, path traces, and health scorecards rendered directly in the OpenClaw chat… | automateyournetwork/ | 674 | — | ~941 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 113 | 113.Oh My Opencode Multi-agent orchestration plugin for OpenCode. An agent skill from LeoYeAI/openclaw-master-skills. | LeoYeAI/ | 2.2k | — | ~5.1k | Automated safety check: Notes | MIT | 2 mo ago |
| 114 | Design, audit, and troubleshoot production monitoring and observability using user-impact checks, layered telemetry, USE/RED, SLI/SLO/SLA, error budgets, cardinality controls, actionable alerting… | AnastasiyaW/ | 154 | — | ~4.1k | Automated safety check: Pass | MIT | 4 days ago |
| 115 | A skill your agent uses when you need to implement or improve Java metrics observability with Micrometer — including meter design, naming/tag conventions, cardinality control… | jabrena/ | 445 | — | ~868 | Automated safety check: Pass | Apache-2.0 | today |
| 116 | [OMX] Clean-room interview-driven planner: Metis clarifies, Momus challenges, Oracle synthesizes, then hands off to $ultragoal/$team. | yangyuan-zhen/ | 315 | — | ~4.6k | Automated safety check: Pass | AGPL-3.0 | 16 days ago |
| 117 | Observability and SRE expert. An agent skill from majiayu000/spellbook. | majiayu000/ | 286 | — | ~3.3k | Automated safety check: Pass | MIT | today |
| 118 | 118.Promql Generator Generate/create/write PromQL queries, metric expressions, alerting rules, recording rules, Prometheus dashboards. | akin-ozer/ | 319 | — | ~9.2k | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 119 | Observability patterns for Python applications. An agent skill from aiskillstore/marketplace. | aiskillstore/ | 430 | 1 repo | ~1.3k | Automated safety check: Pass | No licence | today |
| 120 | Build AI-focused SRE incident response practices for LLM outages, degraded quality, runaway cost events, and safety regressions. | majiayu000/ | 666 | 3 repos | ~2.8k | Automated safety check: Pass | MIT | today |
| 121 | This skill provides AWS cost optimization, monitoring, and operational best practices with integrated MCP servers for billing analysis, cost estimation, observability, and security assessment. | Microck/ | 401 | 1 repo | ~2.5k | Automated safety check: Pass | Unknown | 1 mo ago |
| 122 | 122.Datapages Server Configure the Datapages server entry point: NewServer type arguments, the message broker, server options, static assets, TLS and Prometheus metrics. | romshark/ | 113 | — | ~2.2k | Automated safety check: Pass | MIT | 2 days ago |
| 123 | Monitoring and observability patterns for Prometheus metrics, Grafana dashboards, Langfuse v4 LLM tracing (astype, scorecurrentspan, shouldexportspan, LangfuseMedia), and drift detection. | yonatangross/ | 288 | — | ~2.2k | Automated safety check: Pass | MIT | yesterday |
| 124 | 124.Golang Benchmark Golang benchmarking, profiling, and performance measurement. | aiskillstore/ | 430 | 1 repo | ~3.4k | Automated safety check: Pass | MIT | today |
| 125 | 125.Gorm Expert GORM v2 最佳实践与性能优化。适用于:代码审查、慢查询优化、N+1、连接池、 事务管理、分库分表、Prometheus/OTel监控、Session安全、Clause/Upsert、 缓存集成、BaseModel脚手架、SQL→struct生成、多租户隔离。 | LeoYeAI/ | 2.2k | — | ~3.4k | Automated safety check: Pass | MIT | 2 mo ago |
| 126 | 126.Alerting Oncall Set up alerting rules, configure on-call rotations, and manage incident response workflows. | BagelHole/ | 1.1k | — | ~3k | Automated safety check: Pass | MIT | 4 mo ago |
| 127 | 127.Observability OpenTelemetry, distributed tracing, structured logging, metrics (Prometheus, Grafana, Datadog). | TheBeardedBearSAS/ | 107 | — | ~547 | Automated safety check: Pass | MIT | 23 days ago |
| 128 | Monitoring and observability with OpenTelemetry, Prometheus, Grafana dashboards, and structured logging | rohitg00/ | 2.7k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | 4 mo ago |
| 129 | Observability: structured logs, metrics (RED/USE), tracing, SLO/SLI. | softspark/ | 179 | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | today |
| 130 | Query and analyze Claude Code observability data (metrics, logs, traces). | majiayu000/ | 666 | 1 repo | ~1.6k | Automated safety check: Pass | MIT | today |
| 131 | 131.Grafana Helper Use Grafana's GCX CLI for dashboards, datasources, Prometheus metrics, Loki logs, Tempo traces, alert rules, and Grafana resource operations. | shepherdjerred/ | 112 | — | ~1.1k | Automated safety check: Pass | GPL-3.0 | today |
| 132 | 132.Ops Report Generate a 24-hour operational health report for a PostHog service by querying Grafana dashboards and Prometheus metrics. | haacked/ | 134 | — | ~8k | Automated safety check: Notes | No licence | today |
| 133 | 133.Telemetry Operate the observability stack that deploys as one unit: Prometheus scrape configuration, recording and alerting rules, relabeling, retention, and high availability; OpenTelemetry Collector… | magnus919/ | 111 | — | ~3.9k | Automated safety check: Pass | MIT | today |
| 134 | Auto-scale LLM inference clusters on Kubernetes using KEDA, custom GPU metrics, and horizontal pod autoscaling. | BagelHole/ | 1.1k | — | ~2k | Automated safety check: Pass | MIT | 4 mo ago |
| 135 | Observability patterns for logging, monitoring, alerting, and distributed tracing. | TheSoftwareHouse/ | 284 | — | ~2k | Automated safety check: Pass | MIT | 2 days ago |
| 136 | Grafana observability platform — dashboards, Prometheus PromQL, Loki LogQL, alerting, incidents, OnCall schedules, annotations, datasource queries, panel rendering (75+ tools). | automateyournetwork/ | 674 | — | ~2.7k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 137 | Prometheus monitoring — PromQL instant/range queries, metric discovery, metadata, scrape target health, system health checks (6 tools). | automateyournetwork/ | 674 | — | ~2k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 138 | 138.Promql CLI CLI for querying Prometheus and PromQL-compatible engines (Thanos, Cortex, VictoriaMetrics, Grafana Mimir, Grafana Tempo...) — instant queries, range queries, metric discovery (metrics/labels/meta… | samber/ | 228 | — | ~1.9k | Automated safety check: Pass | MIT | 6 days ago |
| 139 | 139.Cloud Monitoring Monitor cloud infrastructure and applications using metrics, logs, and traces to provide real-time observability into performance, health, and reliability. | seb1n/ | 206 | — | ~2.8k | Automated safety check: Pass | MIT | 1 mo ago |
| 140 | 140.Tencentcloud Cls 腾讯云日志服务 CLS 技能。支持 CQL 日志检索、上下文查看、日志主题/日志集查看、机器组与机器状态查看、采集规则查看、日志直方图、指标采集与 PromQL 查询、告警策略与告警历史分析。触发词:CLS、日志检索、日志查询、日志上下文、日志主题、日志集、索引配置、重建索引、机器组、机器状态、采集规则、采集配置、日志直方图、指标查询、PromQL、告警策略、告警历史、告警屏蔽、云日志、log… | infometa/ | 342 | — | ~3.1k | Automated safety check: Notes | No licence | today |
| 141 | 141.Grafana Operate, configure, provision, secure, and troubleshoot Grafana OSS, Enterprise, and Cloud, including dashboards, folders, data sources, annotations, alert rules, contact points, notification… | magnus919/ | 111 | — | ~2.4k | Automated safety check: Pass | MIT | today |
| 142 | Query Prometheus monitoring metrics and alert rules. An agent skill from Kilo-Org/kilo-marketplace. | Kilo-Org/ | 189 | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | 8 days ago |
| 143 | 143.Grafana Expert Grafana expert: dashboard design, panels, alerting, data sources. | theneoai/ | 183 | — | ~4.8k | Automated safety check: Pass | MIT | 4 mo ago |
| 144 | Prometheus expert: PromQL, exporters, alerting rules, recording rules. | theneoai/ | 183 | — | ~4.7k | Automated safety check: Pass | MIT | 4 mo ago |