Developer tool
Prometheus agent skills, page 2
Prometheus skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 49 | Generate Perses dashboards or single panels for GreptimeDB. An agent skill from GreptimeTeam/dashboard. | GreptimeTeam/ | 111 | — | ~3.9k | Automated safety check: Pass | Apache-2.0 | 7 days ago |
| 50 | Create the release notes and upgrade guide for a mariadb-operator release. | mariadb-operator/ | 1k | — | ~3.3k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 51 | 51.Code Review Reviews code for correctness and potential bugs, pinpoints bug locations by file and line, and suggests concrete fixes. | aide-family/ | 253 | — | ~815 | Automated safety check: Pass | No licence | 3 mo ago |
| 52 | 52.Opsany 通过 opsany-mcp-server 连接 OpsAny 运维平台,实现 CMDB 资源查询和操作、模型管理、工单管理、用户管理、作业执行及主机纳管等全栈运维操作。 | unixhot/ | 181 | — | ~2k | Automated safety check: Pass | Unknown | 2 mo ago |
| 53 | End-to-end docker-compose test harness for the minecraft-prometheus-exporter. | dirien/ | 142 | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 54 | Deploys LLMs with vLLM for high-throughput serving, covering the OpenAI-compatible server, offline batch inference, monitoring and a Docker rollout. | Orchestra-Research/ | 13k | 6 repos | ~2.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 55 | Visually verify Eneru browser-dashboard changes against a live daemon or audit an exact deployment. | m4r1k/ | 149 | — | ~1.4k | Automated safety check: Pass | MIT | 3 days ago |
| 56 | Installs, uninstalls, checks and troubleshoots the KubeSphere Gateway extension built on ingress-nginx, including gateways stuck in bad states and Helm or pod failures. | kubesphere/ | 17k | — | ~2.9k | Automated safety check: Pass | Unknown | 2 mo ago |
| 57 | Develops Go microservices with Kratos v2 following official design philosophy, DDD/Clean Architecture layout, Protobuf API, error/config/middleware patterns, and observability. | aide-family/ | 253 | — | ~1.5k | Automated safety check: Pass | No licence | 3 mo ago |
| 58 | 58.Signoz Manage the self-hosted SigNoz observability stack in this GitOps repo. | qjoly/ | 112 | — | ~6.1k | Automated safety check: Pass | WTFPL | today |
| 59 | A skill your agent uses when publishing prometheus-proxy to Maven Central, cutting a release, running a snapshot publish, or bumping the project version — covers the Maven Central coordinates, GPG… | pambrose/ | 157 | — | ~506 | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 60 | 对远程多实例MySQL数据库执行全方位深度巡检,覆盖基础健康、连接负载、性能慢查询、索引冗余、主从复制、容量空间、账号安全、配置风险八大维度,全自动完成巡检扫描、风险识别、问题定级、优化建议、报告归档与飞书推送,适用于生产/测试所有运行中MySQL实例常态化合规巡检。适用场景:用户要求进行 MySQL 全链路健康检查、MySQL 综合巡检、MySQL 风险扫描、MySQL 性能审计、MySQL… | openocta/ | 166 | — | ~2k | Automated safety check: Pass | MIT | 3 mo ago |
| 61 | Bootstrap, create, connect to, operate, secure, scale, upgrade, troubleshoot, inspect, and tear down Alibaba Cloud Container Compute Service (ACS) Agent Sandbox environments. | cinience/ | 397 | — | ~2.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 62 | 当需要为 funboost 创建 Consumer 或 Publisher 的 Mixin 扩展类时使用。触发场景:添加监控、熔断、限流、链路追踪等横切关注点,编写自定义前置/后置处理钩子。关键词:mixin, consumeroverridecls, publisheroverridecls, ConsumerMixin, 自定义消费者, hook, 拦截器, 熔断器, 监控… | ydf0509/ | 892 | — | ~2.1k | Automated safety check: Pass | No licence | 1 mo ago |
| 63 | 63.Oryxos Init 初始化 OryxOS(或同类 JDK 21 + Spring Boot 3.x 企业级单体)的工程地基:Maven 多模块骨架、 结构化日志、Actuator + Prometheus 监控、Spring MVC + 虚拟线程、springdoc OpenAPI、 统一响应体与全局异常/错误码、Google 格式 + 阿里编码规约(Spotless + 阿里 P3C +… | oryx-labs/ | 187 | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 64 | 64.Graft This repo is indexed by graft/. An agent skill from m4r1k/Eneru. | m4r1k/ | 149 | 1 repo | ~2.3k | Automated safety check: Pass | MIT | 3 days ago |
| 65 | Mandatory pre-release deep review for minor/major releases (X.Y.0 / X.0.0). | m4r1k/ | 149 | — | ~1.9k | Automated safety check: Pass | MIT | 3 days ago |
| 66 | MUST USE when investigating performance issues on a ClickHouse-managed Postgres instance. | ClickHouse/ | 544 | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | 8 days ago |
| 67 | Implements backend modules from proto definitions for goddess, marksman, and rabbit apps. | aide-family/ | 253 | — | ~4.1k | Automated safety check: Pass | No licence | 3 mo ago |
| 68 | Builds a picture of whether Prometheus itself is healthy and successfully monitoring its targets, covering readiness, firing alerts, target health and TSDB load. | prometheus/ | 117 | — | ~584 | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 69 | A skill your agent uses when changing Archestra tracing, metrics, OpenTelemetry, Tempo, Grafana, Prometheus, LLM/MCP spans, observability labels, or local observability setup. | archestra-ai/ | 4.3k | — | ~1.2k | Automated safety check: Pass | Unknown | today |
| 70 | 70.Helm Chart A skill your agent uses for Helm chart work - creating charts, modifying existing charts, values design, testing. | astronomer/ | 491 | — | ~6.5k | Automated safety check: Pass | Unknown | today |
| 71 | A skill your agent uses when the user asks about AI Center Coding Agents data, wants to reproduce or extend the Coding Agents dashboards, or asks questions about usage, cost, tokens, sessions… | coralogix/ | 121 | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 72 | 72.Docs Site A skill your agent uses when editing, building, previewing, or deploying the prometheus-proxy documentation site under website/prometheus-proxy — covers the Zensical config, code-snippet resolution… | pambrose/ | 157 | — | ~250 | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 73 | A skill your agent uses when adding tracing, structured logging, metrics, or monitoring to a FrontMCP server. | agentfront/ | 146 | — | ~4.6k | Automated safety check: Pass | Apache-2.0 | today |
| 74 | Diagnose a live, running lean devnet from its Prometheus and container logs, then report findings with proposed fixes and stop for approval. | geanlabs/ | 177 | — | ~3.8k | Automated safety check: Warn | No licence | yesterday |
| 75 | Monitoring and observability strategy, implementation, and troubleshooting. | ahmedasmar/ | 203 | — | ~3.9k | Automated safety check: Pass | No licence | 5 mo ago |
| 76 | Set up metrics collection and visualization with Prometheus and Grafana. | BagelHole/ | 1.1k | — | ~2.5k | Automated safety check: Pass | MIT | 4 mo ago |
| 77 | A skill your agent uses when the user asks to "check data usage", "list TCO policies", "reduce Coralogix costs", "optimize observability spend", "lower our logging bill", "data budget exceeded"… | coralogix/ | 121 | — | ~3.4k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 78 | Quantifies elevated error rates with PromQL, compares them to a baseline, and isolates which jobs or instances an error spike is concentrated in. | prometheus/ | 117 | — | ~592 | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 79 | In-memory caching in Golang using samber/hot — eviction algorithms (LRU, LFU, TinyLFU, W-TinyLFU, S3FIFO, ARC, TwoQueue, SIEVE, FIFO), TTL, cache loaders, sharding, stale-while-revalidate, missing… | samber/ | 3.4k | — | ~2k | Automated safety check: Pass | MIT | 6 days ago |
| 80 | Complete guide to Prometheus setup, metric collection, scrape configuration, and recording rules. | davila7/ | 32k | 11 repos | ~2.6k | Automated safety check: Pass | MIT | today |
| 81 | Finds the metrics and labels behind a Prometheus series explosion using a connected Prometheus MCP server, then proposes relabeling, dropping or recording-rule fixes with measured impact. | prometheus/ | 117 | — | ~547 | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 82 | LiteLLM-RS Observability Architecture. An agent skill from majiayu000/litellm-rs. | majiayu000/ | 116 | — | ~1.3k | Automated safety check: Pass | MIT | today |
| 83 | Author, modify, or review Netdata collectors across Go, IBM, C, Rust and external plugins. | netdata/ | 81k | — | ~1.9k | Automated safety check: Pass | GPL-3.0 | today |
| 84 | DevOps 工程师 Agent — CI/CD 流水线、容器化与 K8s、基础设施即代码、可观测性. An agent skill from peterfei/ai-agent-team. | peterfei/ | 441 | — | ~1k | Automated safety check: Pass | MIT | 3 mo ago |
| 85 | 85.Expert Ops 基础设施运维专家入口。用于 Codex CLI 的 $expert-ops 调用. An agent skill from ReJeCtAll/ExpertTeam-Codex. | ReJeCtAll/ | 113 | — | ~625 | Automated safety check: Pass | MIT | 3 mo ago |
| 86 | Improve and validate the Kaniop Grafana dashboard against repository metrics and the grigri live cluster. | pando85/ | 130 | — | ~987 | Automated safety check: Pass | AGPL-3.0 | yesterday |
| 87 | A skill your agent uses when the user asks to "set up parsing", "create parsing rule", "extract fields from logs", "regex extraction", "log parsing", "enrich logs", "add context to logs", "custom… | coralogix/ | 121 | — | ~3k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 88 | Sets up application monitoring: structured logs, Prometheus metrics, OpenTelemetry tracing, Grafana dashboards, alert rules and load tests with k6 or Artillery. | Jeffallan/ | 12k | — | ~1.6k | Automated safety check: Pass | MIT | 4 days ago |
| 89 | 89.SRE Engineer Defines SLIs, SLOs and error budgets, and sets up golden-signal monitoring, blameless postmortems, toil automation and chaos experiments for production systems. | Jeffallan/ | 12k | — | ~1.7k | Automated safety check: Pass | MIT | 4 days ago |
| 90 | Audits Prometheus recording and alerting rules through a connected Prometheus MCP server, finds gaps and noisy alerts, and drafts improved rule-group YAML. | prometheus/ | 117 | — | ~765 | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 91 | Installs and configures the WizTelemetry Notification extension for KubeSphere: channel setup, alert routing by tenant labels, silences and troubleshooting. | kubesphere/ | 17k | — | ~6.1k | Automated safety check: Pass | Unknown | 2 mo ago |
| 92 | 92.Live Debug Debug the running local stack with traces, logs, and a shared headless browser. | macro-inc/ | 4.6k | — | ~2.4k | Automated safety check: Notes | AGPL-3.0 | today |
| 93 | A skill your agent uses for any question involving telemetry data: "investigate an issue", "debug a problem", "find out why something is slow", "check error rates", "analyze user behavior"… | coralogix/ | 121 | — | ~2.6k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 94 | Finds where a Prometheus metric stops existing, whether at the target, the scrape, relabeling or the query, using the tools of a connected Prometheus MCP server. | prometheus/ | 117 | — | ~587 | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 95 | 检查 Prometheus 数据源的连通性、数据延迟和指标采集健康度。 | kubehan/ | 125 | — | ~229 | Automated safety check: Pass | No licence | 19 days ago |
| 96 | Review and tune Prometheus configuration and performance. An agent skill from prometheus/prometheus-mcp. | prometheus/ | 117 | — | ~724 | Automated safety check: Pass | Apache-2.0 | 3 days ago |