Search
Kubernetes · Monitoring and alerting
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Deploys KubeEye on KubeSphere and writes InspectRule and InspectPlan resources to inspect cluster health, then retrieves the inspection reports. | kubesphere/ | 17k | — | ~3.6k | Automated safety check: Pass | Unknown | 2 mo ago |
| 2 | Installs and configures the WizTelemetry Platform Service extension for KubeSphere, the shared API server behind its observability extensions. | kubesphere/ | 17k | — | ~1.8k | Automated safety check: Pass | Unknown | 2 mo ago |
| 3 | Reviews code for correctness and potential bugs, pinpoints bug locations by file and line, and suggests concrete fixes. | aide-family/ | 253 | — | ~815 | Automated safety check: Pass | No licence | 3 mo ago |
| 4 | Develops Go microservices with Kratos v2 following official design philosophy, DDD/Clean Architecture layout, Protobuf API, error/config/middleware patterns, and observability. | aide-family/ | 253 | — | ~1.5k | Automated safety check: Pass | No licence | 3 mo ago |
| 5 | 5.Signoz Manage the self-hosted SigNoz observability stack in this GitOps repo. | qjoly/ | 112 | — | ~6.1k | Automated safety check: Pass | WTFPL | yesterday |
| 6 | Implements backend modules from proto definitions for goddess, marksman, and rabbit apps. | aide-family/ | 253 | — | ~4.1k | Automated safety check: Pass | No licence | 3 mo ago |
| 7 | Set up metrics collection and visualization with Prometheus and Grafana. | BagelHole/ | 1.2k | — | ~2.5k | Automated safety check: Pass | MIT | 4 mo ago |
| 8 | Installs and configures the WizTelemetry Events extension for KubeSphere, which exports Kubernetes events for storage, with dependency checks and the event query API. | kubesphere/ | 17k | — | ~1.8k | Automated safety check: Pass | Unknown | 2 mo ago |
| 9 | Installs and configures WizTelemetry Logging for KubeSphere, with container log and optional disk log collection, dependency checks and the log query API. | kubesphere/ | 17k | — | ~2.3k | Automated safety check: Pass | Unknown | 2 mo ago |
| 10 | Improve and validate the Kaniop Grafana dashboard against repository metrics and the grigri live cluster. | pando85/ | 132 | — | ~987 | Automated safety check: Pass | AGPL-3.0 | today |
| 11 | Installs clidash, a dependency-free web dashboard that turns the JSON resource listings of a CLI such as NanoClaw's ncl into read-only tabs and tables. | nanocoai/ | 31k | — | ~1.6k | Automated safety check: Pass | MIT | yesterday |
| 12 | Installs and configures the WizTelemetry Notification extension for KubeSphere: channel setup, alert routing by tenant labels, silences and troubleshooting. | kubesphere/ | 17k | — | ~6.1k | Automated safety check: Pass | Unknown | 2 mo ago |
| 13 | Installs the WizTelemetry Ruler extension for KubeSphere and manages event, audit and log alerting rules as RuleGroup and ClusterRuleGroup resources. | kubesphere/ | 17k | — | ~6.5k | Automated safety check: Pass | Unknown | 2 mo ago |
| 14 | A skill your agent uses when the user asks to "write a validator", "add validation", "implement admission control", "write a mutating webhook", "add a mutation handler", "validate incoming… | grafana/ | 282 | — | ~1.7k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 15 | 15.Cloud Devops Cloud infrastructure and DevOps workflow covering AWS, Azure, GCP, Kubernetes, Terraform, CI/CD, monitoring, and cloud-native development. | davila7/ | 33k | 4 repos | ~1.4k | Automated safety check: Pass | MIT | today |
| 16 | Build a unified telemetry pipeline with Grafana Alloy — one OpenTelemetry-compatible binary that collects metrics, logs, traces, and profiles and ships to Grafana Cloud / Prometheus / Loki / Tempo /… | grafana/ | 282 | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 17 | Guides Qdrant monitoring setup including Prometheus scraping, health probes, Hybrid Cloud metrics, alerting, and log centralization. | qdrant/ | 254 | 2 repos | ~874 | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 18 | 18.Dd Apm APM - install, onboard, instrument, enable, set up, configure, traces, services, dependencies, performance analysis, Data Streams Monitoring (DSM), queue lag, pipeline latency. | datadog-labs/ | 177 | — | ~2k | Automated safety check: Pass | MIT | 2 days ago |
| 19 | Install the Datadog Agent on Kubernetes using the Datadog Operator — required before enabling Single Step Instrumentation (SSI), which automatically instruments applications for APM without code… | datadog-labs/ | 177 | — | ~2.1k | Automated safety check: Warn | MIT | 2 days ago |
| 20 | Implements eBPF-based runtime observability and in-kernel enforcement in Kubernetes with Cilium Tetragon, monitoring process execution, file access, network connections, and syscalls, and blocking… | mukul975/ | 34k | — | ~2.1k | Automated safety check: Notes | Apache-2.0 | 1 mo ago |
| 21 | Monitors and troubleshoots GKE TPU workloads, nodes, and node pools using GKE system metrics and PromQL. | google/ | 21k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 22 | 22.Loki Logging Configure Grafana Loki for log aggregation and analysis. An agent skill from sickn33/agentic-awesome-skills. | sickn33/ | 47k | 1 repo | ~2.6k | Automated safety check: Pass | MIT | 2 days ago |
| 23 | Set up metrics collection and visualization with Prometheus and Grafana. | sickn33/ | 47k | 1 repo | ~2.7k | Automated safety check: Pass | MIT | 2 days ago |
| 24 | Triage a degraded or suspect service end to end: read SLO status and burn rate, check active alerting rules and ML anomalies, measure throughput, latency, and error rate, assess dependency health… | elastic/ | 592 | — | ~7.4k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 25 | 当目标涉及云资产(对象存储/云元数据/Serverless)、容器/K8s、运维面板(宝塔/Grafana/Zabbix/Jenkins/GitLab/Nacos等)、消息队列/缓存中间件、CI/CD流水线、第三方回调集成、依赖组件CVE、信息泄露配置时调用。负责未授权访问、弱口令、云配置错误、供应链漏洞与敏感信息挖掘。 | zhaji2333/ | 115 | — | ~688 | Automated safety check: Warn | MIT | 26 days ago |
| 26 | Configures GKE observability, including Cloud Logging, Cloud Monitoring, and managed Prometheus. | google/ | 21k | — | ~4.4k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 27 | Auto-instrument an application's HTTP / gRPC / DB traffic with Grafana Beyla eBPF — no code changes, no SDK, no restart. | grafana/ | 282 | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 28 | Manage a fleet of Grafana Alloy collectors with Fleet Management — author Alloy pipelines once, target them via attribute matchers (env="production", regex region=~"us-."), push remotely via OpAMP… | grafana/ | 282 | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 29 | Ship Kubernetes, host, container, and cloud-provider telemetry into Grafana Cloud — k8s-monitoring Helm chart for K8s clusters (metrics + logs + traces + events + cost), Alloy… | grafana/ | 282 | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 30 | Stand up Grafana Mimir for horizontally scalable, multi-tenant, long-term Prometheus + OTLP metrics storage. | grafana/ | 282 | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 31 | Instrument any app with OpenTelemetry and ship metrics / logs / traces to Grafana Cloud or self-hosted Mimir / Loki / Tempo / Pyroscope. | grafana/ | 282 | — | ~1.7k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 32 | Monitor use when deploying monitoring stacks including Prometheus, Grafana, and Datadog. | jeremylongshore/ | 2.8k | — | ~1.2k | Automated safety check: Pass | MIT | yesterday |
| 33 | Query Coralogix's Service Catalog (APM v2 entities) with the cx service-catalog CLI — discover entity types, list known entities, check their schema, and pull aggregated or timeseries data for… | coralogix/ | 121 | — | ~2.1k | Automated safety check: Pass | Apache-2.0 | 4 days ago |
| 34 | 34.Loki Logging Configure Grafana Loki for log aggregation and analysis. An agent skill from BagelHole/DevOps-Security-Agent-Skills. | BagelHole/ | 1.2k | — | ~2.4k | Automated safety check: Pass | MIT | 4 mo ago |
| 35 | Generate a live Single Step Instrumentation (SSI) onboarding confirmation report — verifies APM instrumentation is working end-to-end with deep links into the Datadog UI. | datadog-labs/ | 177 | — | ~1.1k | Automated safety check: Pass | MIT | 2 days ago |
| 36 | Observability patterns for logging, monitoring, alerting, and distributed tracing. | TheSoftwareHouse/ | 284 | — | ~2k | Automated safety check: Pass | MIT | 6 days ago |
| 37 | Set up GPU monitoring and observability for CoreWeave workloads. | jeremylongshore/ | 2.8k | — | ~1.4k | Automated safety check: Pass | MIT | yesterday |
| 38 | 38.Enable Ssi Enable Single Step Instrumentation (SSI) on Kubernetes — automatically instruments applications for APM without code changes. | datadog-labs/ | 177 | — | ~3.2k | Automated safety check: Warn | MIT | 2 days ago |
| 39 | 39.Grafana Operate, configure, provision, secure, and troubleshoot Grafana OSS, Enterprise, and Cloud, including dashboards, folders, data sources, annotations, alert rules, contact points, notification… | magnus919/ | 115 | — | ~2.4k | Automated safety check: Pass | MIT | yesterday |
| 40 | Query Prometheus monitoring metrics and alert rules. An agent skill from Kilo-Org/kilo-marketplace. | Kilo-Org/ | 190 | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | 12 days ago |