Search
OpenTelemetry · Monitoring and alerting
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Explores and queries OpenTelemetry metrics in Axiom MetricsDB, listing datasets, metrics and tags first and picking the right aggregation for each metric's type. | openclaw/ | 9.5k | — | ~2.6k | Automated safety check: Pass | MIT | yesterday |
| 2 | Specifies how to instrument an opik-backend pipeline with per-stage OpenTelemetry metrics for throughput, latency, errors and queue delay by workspace. | comet-ml/ | 22k | — | ~3.2k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 3 | Writes Terraform alerting policies for AI agents that emit OpenTelemetry metrics, covering reliability, cost, safety, security and quality signals on Google Cloud. | google/ | 21k | — | ~4.2k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 4 | 4.Signoz Manage the self-hosted SigNoz observability stack in this GitOps repo. | qjoly/ | 112 | — | ~6.1k | Automated safety check: Pass | WTFPL | yesterday |
| 5 | Investigates distributed application performance using PostHog APM (OpenTelemetry span) data via MCP. | PostHog/ | 40k | — | ~3.5k | Automated safety check: Pass | Unknown | today |
| 6 | Adds Pydantic Logfire tracing, logging and metrics to Python, JavaScript or TypeScript and Rust projects, with the correct setup order and library extras. | basicmachines-co/ | 4.1k | — | ~2.3k | Automated safety check: Pass | AGPL-3.0 | today |
| 7 | 当需要为 funboost 创建 Consumer 或 Publisher 的 Mixin 扩展类时使用。触发场景:添加监控、熔断、限流、链路追踪等横切关注点,编写自定义前置/后置处理钩子。关键词:mixin, consumeroverridecls, publisheroverridecls, ConsumerMixin, 自定义消费者, hook, 拦截器, 熔断器, 监控… | ydf0509/ | 895 | — | ~2.1k | Automated safety check: Pass | No licence | 2 mo ago |
| 8 | Sets up Arize Phoenix to trace, evaluate and monitor LLM applications, with instrumentation for OpenAI, LangChain and LlamaIndex and a self-hosted server. | Orchestra-Research/ | 13k | 2 repos | ~2.9k | Automated safety check: Pass | MIT | 3 mo ago |
| 9 | A skill your agent uses when changing Archestra tracing, metrics, OpenTelemetry, Tempo, Grafana, Prometheus, LLM/MCP spans, observability labels, or local observability setup. | archestra-ai/ | 4.4k | — | ~1.2k | Automated safety check: Pass | Unknown | today |
| 10 | A skill your agent uses when adding tracing, structured logging, metrics, or monitoring to a FrontMCP server. | agentfront/ | 146 | — | ~4.6k | Automated safety check: Pass | Apache-2.0 | today |
| 11 | Monitoring and observability strategy, implementation, and troubleshooting. | ahmedasmar/ | 203 | — | ~3.9k | Automated safety check: Pass | No licence | 6 mo ago |
| 12 | 12.Ops Monitor OPS on-demand: This skill should be used when the user asks to "datadog", "APM alerts", or… | Lifecycle-Innovations-Limited/ | 542 | 1 repo | ~1.5k | Automated safety check: Notes | MIT | today |
| 13 | LiteLLM-RS Observability Architecture. An agent skill from majiayu000/litellm-rs. | majiayu000/ | 118 | — | ~1.3k | Automated safety check: Pass | MIT | today |
| 14 | Author, modify, or review Netdata collectors across Go, IBM, C, Rust and external plugins. | netdata/ | 81k | — | ~1.9k | Automated safety check: Pass | GPL-3.0 | today |
| 15 | Sets up application monitoring: structured logs, Prometheus metrics, OpenTelemetry tracing, Grafana dashboards, alert rules and load tests with k6 or Artillery. | Jeffallan/ | 12k | — | ~1.6k | Automated safety check: Pass | MIT | 7 days ago |
| 16 | 16.Live Debug Debug the running local stack with traces, logs, and a shared headless browser. | macro-inc/ | 4.6k | — | ~2.4k | Automated safety check: Notes | AGPL-3.0 | today |
| 17 | Find out what a running Mendix app actually does — logs, Prometheus metrics, OpenTelemetry traces and the model catalog, joined across sources. | mendixlabs/ | 129 | — | ~2.8k | Automated safety check: Pass | Apache-2.0 | today |
| 18 | Build a unified telemetry pipeline with Grafana Alloy — one OpenTelemetry-compatible binary that collects metrics, logs, traces, and profiles and ships to Grafana Cloud / Prometheus / Loki / Tempo /… | grafana/ | 282 | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 19 | Go observability — always-on production signals: slog logging, Prometheus metrics, OpenTelemetry tracing, pprof profiling, alerting, Grafana. | context-labs/ | 1.1k | 1 repo | ~3.3k | Automated safety check: Pass | MIT | 6 days ago |
| 20 | Get RED metrics + service maps + frontend RUM + AI/LLM monitoring out of Grafana Cloud — Application Observability (tracesspanmetrics from OTel traces, p50/p95/p99 latency, exemplar-to-trace… | grafana/ | 282 | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 21 | 当需要为 funboost 任务添加监控、链路追踪或告警时使用。触发场景:Prometheus 指标、OpenTelemetry 链路追踪、异常告警通知、周期额度限制、函数结果持久化。关键词:Prometheus, OpenTelemetry, OTel, 告警, 监控, metrics, tracing, AlertNotifier, PeriodicQuota。 | ydf0509/ | 895 | — | ~4.9k | Automated safety check: Pass | No licence | 2 mo ago |
| 22 | Onboard an application into Elastic Observability with the Elastic Distribution of OpenTelemetry (EDOT): route on language and runtime, detect and replace a classic Elastic APM agent, apply the… | elastic/ | 592 | — | ~4.1k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 23 | Intent-based observability + traceability router across layers, boundaries, and signals. | first-fluke/ | 1.3k | — | ~4.9k | Automated safety check: Pass | MIT | yesterday |
| 24 | Unblock hosted Web Console OTLP to localhost:4318. An agent skill from forcedotcom/salesforcedx-vscode. | forcedotcom/ | 1k | — | ~796 | Automated safety check: Pass | BSD-3-Clause | yesterday |
| 25 | Triage a degraded or suspect service end to end: read SLO status and burn rate, check active alerting rules and ML anomalies, measure throughput, latency, and error rate, assess dependency health… | elastic/ | 592 | — | ~7.4k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 26 | Investigates server/infrastructure metric anomalies in PostHog Metrics — from "this metric is rising/dropping/spiking" or a fired alert to a probable cause with evidence. | PostHog/ | 40k | — | ~1.5k | Automated safety check: Pass | Unknown | yesterday |
| 27 | Auto-instrument an application's HTTP / gRPC / DB traffic with Grafana Beyla eBPF — no code changes, no SDK, no restart. | grafana/ | 282 | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 28 | Set up Grafana Cloud Database Observability for MySQL and PostgreSQL — enables pgstatstatements / Performance Schema, creates a least-privilege monitoring user, configures the… | grafana/ | 282 | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 29 | Diagnoses a Grafana Frontend Observability (RUM) session: whether it is healthy, what went wrong, ranked problems with timestamps and evidence, likely cause, how to fix it. | grafana/ | 282 | — | ~2.3k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 30 | Stand up Grafana Mimir for horizontally scalable, multi-tenant, long-term Prometheus + OTLP metrics storage. | grafana/ | 282 | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 31 | Instrument any app with OpenTelemetry and ship metrics / logs / traces to Grafana Cloud or self-hosted Mimir / Loki / Tempo / Pyroscope. | grafana/ | 282 | — | ~1.7k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 32 | Stand up Grafana Tempo as a cost-efficient distributed-tracing backend that only needs object storage, and write TraceQL queries against it. | grafana/ | 282 | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 33 | Add OpenTelemetry traces to an AG2 beta Agent via TelemetryMiddleware (autogen.beta.middleware.builtin). | ag2ai/ | 252 | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 34 | Signals scout for PostHog distributed tracing (APM / OpenTelemetry spans). | PostHog/ | 40k | — | ~4.5k | Automated safety check: Pass | Unknown | yesterday |
| 35 | Sending telemetry data to Grafana Cloud — metrics via Prometheus remote write or OTLP, logs via Loki push or Alloy, traces via OTLP to Tempo, profiles via Pyroscope. | grafana/ | 282 | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 36 | Design, audit, and troubleshoot production monitoring and observability using user-impact checks, layered telemetry, USE/RED, SLI/SLO/SLA, error budgets, cardinality controls, actionable alerting… | AnastasiyaW/ | 154 | — | ~4.1k | Automated safety check: Pass | MIT | yesterday |
| 37 | Observability and SRE expert. An agent skill from majiayu000/spellbook. | majiayu000/ | 287 | — | ~3.3k | Automated safety check: Pass | MIT | 2 days ago |
| 38 | Observability patterns for Python applications. An agent skill from aiskillstore/marketplace. | aiskillstore/ | 433 | 1 repo | ~1.3k | Automated safety check: Pass | No licence | yesterday |
| 39 | Network flow analysis in Dynatrace across three sources: OneAgent flows (host/process/pod-to-peer connections in the defaultnetworkflows Grail bucket), NetFlow/IPFIX/sFlow (via an OpenTelemetry… | Dynatrace/ | 163 | — | ~2k | Automated safety check: Pass | Apache-2.0 | 10 days ago |
| 40 | 40.Gorm Expert GORM v2 最佳实践与性能优化。适用于:代码审查、慢查询优化、N+1、连接池、 事务管理、分库分表、Prometheus/OTel监控、Session安全、Clause/Upsert、 缓存集成、BaseModel脚手架、SQL→struct生成、多租户隔离。 | LeoYeAI/ | 2.2k | — | ~3.4k | Automated safety check: Pass | MIT | 2 mo ago |
| 41 | OpenTelemetry, distributed tracing, structured logging, metrics (Prometheus, Grafana, Datadog). | TheBeardedBearSAS/ | 107 | — | ~547 | Automated safety check: Pass | MIT | 27 days ago |
| 42 | Monitoring and observability with OpenTelemetry, Prometheus, Grafana dashboards, and structured logging | rohitg00/ | 2.7k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | 5 mo ago |
| 43 | Observability: structured logs, metrics (RED/USE), tracing, SLO/SLI. | softspark/ | 179 | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 44 | Set up Apollo.io monitoring and observability. An agent skill from jeremylongshore/tons-of-skills-marketplace. | jeremylongshore/ | 2.8k | — | ~2.3k | Automated safety check: Pass | MIT | yesterday |
| 45 | Set up comprehensive observability for Deepgram integrations. | jeremylongshore/ | 2.8k | — | ~3k | Automated safety check: Pass | MIT | yesterday |
| 46 | A skill your agent uses when you need production monitoring for an Intercom integration — instrumenting API calls with metrics and traces, standing up dashboards, or wiring alerts for error rate… | jeremylongshore/ | 2.8k | — | ~1.7k | Automated safety check: Pass | MIT | yesterday |
| 47 | Set up observability for Klaviyo integrations with metrics, traces, and alerts. | jeremylongshore/ | 2.8k | — | ~1.5k | Automated safety check: Pass | MIT | yesterday |
| 48 | 48.Monitoring A skill your agent uses when setting up uptime and health monitoring, alerts, or on-call basics for a service already in production, so you learn it is down before customers do — health and… | ericrisco/ | 180 | — | ~3.1k | Automated safety check: Pass | MIT | yesterday |