Search
Observability
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 529 | Instrument a Webflow integration around API budgets, errors, cache behavior, webhooks, and deployments without leaking sensitive data. | jeremylongshore/ | 2.8k | — | ~1.2k | Automated safety check: Pass | MIT | yesterday |
| 530 | Monitor Devin Desktop (formerly Windsurf) AI adoption, feature usage, and team productivity metrics. | jeremylongshore/ | 2.8k | — | ~2.1k | Automated safety check: Pass | MIT | yesterday |
| 531 | 531.AI Observability Implement comprehensive observability for LLM applications including tracing (Langfuse/Helicone), cost tracking, token optimization, RAG evaluation metrics (RAGAS), hallucination detection, and… | omer-metin/ | 163 | — | ~578 | Automated safety check: Pass | Apache-2.0 | 8 mo ago |
| 532 | Makes systems debuggable and reliably operable — instrumentation, alerting that is worth waking for, service objectives, and learning from failure. | cbrock84/ | 2k | — | ~931 | Automated safety check: Pass | MIT | 23 days ago |
| 533 | Automatically trace Claude Code conversations to Braintrust for observability. | parcadei/ | 3.9k | — | ~1.5k | Automated safety check: Notes | MIT | 8 mo ago |
| 534 | Audit or improve observability for named production questions, incidents, or opaque operations using privacy-safe telemetry. | swyxio/ | 176 | — | ~931 | Automated safety check: Pass | MIT | 6 days ago |
| 535 | Comprehensive operational review procedures for Amazon Bedrock AgentCore resources aligned with the AWS Well-Architected Framework. | aws/ | 103 | — | ~4k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 536 | Validates and bootstraps Amazon Bedrock AgentCore observability so customers can trace agent reasoning, detect silent failures, and measure performance before an outage. | aws/ | 103 | — | ~3.4k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 537 | A skill your agent uses when someone wants an Amazon EKS cluster graded against best practices. | aws/ | 103 | — | ~3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 538 | A skill your agent uses when checking harness health, setting up observability cadences, understanding snapshot formats, configuring telemetry export, or verifying that the harness's own… | Habitat-Thinking/ | 114 | — | ~1.4k | Automated safety check: Pass | Unknown | 20 days ago |
| 539 | Push records from apps, services, devices, or scripts directly into Unity Catalog Delta tables with Zerobus Ingest: SDKs (Python, TypeScript, Go, Java, Rust, C++, .NET), REST, OTLP, MQTT, Arrow… | databricks/ | 345 | — | ~6.3k | Automated safety check: Pass | Unknown | yesterday |
| 540 | Enabling the built-in OpenTelemetry (OTLP) exporter for Effect-based Golem agents. | golemcloud/ | 1.5k | — | ~2.5k | Automated safety check: Pass | Unknown | yesterday |
| 541 | Adding structured logging and tracing spans to an Effect-based Golem agent. | golemcloud/ | 1.5k | — | ~1.6k | Automated safety check: Pass | Unknown | yesterday |
| 542 | 542.Logging Tracing and logging conventions for the Golem codebase. An agent skill from golemcloud/golem. | golemcloud/ | 1.5k | — | ~1.1k | Automated safety check: Pass | Unknown | yesterday |
| 543 | 543.Dt Sec Insights Query and analyze Dynatrace security data in security.events with DQL: vulnerabilities, threat detections, compliance posture, and scan coverage. | Dynatrace/ | 163 | — | ~7.2k | Automated safety check: Pass | Apache-2.0 | 10 days ago |
| 544 | Read-only audit of MCP definition language across an existing surface — tools, resources, prompts, server instructions. | cyanheads/ | 158 | — | ~4.9k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 545 | 545.Agent CLI Add agent-friendly --json NDJSON output to Python CLI scripts, or scaffold a complete cliutils package for a project. | glebis/ | 391 | — | ~3k | Automated safety check: Pass | MIT | 3 days ago |
| 546 | 546.Error Handling A skill your agent uses when designing the reaction to a class of failures — typed error taxonomies, retry/backoff/timeout policy, circuit breakers, React/Next error boundaries, and the user-message… | ericrisco/ | 180 | — | ~3.2k | Automated safety check: Pass | MIT | yesterday |
| 547 | 547.Monitoring A skill your agent uses when setting up uptime and health monitoring, alerts, or on-call basics for a service already in production, so you learn it is down before customers do — health and… | ericrisco/ | 180 | — | ~3.1k | Automated safety check: Pass | MIT | yesterday |
| 548 | 548.Observability A skill your agent uses when instrumenting a service from the inside so an incident can be explained from telemetry alone — wiring OpenTelemetry logs, metrics and traces, standing up a Collector… | ericrisco/ | 180 | — | ~3.8k | Automated safety check: Pass | MIT | yesterday |
| 549 | OpenTelemetry observability - tracing, metrics, logs, instrumentation, and context propagation patterns When user works with OpenTelemetry, adds tracing/metrics/logging, configures exporters, or… | shepherdjerred/ | 112 | — | ~4.2k | Automated safety check: Pass | GPL-3.0 | yesterday |
| 550 | 550.Operate Devops Plan and implement infrastructure, CI/CD, container, deployment, observability, and operational configuration changes with least privilege, staged validation, and rollback awareness. | hashgraph-online/ | 1.3k | — | ~618 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 551 | A skill your agent uses when connecting observed behavior, logs, metrics, request IDs, run IDs, screenshots, traces, external dependency results, or artifacts into a runtime evidence loop. | hashgraph-online/ | 1.3k | — | ~950 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 552 | A skill your agent uses when adding, changing, debugging, or reviewing Rust service error handling and observability, especially when separating domain errors from HTTP responses, adding… | hashgraph-online/ | 1.3k | — | ~969 | Automated safety check: Pass | MIT | yesterday |
| 553 | Provides AWS CloudFormation patterns for CloudWatch monitoring, metrics, alarms, dashboards, logs, and observability. | giuseppe-trisciuoglio/ | 357 | — | ~3.7k | Automated safety check: Notes | MIT | 1 mo ago |
| 554 | Provides comprehensive patterns for deploying Next.js applications to production. | giuseppe-trisciuoglio/ | 357 | — | ~2.3k | Automated safety check: Notes | MIT | 1 mo ago |
| 555 | 555.Agent Skills Datadog skills for AI agents. An agent skill from datadog-labs/agent-skills. | datadog-labs/ | 177 | — | ~951 | Automated safety check: Pass | MIT | 2 days ago |
| 556 | Add capabilities to an existing Agent Kernel project. An agent skill from yaalalabs/agent-kernel. | yaalalabs/ | 192 | — | ~13k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 557 | 557.Backend Backend engineering judgment, distilled from a stronger model - invoke when CHOOSING a tech stack, language, database, queue, or architecture; designing a service, API, business logic, or schema… | telagod/ | 244 | — | ~469 | Automated safety check: Pass | MIT | 2 mo ago |
| 558 | OpenTelemetry SDK and package version lookup across languages. | ollygarden/ | 106 | — | ~487 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 559 | Observability and monitoring validation patterns for dashboards, alerting, log aggregation, APM traces, and SLA/SLO verification. | proffesor-for-testing/ | 495 | — | ~8.3k | Automated safety check: Pass | MIT | yesterday |
| 560 | 560.Spring AI Diagnose and operate Spring AI projects with version-aware Maven or Gradle checks for ChatClient, advisors, retrieval, conversation memory, tool/MCP boundaries, streaming, configuration, and… | magnus919/ | 115 | — | ~1.2k | Automated safety check: Pass | MIT | yesterday |
| 561 | 561.Telemetry Operate the observability stack that deploys as one unit: Prometheus scrape configuration, recording and alerting rules, relabeling, retention, and high availability; OpenTelemetry Collector… | magnus919/ | 115 | — | ~3.9k | Automated safety check: Pass | MIT | yesterday |
| 562 | Inspects the OrchestKit telemetry pipeline for the current project — lists all known telemetry files with write counts, sizes, schema status, growth trend, and orphan detection. | yonatangross/ | 292 | — | ~2.8k | Automated safety check: Notes | MIT | yesterday |
| 563 | Design production observability strategies covering SLI/SLOs, metrics, logs, traces, dashboards, and alert quality. | aAAaqwq/ | 105 | 1 repo | ~3.3k | Automated safety check: Pass | MIT | 2 days ago |
| 564 | World-Class Technology & Data Playbook. An agent skill from LeoYeAI/openclaw-master-skills. | LeoYeAI/ | 2.2k | — | ~6.8k | Automated safety check: Pass | MIT | 2 mo ago |
| 565 | Observability patterns for logging, monitoring, alerting, and distributed tracing. | TheSoftwareHouse/ | 284 | — | ~2k | Automated safety check: Pass | MIT | 5 days ago |
| 566 | Sets up CloudWatch observability for the first time - Omni (CloudWatch Application Observability) and classic CloudWatch. | aws/ | 2.8k | — | ~5k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 567 | Scan project root for file tree, dominant languages, and LOC estimate. | digipulse-engineering/ | 163 | — | ~1.4k | Automated safety check: Pass | Unknown | 12 days ago |
| 568 | Set up observability for Claude API integrations with metrics, logging, and alerting for latency, cost, errors, and token usage. | jeremylongshore/ | 2.8k | — | ~1.8k | Automated safety check: Pass | MIT | yesterday |
| 569 | Monitor Apple Notes automation health and performance metrics. | jeremylongshore/ | 2.8k | — | ~1.8k | Automated safety check: Pass | MIT | yesterday |
| 570 | Monitor Clay enrichment pipeline health, credit consumption, and data quality metrics. | jeremylongshore/ | 2.8k | — | ~2.2k | Automated safety check: Pass | MIT | yesterday |
| 571 | Set up GPU monitoring and observability for CoreWeave workloads. | jeremylongshore/ | 2.8k | — | ~1.4k | Automated safety check: Pass | MIT | yesterday |
| 572 | Track: documents indexed per run (total + new + updated + deleted), indexing errors and retries, search API latency, zero-result query rate, stale content age distribution. | jeremylongshore/ | 2.8k | — | ~1.4k | Automated safety check: Pass | MIT | yesterday |
| 573 | Execute Langfuse primary workflow: Tracing LLM calls and spans. | jeremylongshore/ | 2.8k | — | ~2.3k | Automated safety check: Pass | MIT | yesterday |
| 574 | Set up comprehensive observability for Lokalise integrations with metrics, traces, and alerts. | jeremylongshore/ | 2.8k | — | ~2.9k | Automated safety check: Pass | MIT | yesterday |
| 575 | Retell AI observability — AI voice agent and phone call automation. | jeremylongshore/ | 2.8k | — | ~515 | Automated safety check: Pass | MIT | yesterday |
| 576 | Build Salesforce integration observability across application traces, platform status, limits, async jobs, events, logs, and business reconciliation. | jeremylongshore/ | 2.8k | — | ~1.1k | Automated safety check: Pass | MIT | yesterday |