Agent skill

Oma Observability

by first-fluke in first-fluke/oh-my-agent

Intent-based observability + traceability router across layers, boundaries, and signals.

MITAuto-check passedDevOps & Cloud

Install Oma Observability

skills CLI
$ npx skills add first-fluke/oh-my-agent --skill oma-observability -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install first-fluke/oh-my-agent oma-observability --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/first-fluke/oh-my-agent.git skills-src && mkdir -p .claude/skills && cp -r skills-src/benchmarks/runs/oma/.agents/skills/oma-observability .claude/skills/oma-observability && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
oma-observability
GitHub stars
1.3k
Token cost
~4.9k tokens
SKILL.md length
1,700 words
Files
35
Skills in repo
57
Repo updated
First seen
Licence
MIT

At a glance

Intent-based observability + traceability router across layers, boundaries, and signals.

  • Works in 3 steps: Classify the intent: setup, migrate,… → Identify layers, boundaries, signals,… → Load only the relevant resource guide(s).
  • Incident forensics
  • SKILL.md covers Scheduling, Structural Flow and Logical Operations
  • Calls git; reaches landscape.cncf.io

What it does

Oma Observability is an agent skill from first-fluke/oh-my-agent. Intent-based observability + traceability router across layers, boundaries, and signals. Routes to vendor-specific skills via category taxonomy; owns transport tuning, meta-observability, incident forensics. Use for observability, traceability, telemetry, APM, RUM, metrics, logs, traces, profiles, SLO, incident forensics, tracing architecture work.

Its SKILL.md is about 4.9k tokens, which your agent loads only when the skill is triggered. The skill folder holds 38 other files (for example `resources/anti-patterns.md`, `resources/boundaries/cross-application.md` and `resources/boundaries/multi-tenant.md`).

It sits in DevOps & Cloud, covering Observability, Monitoring and alerting and Site reliability engineering. It works with OpenTelemetry. The repository describes itself as: Mechanical verification for AI coding agents — skills pack or full harness (stop-hook gates, artifact checks, independent judges). The licence is MIT.

When your agent uses it

  • Incident forensics
  • Tracing architecture work

Example prompts

  • “/oma-observability”

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. Classify the intent: setup, migrate, investigate, alert, trace, tune, or route.
  2. Identify layers, boundaries, signals, and vendor category.
  3. Load only the relevant resource guide(s).

What it can do on your machine

Read from SKILL.md and the folder at commit f65bbc0. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • git

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • landscape.cncf.io

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Oma Observability loads about 4.9k tokens when it runs. Until then it costs about 92 tokens; SKILL.md has 1,700 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~92
When it runs · the whole SKILL.md, loaded when a task matches
~4.9k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from first-fluke/oh-my-agent at commit f65bbc0, republished under its MIT licence (© first-fluke). 1,700 words, ~4,890 tokens.

Download SKILL.mdSave it as .claude/skills/oma-observability/SKILL.md (or your agent's skills folder). This skill also uses 34 other files; get the full folder from GitHub.
name
oma-observability
description
Intent-based observability + traceability router across layers, boundaries, and signals. Routes to vendor-specific skills via category taxonomy; owns transport tuning, meta-observability, incident forensics. Use for observability, traceability, telemetry, APM, RUM, metrics, logs, traces, profiles, SLO, incident forensics, tracing architecture work.

Observability Agent - Intent-based Router

Scheduling

Goal

Route, design, tune, and review observability work across MELT+P signals, layers, boundaries, vendor categories, transport choices, meta-observability, and incident forensics.

Intent signature
  • User asks for observability, telemetry, OTel, metrics, logs, traces, profiles, SLOs, RUM, APM, incident forensics, trace propagation, transport tuning, or observability-as-code.
  • User needs vendor/category routing or observability architecture instead of a single vendor's already-covered setup.
When to use
  • Setting up an observability pipeline (OTel SDK + Collector + vendor backend)
  • Designing traceability across service and domain boundaries (W3C propagators, baggage, multi-tenant, multi-cloud)
  • Tuning transport layer (UDP/MTU, OTLP gRPC vs HTTP, Collector DaemonSet vs sidecar topology)
  • Running incident forensics (6-dimension localization: code / service / layer / host / region / infra)
  • Selecting a vendor category (OSS full-stack vs commercial SaaS vs high-cardinality specialist vs profiling specialist)
  • Implementing observability-as-code (Grafana Jsonnet dashboards, PrometheusRule CRD, OpenSLO YAML, SLO burn-rate alerts)
  • Meta-observability (pipeline self-health, clock skew detection, cardinality guardrails, retention matrix)
  • Covering the MELT+P signal set: metrics, logs, traces, profiles (OTEP 0239), cost (OpenCost), audit (SOC2/ISO), privacy (GDPR/PIPA)
  • Migrating off deprecated tools (Fluentd → Fluent Bit or OTel Collector, per CNCF 2025-10 guide)
When NOT to use
  • LLM ops (prompt versioning, evals, gen_ai span deep dive) — use Langfuse, Arize Phoenix, LangSmith, or Braintrust directly
  • Data pipeline lineage — use OpenLineage + Marquez, dbt test, or Airflow lineage backends
  • IoT / hardware / datacenter physical-layer telemetry (IPMI, BMC, SNMP) — use vendor DCIM tooling (Nlyte, Sunbird, Device42)
  • Chaos engineering orchestration — use Chaos Mesh, Litmus, Gremlin, or ChaosToolkit (this skill consumes their telemetry; it does not orchestrate chaos)
  • GPU / TPU infrastructure observability — use NVIDIA DCGM Exporter + Prometheus
  • Software supply chain (SBOM, attestation) — use sigstore (cosign / rekor), in-toto framework, SLSA level attestations
  • Incident response workflow (on-call rotation, paging, escalation) — use PagerDuty, OpsGenie, or Grafana OnCall
  • Single-vendor setup already fully covered by that vendor's own published skill — invoke the vendor skill directly
Expected inputs
  • Observability intent, target system, architecture boundary, signals, vendor context, and incident symptoms if any
  • Existing OTel/collector/vendor configs, dashboards, SLOs, trace/log/metric examples, or deployment topology
Expected outputs
  • Routed observability guidance, setup/migration/tuning plan, incident-forensics path, alerting/SLO guidance, or observability-as-code recommendations
  • Transport, meta-observability, privacy, audit, and retention checks
  • Vendor delegation target when appropriate
Dependencies
  • OTel/W3C/CNCF references and resources under resources/
  • Vendor categories, matrix, standards, incident forensics, meta-observability, transport, layers, boundaries, and signal guides
Control-flow features
  • Branches by intent, vendor category, layer/boundary/signal matrix, transport topology, privacy/audit risk, and incident localization dimension
  • May read/write observability config and docs; generally delegates vendor-specific implementation
  • Requires live status verification for load-bearing CNCF/vendor currency

Structural Flow

Entry
  1. Classify the intent: setup, migrate, investigate, alert, trace, tune, or route.
  2. Identify layers, boundaries, signals, and vendor category.
  3. Load only the relevant resource guide(s).
Scenes
  1. PREPARE: Classify intent and matrix coverage.
  2. ACQUIRE: Read configs, topology, telemetry examples, or incident signals.
  3. REASON: Route vendor/category, tune transport, assess meta-observability, or localize incident.
  4. ACT: Produce setup/migration/tuning/alert/trace/forensics guidance or config changes.
  5. VERIFY: Check pipeline health, clock skew, cardinality, retention, privacy, and audit concerns.
  6. FINALIZE: Report route, evidence, risks, and handoff references.
Transitions
  • If a vendor-owned skill fully covers setup, delegate instead of duplicating docs.
  • If Fluentd appears, recommend Fluent Bit or OTel Collector migration.
  • If incident investigation is requested, use 6-dimensional localization.
  • If transport tuning appears, load transport-specific resources.
Failure and recovery
  • If live CNCF/vendor status is load-bearing, verify current status.
  • If telemetry samples are missing, provide instrumentation/collection steps before analysis.
  • If scope belongs to out-of-scope domains, route to external authoritative tools.
Exit
  • Success: observability path is routed, evidence-backed, and checks are explicit.
  • Partial success: missing telemetry, stale vendor status, or external-domain handoff is explicit.

Logical Operations

Actions
ActionSSL primitiveEvidence
Classify observability intentSELECTIntent rules
Read telemetry/config evidenceREADOTel/vendor configs, dashboards, samples
Route vendor/categorySELECTVendor categories
Infer coverage gapsINFERMatrix and signal/boundary mapping
Validate meta-observabilityVALIDATEClock, cardinality, retention, health
Write guidance/configWRITEOaC/config/docs when requested
Notify resultNOTIFYRouted recommendation
Tools and instruments
  • OTel/CNCF/W3C standards references
  • Vendor categories, matrix, incident forensics, meta-observability, transport and signal guides
  • Optional CLI/config tooling from the target stack
Canonical workflow path
text
1. Classify intent: setup, migrate, investigate, alert, trace, tune, or route.
2. Select layer/boundary/signal coverage from `resources/matrix.md`.
3. Load the specific vendor, transport, incident, or signal guide before producing guidance.

When CNCF/vendor status is load-bearing, verify live state at https://landscape.cncf.io.

Resource scope
ScopeResource target
CODEBASEObservability config, dashboards, alert rules, instrumentation
LOCAL_FSResource guides and generated docs
NETWORKVendor/CNCF status and telemetry backends when checked
USER_DATAIncident symptoms, logs, metrics, traces, profiles
Preconditions
  • Observability intent and system boundary are identifiable.
  • Relevant telemetry/config evidence is available or missing evidence is stated.
Effects and side effects
  • May recommend or modify observability config, dashboards, alerts, and instrumentation docs.
  • May route to vendor-owned skills or external tools.
Guardrails
  1. Classify intent before routing: every query goes through intent classification — setup | migrate | investigate | alert | trace | tune | route
  2. Category-first, not vendor-registry: delegate to vendor-owned skills via resources/vendor-categories.md; do not duplicate their documentation
  3. Transport tuning is the moat: UDP/MTU thresholds, OTLP protocol selection, Collector topology, and sampling recipes are in-skill depth that other skills do not cover
  4. Meta-observability is non-negotiable: always validate pipeline self-health, clock sync (< 100 ms drift), cardinality, and retention before declaring setup complete
  5. CNCF-first preference: Prometheus, Jaeger, Thanos, Fluent Bit, OpenFeature (Graduated 2024-11), Flagger, Falco (Graduated); OpenTelemetry, Cortex, OpenCost (Incubating)
  6. Fluentd is deprecated: per CNCF 2025-10 migration guide, recommend Fluent Bit or OTel Collector for all new and migration work
  7. W3C Trace Context as default propagator: translate per cloud (AWS X-Ray X-Amzn-Trace-Id, GCP Cloud Trace, Datadog, Cloudflare, Linkerd) via boundaries/cross-application.md
  8. Privacy before features: PII redaction, sampling-aware baggage rules, and compliance (SOC2/ISO immutable audit + GDPR/PIPA erasure) are applied at collection, not only at storage
  9. Domain-level trust: all vendor and tool references are timestamped as of 2026-Q2; verify live status at https://landscape.cncf.io
  10. No stub in final deliverable: scaffolds are editing anchors only during build phase; remove before output
Out of Scope (use external tools)

The combinations below are outside this skill's boundary. The external tools listed are authoritative for each domain.

DomainExternal tools
LLM ops / gen_ai observabilityLangfuse, Arize Phoenix, LangSmith, Braintrust
Data pipeline lineageOpenLineage + Marquez, dbt test, Apache Airflow lineage
L1/L2 physical / datacenter hardwareNlyte, Sunbird, Device42; SNMP exporters where Prometheus bridge is needed
L5 Session / L6 Presentation full TLS inspectionWireshark (packet-level), Cloudflare Radar (TLS ecosystem data), vendor TLS inspection tooling
Chaos engineering orchestrationChaos Mesh, Litmus, Gremlin, ChaosToolkit
GPU / AI infra (DCGM, NVIDIA)NVIDIA DCGM Exporter + Prometheus; OTel GPU semconv (Development, not production-ready)
Software supply chain (SBOM, attestation)sigstore (cosign / rekor), in-toto framework, SLSA level attestations
Incident response workflow (paging, rotation)PagerDuty, OpsGenie, Grafana OnCall
Fluentd (primary tool)Deprecated CNCF 2025-10 — use Fluent Bit or OTel Collector
Show full SKILL.md (616 more words)Show less
Architecture (4 x 4 x 7 matrix)
                  User / Other Skill Query
                            |
                            v
              +-----------------------------+
              |      Intent Classifier      |
              |  setup | migrate | investigate
              |  alert | trace | tune | route|
              +-----------------------------+
                            |
                            v
              +-----------------------------+
              |      Vendor Router          |
              |  category-first delegation  |
              +-----------------------------+
                            |
                            v
              +-----------------------------+
              |   vendor-categories.md      |
              |   (a) OSS Full-Stack        |
              |   (b) Commercial SaaS APM   |
              |   (c) High-Cardinality      |
              |   (d) Profiling Specialist  |
              |   (e) SIEM / Enterprise Logs|
              |   (f) FinOps / Cost         |
              |   (g) Feature Flags/Rollout |
              |   (h) Log Pipeline          |
              |   (i) Time Series Storage   |
              |   (j) Crash Analytics       |
              +-----------------------------+
                            |
                            v
              +-----------------------------+
              |  Matrix Coverage Selector   |
              |  4 Layers x 4 Boundaries    |
              |  x 7 Signals = 112 cells    |
              +-----------------------------+
                            |
                            v
              +-----------------------------+
              |  Transport Depth /          |
              |  Meta-observability         |
              |  UDP, OTLP, Collector,      |
              |  cardinality, clock skew    |
              +-----------------------------+
                            |
                            v
              +-----------------------------+
              |  Incident Forensics         |
              |  6-dim localization:        |
              |  code/service/layer/host/   |
              |  region/infra               |
              +-----------------------------+

Layers (4): L3-network, L4-transport, mesh, L7-application Boundaries (4): multi-tenant, cross-application, slo, release Signals (7): metrics, logs, traces, profiles, cost, audit, privacy

See resources/matrix.md for the full 112-cell coverage map with N/A markers for invalid combinations.

Routes (Intent)
IntentPrimary targetFallback
setupresources/vendor-categories.md → vendor-owned skillGeneric OTel semconv in resources/standards.md
migrateCNCF 2025-10 guide + resources/vendor-categories.md §(h)OTel Collector bridge config
investigateresources/incident-forensics.md (MRA + 6-dim localization)signals/traces.md + signals/logs.md
alertboundaries/slo.md (burn-rate alert rules)resources/observability-as-code.md
traceboundaries/cross-application.md (propagator matrix)layers/mesh.md (zero-code auto-instrumentation)
tunetransport/ (4 files: UDP/MTU, OTLP, topology, sampling)resources/meta-observability.md (cardinality guardrails)
routeboundaries/multi-tenant.md + transport/collector-topology.mdboundaries/cross-application.md (data residency)
Invocation

Standalone:

/oma-observability "set up OTel stack on Kubernetes"
/oma-observability --migrate "move from Fluentd to Fluent Bit"
/oma-observability --investigate "5xx spike in ap-northeast-2"
/oma-observability --alert "configure SLO burn-rate alert for checkout API"
/oma-observability --trace "W3C propagator across AWS + GCP boundary"
/oma-observability --tune "UDP statsd MTU throughput limit"
/oma-observability --route "multi-tenant log isolation with data residency"

Shared invocation (from other skills):

  1. State intent: setup | migrate | investigate | alert | trace | tune | route
  2. Pass the user query string
  3. Receive routed guidance or a vendor-skill delegation target
How to Execute

Follow resources/execution-protocol.md step by step. See resources/examples.md for end-to-end walkthroughs. Use resources/intent-rules.md for intent classification reference. Use resources/matrix.md for coverage navigation across layers, boundaries, and signals. Use resources/vendor-categories.md for vendor delegation and category selection. Before submitting, run resources/checklist.md.

Integrations with OMA Ecosystem

Integration status (2026-Q2): rows below describe recommended handoff patterns from the oma-observability side. As of this version, reciprocal cross-references from the other skills' SKILL.md files are not yet in place — this is a v1.1 follow-up item. Users invoking the other skills directly will need to surface this integration manually until the reciprocal links land.

SkillIntegration pointReciprocal link status
oma-debugOn failure: pull traces + logs by request_id → trigger resources/incident-forensics.md 6-dim localization playbook⏳ pending (v1.1)
oma-qaCanary post-deploy loop via chrome-devtools MCP: console errors + Core Web Vitals trend; INP/LCP/CLS from layers/L7-application/web-rum.md⏳ pending (v1.1)
oma-tf-infraTerraform modules for OTel Collector, Grafana, and Loki stack provisioning⏳ pending (v1.1)
oma-scmDeployment SHA → service.version OTel attribute + release marker events; see boundaries/release.md⏳ pending (v1.1)
oma-backendPropagator and baggage rules cross-referenced in backend.md ruleset; DB N+1 + Kafka patterns in signals/traces.md⏳ pending (v1.1)
oma-frontendlayers/L7-application/web-rum.md INP/LCP/CLS checklist cross-referenced in frontend.md ruleset⏳ pending (v1.1)
oma-mobilelayers/L7-application/mobile-rum.md offline-queuing pattern cross-referenced in mobile.md ruleset⏳ pending (v1.1)
oma-dbsignals/traces.md DB patterns (N+1, connection pool) cross-referenced in database.md ruleset⏳ pending (v1.1)
Versioning & Deprecation
  • Spec version pinning: otel_spec / otel_semconv keys in each file's frontmatter document the assumed version. If content depends on a specific attribute stability tier, the tier is stated inline.
  • Update triggers (not scheduled):
    • OTel semconv promotion (Development → RC → Stable) affecting attributes cited in this skill → update resources/standards.md and the affected file, bump minor version.
    • Attribute deprecation → replace across all citing files; migration note in resources/standards.md.
    • CNCF status change for a vendor/project named in vendor-categories.md (Graduated / Archived / acquired) → update the vendor table.
  • Authoritative live state: https://landscape.cncf.io for CNCF project status. This skill does not promise to track it on any schedule — verify at use time if the information is load-bearing.
  • No per-file review stamps: earlier drafts carried last_reviewed / next_review frontmatter. Those were removed because no automated enforcement exists; relying on voluntary manual review produces stale stamps that misrepresent currency. Git history (git log path/to/file) is the source of truth for when a file was last changed.
Contribution Protocol
  • Do NOT pre-declare future OMA skill names in user-facing documentation. If OMA-native coverage becomes warranted for an out-of-scope domain, evaluate and name it at that point.
  • File edits follow the ownership matrix in docs/plans/designs/005-oma-observability.md §Ownership. CTO co-signs changes to standards.md, matrix.md, anti-patterns.md.
  • Run resources/checklist.md §1 Setup validation before merging.

References

  • Execution steps: resources/execution-protocol.md
  • Intent classification: resources/intent-rules.md
  • Coverage matrix: resources/matrix.md
  • Standards (OTel spec, W3C, ISO): resources/standards.md
  • Vendor categories: resources/vendor-categories.md
  • Incident forensics: resources/incident-forensics.md
  • Meta-observability: resources/meta-observability.md
  • Observability-as-code: resources/observability-as-code.md
  • Anti-patterns (18 items): resources/anti-patterns.md
  • Checklist: resources/checklist.md
  • Examples: resources/examples.md
  • Transport:
    • resources/transport/udp-statsd-mtu.md
    • resources/transport/otlp-grpc-vs-http.md
    • resources/transport/collector-topology.md
    • resources/transport/sampling-recipes.md
  • Layers:
    • resources/layers/L3-network.md
    • resources/layers/L4-transport.md
    • resources/layers/mesh.md
    • resources/layers/L7-application/web-rum.md
    • resources/layers/L7-application/mobile-rum.md
    • resources/layers/L7-application/crash-analytics.md
  • Boundaries:
    • resources/boundaries/multi-tenant.md
    • resources/boundaries/cross-application.md
    • resources/boundaries/slo.md
    • resources/boundaries/release.md
  • Signals:
    • resources/signals/metrics.md
    • resources/signals/logs.md
    • resources/signals/traces.md
    • resources/signals/profiles.md
    • resources/signals/cost.md
    • resources/signals/audit.md
    • resources/signals/privacy.md

© first-fluke, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 34 other files in benchmarks/runs/oma/.agents/skills/oma-observability of first-fluke/oh-my-agent.

  • SKILL.md
  • resources/anti-patterns.md
  • resources/boundaries/cross-application.md
  • resources/boundaries/multi-tenant.md
  • resources/boundaries/release.md
  • resources/boundaries/slo.md
  • resources/checklist.md
  • resources/examples.md
  • resources/execution-protocol.md
  • resources/incident-forensics.md
  • resources/intent-rules.md
  • resources/layers/L3-network.md
  • resources/layers/L4-transport.md
  • resources/layers/L7-application/crash-analytics.md
  • resources/layers/L7-application/mobile-rum.md
  • resources/layers/L7-application/web-rum.md
  • resources/layers/mesh.md
  • … and 18 more

Open the folder on GitHubat commit f65bbc0

Compare with similar skills

Oma Observability next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Oma Observability compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Oma Observability this skillfirst-fluke/oh-my-agent1.3k—~4.9kAutomated safety check: PassMIT
Monitoring Observabilityahmedasmar/devops-claude-skills203—~3.9kAutomated safety check: PassNone
Observability Sre Triageelastic/agent-skills592—~7.4kAutomated safety check: PassApache-2.0
Observability MonitoringAnastasiyaW/codex-claude-code-config154—~4.1kAutomated safety check: PassMIT
Observability Sremajiayu000/spellbook286—~3.3kAutomated safety check: PassMIT
Observability Patternssoftspark/ai-toolkit179—~2.2kAutomated safety check: PassApache-2.0

Similar skills

  • Monitoring Observability

    ahmedasmar/devops-claude-skills

    Monitoring and observability strategy, implementation, and troubleshooting.

    203 GitHub stars~3.9k tokensUpdated 6 mo ago
    DevOps & CloudAuto-check passed
  • Observability Sre Triage

    elastic/agent-skills

    Official

    Triage a degraded or suspect service end to end: read SLO status and burn rate, check active alerting rules and ML anomalies, measure throughput, latency, and error rate, assess dependency health…

    592 GitHub stars~7.4k tokensUpdated yesterday
    DevOps & CloudAuto-check passed
  • Observability Monitoring

    AnastasiyaW/codex-claude-code-config

    Design, audit, and troubleshoot production monitoring and observability using user-impact checks, layered telemetry, USE/RED, SLI/SLO/SLA, error budgets, cardinality controls, actionable alerting…

    154 GitHub stars~4.1k tokensUpdated today
    DevOps & CloudAuto-check passed
  • Observability Sre

    majiayu000/spellbook

    Observability and SRE expert. An agent skill from majiayu000/spellbook.

    286 GitHub stars~3.3k tokensUpdated today
    DevOps & CloudAuto-check passed
  • Observability Patterns

    softspark/ai-toolkit

    Observability: structured logs, metrics (RED/USE), tracing, SLO/SLI.

    179 GitHub stars~2.2k tokensUpdated yesterday
    DevOps & CloudAuto-check passed
  • Monitoring

    ericrisco/rsc-harness

    A skill your agent uses when setting up uptime and health monitoring, alerts, or on-call basics for a service already in production, so you learn it is down before customers do — health and…

    167 GitHub stars~3.1k tokensUpdated today
    DevOps & CloudAuto-check passed

More from first-fluke/oh-my-agent

All 57 skills in this repo
  • OMA Multi-Agent Orchestration

    first-fluke/oh-my-agent

    Decomposes a complex feature into tasks, dispatches parallel specialist agents with durable state, and supervises verification, QA review and retries.

    1.3k GitHub stars~4.1k tokensUpdated today
    Auto-check passed
  • Oma Video

    first-fluke/oh-my-agent

    Create short, explainer, or recorded-demo videos through the OMA video CLI.

    1.3k GitHub starsUsed in 1 repo~1.7k tokens
    Auto-check passed
  • OMA Multi-Agent Orchestrator

    first-fluke/oh-my-agent

    Splits a complex feature into prioritized tasks, spawns specialist CLI subagents in parallel, tracks them through shared memory and verifies each result.

    1.3k GitHub stars~3.1k tokensUpdated today
    Auto-check passed
  • Architecture Decisions and ADRs

    first-fluke/oh-my-agent

    Evaluates system boundaries and tradeoffs and writes architecture recommendations, option comparisons or ADRs, with a Mermaid diagram when structure changes.

    1.3k GitHub stars~2.6k tokensUpdated today
    Auto-check passed
  • OMA Backend Agent

    first-fluke/oh-my-agent

    Backend specialist for APIs, database work, authentication and migrations that follows clean architecture with router, service and repository layers.

    1.3k GitHub stars~2.4k tokensUpdated today
    Auto-check passed
  • oma Bootstrap

    first-fluke/oh-my-agent

    Installs or checks the oma CLI and its runtimes (bun, uv, serena) in a fresh workspace so that oma-* skills can run their commands.

    1.3k GitHub stars~719 tokensUpdated today
    Auto-check passed

Works with

Categories

Questions about Oma Observability

What does Oma Observability do?

Intent-based observability + traceability router across layers, boundaries, and signals. Oma Observability is an agent skill from first-fluke/oh-my-agent. Intent-based observability + traceability router across layers, boundaries, and signals.

When should I use Oma Observability?

Oma Observability fits situations like: incident forensics; tracing architecture work.

How do I install Oma Observability in Claude Code?

Run `npx skills add first-fluke/oh-my-agent --skill oma-observability -a claude-code`. Or copy the skill folder (benchmarks/runs/oma/.agents/skills/oma-observability in first-fluke/oh-my-agent) into .claude/skills/oma-observability in your project. Claude Code loads it when a task matches its description.

How do I install Oma Observability in Codex?

Run `npx skills add first-fluke/oh-my-agent --skill oma-observability -a codex`. Or copy the skill folder (benchmarks/runs/oma/.agents/skills/oma-observability in first-fluke/oh-my-agent) into .agents/skills/oma-observability in your project. Codex loads it when a task matches its description.

Can I use Oma Observability in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add first-fluke/oh-my-agent --skill oma-observability -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/oma-observability, .gemini/skills/oma-observability, .github/skills/oma-observability and .opencode/skills/oma-observability in your project.

What does Oma Observability need to run?

Going by SKILL.md and its folder, Oma Observability needs the command-line tools its instructions call (git).

Does Oma Observability access the network?

SKILL.md names 1 domain. In commands or code: landscape.cncf.io; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.

Is Oma Observability safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Oma Observability use?

Oma Observability is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Oma Observability use?

About 4.9k tokens (SKILL.md is roughly 20k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Oma Observability?

Skills that share tags, products or a category with Oma Observability: Monitoring Observability (ahmedasmar/devops-claude-skills, 203 stars), Observability Sre Triage (elastic/agent-skills, 592 stars), Observability Monitoring (AnastasiyaW/codex-claude-code-config, 154 stars) and Observability Sre (majiayu000/spellbook, 286 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Oma Observability?

first-fluke (a GitHub organization) maintains it in first-fluke/oh-my-agent, which has 1,338 GitHub stars. The repository holds 57 skills in this directory. The repository was last updated on October 8, 2026.

Source: first-fluke/oh-my-agent on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.