Observability Slo

Observability and reliability engineering — structured logging, metrics, distributed tracing, correlation ids, error tracking, dashboards, SLIs/SLOs and error budgets, alerting that pages on symptoms rather than causes, on-call practice, incident response and blameless postmortems. Use when the user says "logging", "monitoring", "observability", "metrics", "tracing", "Prometheus", "Grafana", "Datadog", "Sentry", "OpenTelemetry", "SLO", "SLA", "uptime", "alerting", "on-call", "incident", "postmortem", "how do we know if it breaks", "it broke and we didn't notice" or "debugging production"; and as a pass in any project audit. By Devleck.

Kin9Zeus d1ffcd1 3 files · 28.6 KB Updated

File contents

Kin9Zeus/senior-engineer-skills/tree/main/skills/observability-slo commit d1ffcd1704

Frequently asked questions

npx skillmds@latest add kin9zeus/observability-slo