4 — Incident Detection
Purpose
[stub: incident-detection]
Metadata
| Author | Amit Singh |
| Scope | observability |
Local graph
Related notes
1 — SLIs
Covers choosing a Service Level Indicator that actually reflects user-perceived reliability, not just what's easiest to measure.
3 — Error Budgets
Covers treating the error budget as a spendable risk resource that governs release velocity, not a compliance scorecard.
6 — Postmortems
Covers writing a blameless postmortem that traces the incident timeline back to instrumentation and observability gaps, not just the code fix.
7 — Chaos Engineering
Covers using deliberate fault injection to validate that observability signals actually fire the way an incident response plan assumes.