DevOps Engineering Case Studies

SRE

Incident response, postmortems, SLOs, error budgets and MTTR. This section is about the operating discipline rather than any particular technology.

Written

None yet.

Planned

Case study Type
A production outage, start to finish, with a blameless postmortem Incident
Designing SLIs and SLOs for a service nobody has measured Architecture
Error budget policy and what it actually changes about shipping Architecture
Reducing MTTR: where the minutes really go Optimisation
Capacity planning from first principles Architecture

Nothing here yet — content is on the roadmap.