This week, Sean sat down with Emily Arnott of Blameless, who is making it her mission to spread βthe Gospel of SRE.β Their discussion covered the philosophy underpinning Site Reliability Engineering, its origins in the world of manufacturing, and a few detailed scenarios for how this approach plays out in real-world incident response teams. π How does an approach that bakes operational wisdom into the development and release process from the beginning change things for teams? π What is the typical structure of an SRE team, and the incident command role? (And how is it similar to a military training team?) π How does an SRE approach differ, when βthe bodies hit the floorβ and the team needs to rapidly triage incidents? π How is the evolving language and taxonomy around SRE changing our paradigms for how we think about incident response? Emily on LinkedIn: https://www.linkedin.com/in/emily-arn... Blameless: https://www.blameless.com/ For further reading: ποΈβπ¨οΈ So you Want an SRE Tool. Do you Build, Buy, or Open Source?: https://dev.to/blameless/so-you-want-... ποΈβπ¨οΈ SRE: From Theory to Practice | What's difficult about incident command: https://dev.to/blameless/sre-from-the... ποΈβπ¨οΈ DevOps & SRE Words Matter: How Our Language has Evolved: https://dev.to/blameless/devops-sre-w... ποΈβπ¨οΈ SRECon presentation: https://www.usenix.org/conference/sre