Fifteen years building and scaling distributed systems, across the industry's move from bare-metal servers and monoliths to cloud-native, Kubernetes and GitOps. I take on the work teams postpone: paying down technical debt, building deployment from nothing, and keeping systems running.
Six shifts the industry has since packaged and sells ready-made. I worked through them before anything was ready-made.
DevOps When It Was Two Departments
2016
Development and operations had different targets and different definitions of done. An incident is almost never purely technical.
Development and operations had different targets and different definitions of the word "done" — one was measured on shipping, the other on nothing breaking, and each was rewarded for the other’s caution being wrong.
Standing between them is where I learned the part that has not changed: an incident is almost never purely technical. The technical cause is real, and the reason nobody caught it earlier is organisational.