Conference Program

iSAQB Software Architecture Gathering 2026

  • Session (45min)
  • Introductory
  • 18 Nov 2026
  • 16:15-17:00
  • Room Leipzig

Incident Recovery Through a GreenOps Lens

by Sarah Hsu

GreenOps and incident management are usually treated as separate concerns: one saves the planet, the other saves the weekend. But they often share the same root cause: waste. Carbon waste and system instability tend to come from the same architectural flaws.

When a P0 hits, the playbook is predictable. Scale up, add redundancy, restore service. Efficiency takes a back seat to uptime, as it should. The problem is what happens next. Emergency response leaves a carbon hangover behind: over-provisioned clusters, zombie infrastructure, and resource sprawl that stay in production long after the crisis is over. It is hidden technical debt that masks systemic weakness and slows every team shipping into that environment afterwards.

This talk reframes GreenOps as a diagnostic tool rather than a sustainability initiative. Engineering leaders will leave with a get-started-tomorrow framework covering:
– How to repurpose metrics your teams already track to expose architectural friction that uptime monitoring hides, and why waste is a leading indicator of instability.
– How emergency scaling decisions accumulate into long-term structural debt, and how to spot the patterns in your own post-incident reviews.
– A practical shift in post-incident questions, from “why did it break?” to “what did we leave behind?”, and how to make that shift stick culturally without adding process overhead.

The goal is not to slow your incident response down. It is to make sure the cleanup leaves the system stronger, greener, and more resilient before the next incident hits.