It's Never the Thing That Broke: Learning From Incidents at (AI) Speed

QCon San Francisco 2026

Session

It's Never the Thing That Broke: Learning From Incidents at (AI) Speed

Monday Nov 16 / 05:05PM PST, Seacliff ABC at Hyatt Regency, San Francisco

Register

$2,835, Conference (3 days)
Current pricing ends September 8th

Abstract

When something breaks, our first instinct is to focus on the thing that failed. But in complex systems, the failure is usually the result of many things going right, changing, and interacting in ways we didn't anticipate. Learning From Incidents has taught us to look beyond the failure itself; to understand the decisions, adaptations, and conditions that made it possible. But what happens when the system keeps changing while we're trying to understand it?

AI is making this problem harder. Models are updated, agents get new tools and permissions, prompts change, and system behavior has shifted. By the time we've investigated an incident and acted on what we learned, the system may already be different.

This talk looks at what we can carry forward from years of Learning From Incidents practice, where those approaches start to strain, and what it means to build a learning process that can keep up with systems that don't sit still.

Register

$2,835, Conference (3 days). Current pricing ends September 8th. All pass options.

76% senior dev or higher
1:11 speaker ratio
60+ practitioners

QCon San Francisco 2026 is a three day conference for senior software engineers, architects and team leads. An international program committee of working engineers selects every session. Patterns and practices, not products and pitches.

Share

From the same track

Monday 16 November

10:35 Seacliff ABC Session The Freeze Paradox: Why Stopping Deployments Doesn't Stop Failures Prachi Jain, Sandhya Narayan 11:45 Seacliff ABC Session Saturation: How Your Software Will Fail at Scale Lorin Hochstein Staff Software Engineer @Airbnb, Writes @surfingcomplexity.blog, Previously @Netflix and Member of the Resilience in Software Foundation 13:35 Seacliff D Unconference Unconference: Resilience Engineering 14:45 Seacliff ABC Session Adapt or Drift: Resilience Engineering When AI Moves the Operating Point Andrew Hatch Engineering Leader and SRE Manager @Cisco ThousandEyes, With 25+ Years Building Software, Operations, SRE, and Platform Teams Across Australia, India, and the United States 15:55 Seacliff ABC Session Engineering Boundaries for Outages You Can't Prevent Em Ruppe Technical Incident Commander @Chime, Previously @SendGrid and @Twilio, and Product and Training @Jeli.io 17:05 Seacliff ABC Session It's Never the Thing That Broke: Learning From Incidents at (AI) Speed Vanessa Huerta Granda Resilience Engineering Manager @Enova, Co-Author of the Howie Guide on Post Incident Analysis, Board Member for the Resilience in Software Foundation

Current pricing ends September 8th
$2,835, Conference (3 days)

Register