Part II. Applying Agentic Reliability to Core Workflows
Five chapters. Each takes one familiar SRE workflow and re-renders it through the agentic lens. Each chapter opens by pointing at one beat of the DRAL loop.
Part I was about limits. Why human-paced reliability cannot scale to where systems have gone. What observability has to become before it can serve as a substrate for autonomous decision-making. How resilience changes when failure is treated as input rather than exception. How the loop that learns from failure is bounded by accountability rather than by ambition. By the end of those three chapters, one claim should be settled. Agentic reliability is not something you add on top of existing practices. It only works when it is deliberately designed into the system.
Part II is where that design becomes operational. Each chapter takes a familiar SRE workflow that the reader already runs every week, and shows what changes when the architecture from Chapter 4 is brought to bear on it. Chapter 5 ...
Become an O’Reilly member and get unlimited access to this title plus top books and audiobooks from O’Reilly and nearly 200 top publishers, thousands of courses curated by job role, 150+ live events each month,
and much more.
Read now
Unlock full access