Agents are distributed systems — and almost nobody treats them that way
A multi-agent system inherits every classic distributed-systems problem — partial failure, shared state, retries that duplicate actions, ordering, consistency — with one aggravating factor: every node is non-deterministic. The talk applies to agents the toolkit engineering already masters (idempotency, sagas, circuit breakers, observability, SLOs and error budgets) and proposes what SRE for agents would look like: how to design, measure and operate systems that may answer the same question differently.
What is the SLO of a system that can give two different answers to the same input?
Ideal for: platform, SRE, architecture and AI engineering teams.