Session

The Hidden Reliability Crisis: Why Observability Fails in Modern Distributed Systems

Modern distributed systems generate more telemetry than ever before — yet outages are becoming harder to detect, diagnose, and prevent.

This talk explores why traditional observability approaches are failing in large-scale cloud environments, especially with microservices, Kubernetes, and multi-cloud architectures.

We will examine:

The gap between metrics and real system behavior
Why logs and traces miss cascading failures
Alert fatigue and the illusion of visibility
The impact of complex service dependencies

We introduce a practical approach to improving reliability:

Designing signal-driven observability instead of data-driven
Mapping service dependencies to failure patterns
Improving incident response with better context and correlation

This session is based on real-world experience managing production systems at scale and focuses on actionable improvements rather than tooling.

Charit Upadhyay

Adobe, Senior Site Reliability Engineer

San Francisco, California, United States

Actions

Please note that Sessionize is not responsible for the accuracy or validity of the data provided by speakers. If you suspect this profile to be fake or spam, please let us know.

Jump to top