Session
When the Cluster Bleeds: High-Severity Incident Diagnostics for EKS Workloads at Scale
When production Kubernetes clusters go down, clock cycles turn into dollar signs. Moving beyond basic kubectl logs and describe commands, this session dives deep into the trenches of real-world, high-severity EKS and AWS Batch incidents.
Led by a Senior Consultant who specializes in rapid incident restoration, we will break down a repeatable, calm-under-pressure diagnostic framework for containerized workloads. Attendees will walk away with architectural blueprints and operational runbooks covering:
- Tracing silent failures in multi-tenant EKS control planes and VPC CNI bottlenecks.
- Decoupling application bugs from underlying AWS infrastructure drift.
- Real-world post-mortems of root-cause analyses that saved customer trust and minimized downtime.
Cloud Native Benefit: Shifts the focus from theoretical Day-1 deployment to the harsh realities of Day-2 operations, helping platform engineers build resilient, highly observable infrastructure.
Please note that Sessionize is not responsible for the accuracy or validity of the data provided by speakers. If you suspect this profile to be fake or spam, please let us know.
Jump to top