Session

Keep Calm and Flink On: Disaster-Proofing Apache Flink

As real-time data processing becomes crucial for modern enterprises, building resilient and fault-tolerant Apache Flink applications is crucial. This presentation explores the foundational and advanced strategies for ensuring high availability, reliability, and disaster recovery in Flink environments.

We begin by demystifying Flink's checkpointing mechanism, its role in fast failure recovery, and common pitfalls to avoid. Beyond checkpointing, we delve into high-availability architectures, from active-passive to active-active setups, and their impact on system robustness. Finally, we present Bloomberg's solution for Kubernetes multi-cluster failover using Karmada, which not only allows it to failover workloads but also resume them with their state intact – further improving application resiliency.

By the end of this talk, attendees will gain actionable insights to build resilient, fault-tolerant Flink applications that can withstand real-world disruptions.

Niharika Sakuru

Software Engineer at Bloomberg LP

New York City, New York, United States

Actions

Please note that Sessionize is not responsible for the accuracy or validity of the data provided by speakers. If you suspect this profile to be fake or spam, please let us know.

Jump to top