Session
Beyond the DC Walls: Building a Nationwide GPU Fabric with KubeVirt and Wide-Area L2
Cloud-native AI infrastructure faces a "physical wall" of power and land limits. To solve this, NTT and SKT built a "Nationwide GPU Fabric", merging a 1,000km+ high-speed L2 network (IOWN APN) with Kubernetes and KubeVirt to operate distributed DCs (Sapporo to Fukuoka) as a single cluster.
We deep-dive into the architecture, performance, and future of GPU orchestration beyond a single DC.
- Distributed vs. Traditional DCs: Achievable workloads and trade-offs across a 1,000km+ fabric from Sapporo to Fukuoka.
- Fabric-Aware Topology Alignment: We demonstrate how our solution maps long-haul IOWN APN constraints into KubeVirt VMs, maintaining high-performance GPU peer communication (NCCL) across 1,000km as a single, virtualized cluster.
- DRA Feature: We share KubeVirt’s static limits and the potential of DRA for dynamic orchestration.
Attendees walk away with a concrete blueprint to shatter the "walls" of a single DC — a scalable solution for global urban AI infrastructure challenges.
Please note that Sessionize is not responsible for the accuracy or validity of the data provided by speakers. If you suspect this profile to be fake or spam, please let us know.
Jump to top