Uploaded August 2026 | Updated September 2026, 3 weeks ago
Don't miss out! Join us at our next KubeCon + CloudNativeCon events in Shanghai, China (8-9 September, 2026) and Salt Lake City, United States (Nov 9–12, 2026). Connect with our current graduated, incubating, and sandbox projects as the community gathers to further the education and advancement of cloud native computing. Learn more at kubecon.io
Beyond the DC Walls: Building a Nationwide GPU Fabric With KubeVirt and Wide-Area L2 - Kazuki Sato, NTT Docomo Business, Inc. & Jian Li, SK Telecom
Cloud-native AI infrastructure faces a "physical wall" of power and land limits. To solve this, NTT and SKT built a "Nationwide GPU Fabric", merging a 1,000km+ high-speed L2 network (IOWN APN) with Kubernetes and KubeVirt to operate distributed DCs (Sapporo to Fukuoka) as a single cluster.
We deep-dive into the architecture, performance, and future of GPU orchestration beyond a single DC.
- Distributed vs. Traditional DCs: Achievable workloads and trade-offs across a 1,000km+ fabric from Sapporo to Fukuoka.
- Fabric-Aware Topology Alignment: We demonstrate how our solution maps long-haul IOWN APN constraints into KubeVirt VMs, maintaining high-performance GPU peer communication (NCCL) across 1,000km as a single, virtualized cluster.
- DRA Feature: We share KubeVirt’s static limits and the potential of DRA for dynamic orchestration.
Attendees walk away with a concrete blueprint to shatter the "walls" of a single DC — a scalable solution for global urban AI infrastructure challenges.
Don't miss out! Join us at our next KubeCon + CloudNativeCon events in Shanghai, China (8-9 September, 2026) and Salt Lake City, United States (Nov 9–12, 2026). Connect with our current graduated, incubating, and sandbox projects as the community gathers to further the education and advancement of cloud native computing. Learn more at kubecon.io
Beyond the DC Walls: Building a Nationwide GPU Fabric With KubeVirt and Wide-Area L2 - Kazuki Sato, NTT Docomo Business, Inc. & Jian Li, SK Telecom
Cloud-native AI infrastructure faces a "physical wall" of power and land limits. To solve this, NTT and SKT built a "Nationwide GPU Fabric", merging a 1,000km+ high-speed L2 network (IOWN APN) with Kubernetes and KubeVirt to operate distributed DCs (Sapporo to Fukuoka) as a single cluster.
We deep-dive into the architecture, performance, and future of GPU orchestration beyond a single DC.
- Distributed vs. Traditional DCs: Achievable workloads and trade-offs across a 1,000km+ fabric from Sapporo to Fukuoka.
- Fabric-Aware Topology Alignment: We demonstrate how our solution maps long-haul IOWN APN constraints into KubeVirt VMs, maintaining high-performance GPU peer communication (NCCL) across 1,000km as a single, virtualized cluster.
- DRA Feature: We share KubeVirt’s static limits and the potential of DRA for dynamic orchestration.
Attendees walk away with a concrete blueprint to shatter the "walls" of a single DC — a scalable solution for global urban AI infrastructure challenges.










