Scaling in Kubernetes Safely on On-Prem KaaS Across 1,300+ Clusters and 40,000+ Nodes - S. Yoshimura @cncf
Scaling in Kubernetes Safely on On-Prem KaaS Across 1,300+ Clusters and 40,000+ Nodes - S. Yoshimura  @cncf
Uploaded August 2026 | Updated September 2026, 2 weeks ago
Don't miss out! Join us at our next KubeCon + CloudNativeCon events in Shanghai, China (8-9 September, 2026) and Salt Lake City, United States (Nov 9–12, 2026). Connect with our current graduated, incubating, and sandbox projects as the community gathers to further the education and advancement of cloud native computing. Learn more at kubecon.io

Scaling in Kubernetes Safely on On-Prem KaaS Across 1,300+ Clusters and 40,000+ Nodes - Shota Yoshimura, LY Corporation

Many Kubernetes talks focus on scaling out to support growth, but large platforms also need safe ways to scale in. This is especially true for on-premises Kubernetes as a Service, where unused node capacity affects hardware procurement, rack capacity, and long-term operations.

This talk shares how we built a custom controller to consolidate underutilized worker nodes safely across 1,300+ clusters and 40,000+ nodes. At this scale, underutilized nodes were no longer isolated inefficiencies but a meaningful capacity problem. We needed a more conservative model than a general-purpose autoscaler could provide, so the controller uses observed node usage, seven-day peak metrics, minimum replica guarantees, and machine-group-level safety checks before taking action.

We will also cover safeguards including graceful shutdown expectations, DaemonSet termination ordering, and node deletion behavior designed to minimize user impact.
Scaling in Kubernetes Safely on On-Prem KaaS Across 1,300+ Clusters and 40,000+ Nodes - S. YoshimuraNavigating the Identity Abyss in the AI‑Native Era - Yuichi Nakamura, HitachiRoot Without Risk: A Decade-Long Quest for True Container Isolation - Sumir BrootaCNCF Ambassador Prithvi Raj on 14 Years of KubeCon + CloudNativeConLonghorn: Intro, Deep Dive and Q&A - Derek Su, SUSEThe Day is Made: Seema Saharan, CNCF AmbassadorConfiguring Argo CD Is Messy: Heres How We Can Fix It - Michael Crenshaw, IntuitChatLoopBackOff Episode 79: Capsule with Thomas and OliverSandbox for Agentic Application With Cloud Native Stack - Xu Wang, Ant Group & Yu Hu, JAPAN AIThe Next Evolution of Kubernetes: GPU-Centric Infrastructure for AI Workloads - Takao Indoh, FujitsuContainer Forensics for Kubernetes: Building an Evidence Pipeline With Open Source… J. Wu & P. GargPreferred Networks on Sharing Expensive AI Accelerators | KubeCon + CloudNativeCon Japan
CNCF [Cloud Native Computing Foundation] |

Scaling in Kubernetes Safely on On-Prem KaaS Across 1,300+ Clusters and 40,000+ Nodes - S. Yoshimura

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER