Uploaded July 2025 | Updated September 2026, 2 weeks ago
AI workloads are growing fast and so are the operational challenges.
At Fully Connected 2025, Robert Nishihara, Co-founder of Anyscale, shared a clear path forward: a unified stack of Kubernetes, Ray, PyTorch, and vLLM.
His talk walks through live demos featuring autoscaling GPU pools and sub-10ms generation latency, showing what’s possible when these tools work together.
The full session is now available on demand. Worth a watch: lnkd.in/gUZHC8hJ
#shorts
AI workloads are growing fast and so are the operational challenges.
At Fully Connected 2025, Robert Nishihara, Co-founder of Anyscale, shared a clear path forward: a unified stack of Kubernetes, Ray, PyTorch, and vLLM.
His talk walks through live demos featuring autoscaling GPU pools and sub-10ms generation latency, showing what’s possible when these tools work together.
The full session is now available on demand. Worth a watch: lnkd.in/gUZHC8hJ
#shorts










