Uploaded April 2026 | Updated September 2026, 2 weeks ago
Please subscribe to our YouTube channel @ youtube.com/@DevoxxForever
Subscribe to LinkedIn @ linkedin.com/company/voxxed-days-amsterdam
Follow us on Twitter @ twitter.com/voxxedamsterdam
Welcome to Voxxed Days Amsterdam! We’re thrilled to have you on board for what promises to be an inspiring event.
In short an opening session, we’ll set the stage for an great day of learning, networking, and tech-driven insights. Here’s what’s coming up:
What to expect – A quick tour of the agenda and highlights you won’t want to miss.
How to make the most of your day – Tips on engaging with speakers, connecting with fellow attendees, and exploring the venue.
Why we’re here – The spirit of Voxxed Days and how we’re building a thriving tech community together.
So grab your coffee, settle in, and get ready, because Voxxed Days Amsterdam is officially kicking off!
Please subscribe to our YouTube channel @ youtube.com/@DevoxxForever
Subscribe to LinkedIn @ linkedin.com/company/voxxed-days-amsterdam
Follow us on Twitter @ twitter.com/voxxedamsterdam
Welcome to Voxxed Days Amsterdam! We’re thrilled to have you on board for what promises to be an inspiring event.
In short an opening session, we’ll set the stage for an great day of learning, networking, and tech-driven insights. Here’s what’s coming up:
What to expect – A quick tour of the agenda and highlights you won’t want to miss.
How to make the most of your day – Tips on engaging with speakers, connecting with fellow attendees, and exploring the venue.
Why we’re here – The spirit of Voxxed Days and how we’re building a thriving tech community together.
So grab your coffee, settle in, and get ready, because Voxxed Days Amsterdam is officially kicking off!


![[VDBUH2026] Abdel Sghiouar - Optimizing LLM Inference for the Rest of Us
Not every organization operates with the hyperscale resources of Anthropic, Google, or OpenAI. For the majority of businesses integrating Large Language Models (LLMs) into their critical paths, the high costs and scarcity of GPU/TPU accelerators present a significant challenge. Striking the balance between performance, availability, scalability, and cost-efficiency is a must.
While Kubernetes is a ubiquitous runtime for modern workloads, deploying LLM inference effectively demands a specialized approach. This session dives deep into practical strategies for optimizing your Kubernetes clusters and LLM Inference workloads to run efficiently and cost effectively. We will explore:
– Container and Model Optimization
– Accelerator Management
– Data & Storage
– Network & Load Balancing
– Observability
Attendees will leave with practical techniques for maximizing cost/performance for LLM inference for their AI-powered applications on Kubernetes. [VDBUH2026] Abdel Sghiouar - Optimizing LLM Inference for the Rest of Us](https://i.ytimg.com/vi/G58PbxBXC8c/mqdefault.jpg)







