Uploaded October 2024 | Updated September 2026, 2 weeks ago
Phaidra is reshaping the landscape of industrial and data center optimization with AI-driven controls. In this illuminating session, Brandon Hernandez and Jerry Luo unveil Phaidra's innovative approach to building a multi-tenant data processing platform on Ray for Reinforcement Learning agents.
The presenters delve into the architecture of their Ray-based platform, which forms the backbone of their data ingestion, transformation, and feature processing operations. They explore how this system has accelerated customer onboarding and streamlined model deployment in production environments. Hernandez and Luo also address the challenges they faced in achieving isolation and efficient resource utilization for multi-tenancy, offering valuable insights for organizations looking to scale their AI workloads while maintaining performance and security.
--
Interested in more?
- Watch the full Day 1 Keynote: youtu.be/jwZHJthQvXo
- Watch the full Day 2 Keynote youtu.be/Lury2ad6KG8
- Check out the Ray Summmit Breakout sessions youtube.com/playlist?list=PLzTswPQNepXntmT8jr9WaNfqQ60QwW7-U&si=qPw-_SxT9lVmbRGE
--
đź”— Connect with us:
- Subscribe to our YouTube channel: youtube.com/@anyscale
- Twitter: https://x.com/anyscalecompute
- LinkedIn: linkedin.com/company/joinanyscale
- Website: anyscale.com
Phaidra is reshaping the landscape of industrial and data center optimization with AI-driven controls. In this illuminating session, Brandon Hernandez and Jerry Luo unveil Phaidra's innovative approach to building a multi-tenant data processing platform on Ray for Reinforcement Learning agents.
The presenters delve into the architecture of their Ray-based platform, which forms the backbone of their data ingestion, transformation, and feature processing operations. They explore how this system has accelerated customer onboarding and streamlined model deployment in production environments. Hernandez and Luo also address the challenges they faced in achieving isolation and efficient resource utilization for multi-tenancy, offering valuable insights for organizations looking to scale their AI workloads while maintaining performance and security.
--
Interested in more?
- Watch the full Day 1 Keynote: youtu.be/jwZHJthQvXo
- Watch the full Day 2 Keynote youtu.be/Lury2ad6KG8
- Check out the Ray Summmit Breakout sessions youtube.com/playlist?list=PLzTswPQNepXntmT8jr9WaNfqQ60QwW7-U&si=qPw-_SxT9lVmbRGE
--
đź”— Connect with us:
- Subscribe to our YouTube channel: youtube.com/@anyscale
- Twitter: https://x.com/anyscalecompute
- LinkedIn: linkedin.com/company/joinanyscale
- Website: anyscale.com






![[Ray Meetup] Ray + vLLM in Action: Lessons from Pinterest and Large Scale Distributed Inference
Listen in to our Ray Meetup where we explored batch inference at scale with Ray and vLLM! Learn how Pinterest scales batch inference using Ray, and get a first look at Anyscale’s latest tools—Ray Serve and Data LLM—for orchestrating large-scale LLM inference. We’ll cover topics like batch inference, prefill-decode disaggregation, DP/EP parallelism, and custom request routing.
Speakers:
Chia-Wei Chen, Software Engineer, ML Training Infra, Pinterest
Kourosh Hakhamaneshi, AI Lead, Anyscale [Ray Meetup] Ray + vLLM in Action: Lessons from Pinterest and Large Scale Distributed Inference](https://i.ytimg.com/vi/HDSy09hrm2I/mqdefault.jpg)



