MUVERA with Rajesh Jayaram and Roberto Esposito - Weaviate Podcast #123! @Weaviate
MUVERA with Rajesh Jayaram and Roberto Esposito - Weaviate Podcast #123!  @Weaviate
Uploaded May 2025 | Updated September 2026, 1 hour ago
Multi-vector retrieval offers richer, more nuanced search, but often comes with a significant cost in storage and computational overhead. How can we harness the power of multi-vector representations without breaking the bank? Rajesh Jayaram, the first author of the groundbreaking MUVERA algorithm from Google, and Roberto Esposito from Weaviate, who spearheaded its implementation, reveal how MUVERA tackles this critical challenge.

Dive deep into MUVERA, a novel compression technique specifically designed for multi-vector retrieval. Rajesh and Roberto explain how it leverages contextualized token embeddings and innovative fixed dimensional encodings to dramatically reduce storage requirements while maintaining high retrieval accuracy. Discover the intricacies of quantization within MUVERA, the interpretability benefits of this approach, and how LSH clustering can play a role in topic modeling with these compressed representations.

This conversation explores the core mechanics of efficient multi-vector retrieval, the challenges of benchmarking these advanced systems, and the evolving landscape of vector database schemas designed to handle such complex data. Rajesh and Roberto also share their insights on the future directions in artificial intelligence where efficient, high-dimensional data representation is paramount.
Whether you're an AI researcher grappling with the scalability of vector search, an engineer building advanced retrieval systems, or fascinated by the cutting edge of information retrieval and AI frameworks, this episode delivers unparalleled insights directly from the source. You'll gain a fundamental understanding of MUVERA, practical considerations for its application in making multi-vector retrieval feasible, and a clear view of future directions in AI.

Links:
MUVERA: arxiv.org/abs/2405.19504
CRISP: arxiv.org/pdf/2505.11471
ColBERT: arxiv.org/abs/2004.12832
ColPali: arxiv.org/abs/2407.01449
Multi-Vector Embeddings with Weaviate (Tutorial): weaviate.io/developers/weaviate/tutorials/multi-vector-embeddings

Chapters
0:00 Welcome Rajesh and Roberto
2:10 Intro to Multi-Vector Retrieval
7:53 Contextualized Token Embeddings
14:04 Interpretability of Multi-Vector Retrieval
17:46 Multi-Vector Storage Cost
20:10 MUVERA Deep Dive
32:30 Fixed Dimensional Encodings
55:32 Quantization in MUVERA
1:00:04 LSH Clustering for Topic Modeling
1:02:44 Benchmarks for Multi-Vector Retrieval
1:06:33 Future of Vector Database Schemas
1:09:28 Directions for the Future of AI
MUVERA with Rajesh Jayaram and Roberto Esposito - Weaviate Podcast #123!Booking.coms Partner-to-Guest Messaging AgentRAGKit with Kyle Davis - Weaviate Podcast #93!Box AI with Ben Kus and Bob van Luijt - Weaviate Podcast #120!Subjectivity in AI with Dan Shipper: AI-Native Databases #4Founding Weaviate with Bob van Luijt and Etienne Dilocker - Weaviate Podcast #140!Scaling Pandas with Devin Petersohn - Weaviate Podcast #101!Generate multimodal datasets for your demos, Proof of Concept (PoC), or project with Generator9000!Think Deepseek-R1 was trained like other top LLMs? Think again.AI in Education with Rose E. Wang - Weaviate Podcast #106!Compound AI Systems with Philip Kiely - Weaviate Podcast #105!THIS is how you build agentic RAG systems that work
Weaviate vector database |

MUVERA with Rajesh Jayaram and Roberto Esposito - Weaviate Podcast #123!

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER