Pyversity with Thomas van Dongen - Weaviate Podcast #132! @Weaviate
Pyversity with Thomas van Dongen - Weaviate Podcast #132!  @Weaviate
Uploaded December 2025 | Updated September 2026, 5 hours ago
Hey everyone! Thanks so much for watching the 132nd episode of the Weaviate Podcast with Thomas van Dongen, head of AI engineering at Springer Nature! Thomas is also the creator of Pyversity, a fast, lightweight library for diversifying retrieval results! Retrieval systems often return highly similar items. Pyversity efficiently re-ranks these results to encourage diversity, surfacing items that remain relevant but less redundant. It implements several popular diversification strategies such as MMR, MSD, DPP, and Cover with a clear, unified API.

Check out Pyversity! github.com/Pringled/pyversity

Chapters:
0:00 Welcome Thomas!
0:30 An Introduction to Diversity in Vector Space
6:32 Diversification Strategies
15:42 Evaluating Diversity
21:36 Embedding Models for Diversity
27:25 LLMs for Diversity
33:20 The most representative set
36:50 Scientific Literature Mining
39:05 Thoughts on Chunking
42:35 Synthetic Data for Information Retrieval
46:00 Chatting with Scientific Papers
51:25 Future Directions and State of AI
Pyversity with Thomas van Dongen - Weaviate Podcast #132!7. AI Agents FrameworksAI Renaissance Berlin - AI BuzzwordsAgentic Topic Modeling with Maarten Grootendorst - Weaviate Podcast #126!Dive into Chunking Strategies for RAG with Zain 💚Agent Experience with Matt Biilmann, Sebastian Witalec, and Charles Pierse - Weaviate Podcast #116!AI Agent Optimization: Understanding the Pareto FrontierBooking.com and Weaviate with Başak Eskili - Weaviate Podcast #138!6. AI Agents Explained: Memory & Storage2. The Building Blocks of AI AgentsHow to pick the best LLM in 2025Agentic RAG: An End-to-End Open Source Framework
Weaviate vector database |

Pyversity with Thomas van Dongen - Weaviate Podcast #132!

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER