Uploaded February 2025 | Updated September 2026, 1 hour ago
Hey everyone! Thank you so much for watching the 114th episode of the Weaviate Podcast featuring Amanpreet Singh, Co-Founder and CTO of Contextual AI! Contextual AI is at the forefront of production-grade RAG agents! I learned so much from this conversation! We began by discussing the vision of RAG 2.0, jointly optimizing generative and retrieval models! This then lead us to discuss Agentic RAG and how the RAG 2.0 roadmap is evolving with emerging perspectives on tool use. Amanpreet continues to further motivate the importance of continual learning of the model and the prompt / few-shot examples -- discussing the limits of prompt engineering. Personally I have to admit I think I have been a bit too bullish on only tuning instructions / examples, Amanpreet made an excellent case for updating the weights of the models as well -- citing issues such as parametric knowledge conflicts, and later on discussing how Mechanistic Interpretability is used to audit models and their updates in enterprise settings. We then discussed Contextual AI's LMUnit for evaluating these systems. This then lead us into my favorite part of the podcast, a deep dive into RL algorithms for LLMs. I highly recommend checking out the links below to learn more about Contextual's innovations on APO and KTO! We then discuss the importance of domain specific data, Mechanistic Interpretability, return to another question on RAG 2.0, and conclude with Amanpreet's most exciting future directions for AI! I hope you enjoy the podcast!
Learn more about Contextual AI! - contextual.ai
Contextual AI Platform: contextual.ai/blog/contextual-ai-platform-generally-available
Links:
Anchored Preference Optimization: arxiv.org/abs/2408.06266
KTO: Model Alignment as Prospect Theoretic Optimization: arxiv.org/pdf/2402.01306
RETRO: arxiv.org/abs/2112.04426
Fusion-in-Decoder: arxiv.org/pdf/2007.01282
RAG: arxiv.org/pdf/2005.11401
LMUnit: contextual.ai/lmunit
Chapters
0:00 Welcome Amanpreet!
0:28 Founding Story and RAG 2.0
5:00 Agentic RAG
10:25 The Limits of Prompt Engineering
19:55 LMUnit Evals
28:35 RL for LLMs
40:00 Model Specialization
46:12 Mechanistic Interpretability
49:10 Back to RAG 2.0
55:15 Future Directions for AI
Hey everyone! Thank you so much for watching the 114th episode of the Weaviate Podcast featuring Amanpreet Singh, Co-Founder and CTO of Contextual AI! Contextual AI is at the forefront of production-grade RAG agents! I learned so much from this conversation! We began by discussing the vision of RAG 2.0, jointly optimizing generative and retrieval models! This then lead us to discuss Agentic RAG and how the RAG 2.0 roadmap is evolving with emerging perspectives on tool use. Amanpreet continues to further motivate the importance of continual learning of the model and the prompt / few-shot examples -- discussing the limits of prompt engineering. Personally I have to admit I think I have been a bit too bullish on only tuning instructions / examples, Amanpreet made an excellent case for updating the weights of the models as well -- citing issues such as parametric knowledge conflicts, and later on discussing how Mechanistic Interpretability is used to audit models and their updates in enterprise settings. We then discussed Contextual AI's LMUnit for evaluating these systems. This then lead us into my favorite part of the podcast, a deep dive into RL algorithms for LLMs. I highly recommend checking out the links below to learn more about Contextual's innovations on APO and KTO! We then discuss the importance of domain specific data, Mechanistic Interpretability, return to another question on RAG 2.0, and conclude with Amanpreet's most exciting future directions for AI! I hope you enjoy the podcast!
Learn more about Contextual AI! - contextual.ai
Contextual AI Platform: contextual.ai/blog/contextual-ai-platform-generally-available
Links:
Anchored Preference Optimization: arxiv.org/abs/2408.06266
KTO: Model Alignment as Prospect Theoretic Optimization: arxiv.org/pdf/2402.01306
RETRO: arxiv.org/abs/2112.04426
Fusion-in-Decoder: arxiv.org/pdf/2007.01282
RAG: arxiv.org/pdf/2005.11401
LMUnit: contextual.ai/lmunit
Chapters
0:00 Welcome Amanpreet!
0:28 Founding Story and RAG 2.0
5:00 Agentic RAG
10:25 The Limits of Prompt Engineering
19:55 LMUnit Evals
28:35 RL for LLMs
40:00 Model Specialization
46:12 Mechanistic Interpretability
49:10 Back to RAG 2.0
55:15 Future Directions for AI










