AI Dev 25 x NYC | Nitin Kanukolanu: Semantic Caching for LLM Applications @Deeplearningai
AI Dev 25 x NYC | Nitin Kanukolanu: Semantic Caching for LLM Applications  @Deeplearningai
Uploaded December 2025 | Updated September 2026, 2 weeks ago
Nitin Kanukolanu, Applied AI Engineer at Redis, focused on semantic caching during AI Dev 25 x NYC.

As LLMs drive the next wave of applications, compute bottlenecks are becoming a critical challenge. Semantic caching has emerged as a practical strategy to cut costs, reduce latency, and improve consistency in agentic systems. This session covered real-world use cases, explained how semantic caching works, highlighted what to measure in production, and shared strategies for boosting both performance and quality at scale.

Take our course with Redis on this topic: deeplearning.ai/short-courses/semantic-caching-for-ai-agents

--------------

Join us at AI Dev 26 x San Francisco! Tickets: ai-dev.deeplearning.ai
AI Dev 25 x NYC | Nitin Kanukolanu: Semantic Caching for LLM ApplicationsAI Dev 25 | Bilge Yücel: Building and Deploying Agentic Workflows with HaystackThe big brain for the big work...  #aiagents #buildinpublic #deeplearning #jetbrains @JetBrainsTVA new course on Retrieval Augmented Generation (RAG) is here!Learn to post-train LLMs in this free courseIs EU losing the AI race?AI Dev 26 x SF | Paige Bailey: Research to RealityLearn to build effective Agentic AI systems with Andrew NgIs vibe coding real coding?AI Dev 26 x SF | Diamond Bishop: The Next 100 Agents. Building the Agent Native OfficeEnroll in DeepLearning.AIs Data Analytics Professional Certificate!Take back control of your AI coding workflow
DeepLearningAI |

AI Dev 25 x NYC | Nitin Kanukolanu: Semantic Caching for LLM Applications

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER