How to Apply LLMs on Audio Recordings with Multiple Speakers @AssemblyAI
How to Apply LLMs on Audio Recordings with Multiple Speakers  @AssemblyAI
Uploaded March 2024 | Updated September 2026, 3 weeks ago
Get AssemblyAI API key for this tutorial: assemblyai.com/?utm_source=youtube&utm_medium=referral&utm_campaign=yt_mis_60

LLMs work wonders on text data but if you want to use audio or video files instead of text, things get a bit trickier. An easy solution is to transcribe the audio or video files. This would work but you will lose valuable information, especially in multi-speaker situations, like how many people were speaking and who said what.

In this video, we’ll learn how to build a RAG application in 10 minutes that can take multiple speakers into account when answering a question.

Colab notebook: github.com/deepset-ai/haystack-cookbook/blob/main/notebooks/using_speaker_diarization_with_assemblyai.ipynb

AssemblyAI-Haystack Integration docs: assemblyai.com/docs/integrations/haystack/?utm_source=youtube&utm_medium=referral&utm_campaign=yt_mis_60

Blog post of this video: haystack.deepset.ai/blog/level-up-rag-with-speaker-diarization

00:00 Introduction
00:32 Effect of Speaker Labels
01:49 Libraries and example files
04:43 Transcription Pipeline
07:52 RAG Application
10:34 Results
11:52 Try it out yourself!

▬▬▬▬▬▬▬▬▬▬▬▬ CONNECT ▬▬▬▬▬▬▬▬▬▬▬▬

🖥️ Website: assemblyai.com/?utm_source=youtube&utm_medium=referral&utm_campaign=yt_mis_60
🐦 Twitter: twitter.com/AssemblyAI
👽 Reddit: reddit.com/r/assemblyai
▶️ Subscribe: youtube.com/c/AssemblyAI?sub_confirmation=1
🔥 We're hiring! Check our open roles: assemblyai.com/careers

▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬▬

#MachineLearning #DeepLearning
How to Apply LLMs on Audio Recordings with Multiple SpeakersBuild an AI medical note-taker with one APIWe Tested our Speech AI a DJ Booth at the Miami GPGraph Neural Networks – 7 Applications  #machinelearning #deeplearningai #graphml #deeplearningRetrieval Augmented Generation explained on a high-level #RAG #retrievalaugmentedgeneration🎙️Build An AI Voice Agent With DeepSeek R1 (Python)Build an AI Voice Agent in 10 MinutesConversation Intelligence starts with AssemblyAIWhy “Who Said What” Matters (EdgeTier) 🗣️Medical Mode Technical ShowcaseCustom Formatting in AssemblyAI: Control Your Transcription OutputAI That Works for Customers: Synthesia CEO Victor Riparbelli Explains
AssemblyAI |

How to Apply LLMs on Audio Recordings with Multiple Speakers

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER