L19.4.2 Self-Attention and Scaled Dot-Product Attention @SebastianRaschka
L19.4.2 Self-Attention and Scaled Dot-Product Attention  @SebastianRaschka
Uploaded May 2021 | Updated September 2026, 2 weeks ago
Sebastian's books: sebastianraschka.com/books

Slides: sebastianraschka.com/pdf/lecture-notes/stat453ss21/L19_seq2seq_rnn-transformers__slides.pdf

-------

This video is part of my Introduction of Deep Learning course.

Next video: youtu.be/A1eUVxscNq8

The complete playlist: youtube.com/playlist?list=PLTKMiZHVd_2KJtIXOW0zFhFfBaJJilH51

A handy overview page with links to the materials: sebastianraschka.com/blog/2021/dl-course.html

-------

If you want to be notified about future videos, please consider subscribing to my channel: youtube.com/c/SebastianRaschka
L19.4.2 Self-Attention and Scaled Dot-Product AttentionL15.4 Backpropagation Through Time Overview13.4.4 Sequential Feature Selection (L13: Feature Selection)L8.0 Logistic Regression   Lecture OverviewL14.0: Convolutional Neural Networks Architectures   Lecture OverviewL19.5.2.6 BART:  Combining Bidirectional and Auto-Regressive TransformersL5.3 An Iterative Training Algorithm for Linear RegressionDeep Learning News #3, Feb 13 2021Twitter Posts Political Ideology Classification (Student Presentation, Group 15)Build an LLM from Scratch 2: Working with text dataL16.4 A Convolutional Autoencoder in PyTorch   Code ExampleL11.2 How BatchNorm Works
Sebastian Raschka |

L19.4.2 Self-Attention and Scaled Dot-Product Attention

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER