L11.2 How BatchNorm Works @SebastianRaschka
L11.2 How BatchNorm Works  @SebastianRaschka
Uploaded March 2021 | Updated September 2026, 2 weeks ago
Sebastian's books: sebastianraschka.com/books

Slides: sebastianraschka.com/pdf/lecture-notes/stat453ss21/L11_norm-and-init__slides.pdf

BatchNorm paper: Ioffe, S., & Szegedy, C. (2015). Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift. In International Conference on Machine Learning (pp. 448-456). http://proceedings.mlr.press/v37/ioffe15.html

-------

This video is part of my Introduction of Deep Learning course.

Next video: youtu.be/8AUDn7iF2DY

The complete playlist: youtube.com/playlist?list=PLTKMiZHVd_2KJtIXOW0zFhFfBaJJilH51

A handy overview page with links to the materials: sebastianraschka.com/blog/2021/dl-course.html

-------

If you want to be notified about future videos, please consider subscribing to my channel: youtube.com/c/SebastianRaschka
L11.2 How BatchNorm WorksL4.0 Linear Algebra for Deep Learning   Lecture OverviewL4.3 Vectors, Matrices, and BroadcastingL5.2 Relation Between Perceptron and Linear RegressionL10.5.3 (Optional) Dropout Ensemble InterpretationL8.7.1 OneHot Encoding and Multi-category Cross EntropyL4.4 Notational Conventions for Neural NetworksBuild an LLM from Scratch 7: Instruction FinetuningBuild an LLM from Scratch 6: Finetuning for ClassificationL8.7.2 OneHot Encoding and Multi-category Cross Entropy   Code ExampleL15.2 Sequence Modeling with RNNsL18.6: A DCGAN for Generating Face Images in PyTorch   Code Example
Sebastian Raschka |

L11.2 How BatchNorm Works

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER