L12.4 Adam: Combining Adaptive Learning Rates and Momentum @SebastianRaschka
L12.4 Adam: Combining Adaptive Learning Rates and Momentum  @SebastianRaschka
Uploaded March 2021 | Updated September 2026, 2 weeks ago
Sebastian's books: sebastianraschka.com/books

Slides: sebastianraschka.com/pdf/lecture-notes/stat453ss21/L12_optim__slides.pdf

-------

This video is part of my Introduction of Deep Learning course.

Next video: youtu.be/c-SRPvK_zzs

The complete playlist: youtube.com/playlist?list=PLTKMiZHVd_2KJtIXOW0zFhFfBaJJilH51

A handy overview page with links to the materials: sebastianraschka.com/blog/2021/dl-course.html

-------

If you want to be notified about future videos, please consider subscribing to my channel: youtube.com/c/SebastianRaschka
L12.4 Adam: Combining Adaptive Learning Rates and MomentumL19.6 DistilBert Movie Review Classifier in PyTorch   Code ExampleL19.1 Sequence Generation with Word and Character RNNsL12.3 SGD with MomentumL2.1 Artificial NeuronsFinetuning Open-Source LLMsL9.4 Overfitting and UnderfittingL9.5.2 Custom DataLoaders in PyTorch  Code Example13.1 The Different Categories of Feature Selection (L13: Feature Selection)L13.0 Introduction to Convolutional Networks   Lecture OverviewL19.5.2.1 Some Popular Transformer Models: BERT, GPT, and BART   OverviewL5.7 Training an Adaptive Linear Neuron (Adaline)
Sebastian Raschka |

L12.4 Adam: Combining Adaptive Learning Rates and Momentum

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER