A friendly introduction to deep reinforcement learning, Q-networks and policy gradients @SerranoAcademy
A friendly introduction to deep reinforcement learning, Q-networks and policy gradients  @SerranoAcademy
Uploaded May 2021 | Updated September 2026, 2 weeks ago
A video about reinforcement learning, Q-networks, and policy gradients, explained in a friendly tone with examples and figures.

Introduction to neural networks: youtube.com/watch?v=BR9h47Jtqyw

Introduction: (0:00)
Markov decision processes (MDP): (1:09)
Rewards: (5:39)
Discount factor: (8:51)
Bellman equation: (10:48)
Solving the Bellman equation: (12:43)
Deterministic vs stochastic processes: (16:29)
Neural networks: (19:15)
Value neural networks: (21:44)
Policy neural networks: (25:44)
Training the policy neural network: (30:46)
Conclusion: (34:53)

Announcement: Book by Luis Serrano! Grokking Machine Learning. bit.ly/grokkingML
40% discount code: serranoyt
A friendly introduction to deep reinforcement learning, Q-networks and policy gradientsLatent Dirichlet Allocation (Part 1 of 2)Why is DeepSeek so good?The Discrete Fourier TransformNewtons method for approximating zeros of polynomials - Math for ML with Deeplearning.aiProximal Policy Optimization (PPO) - How to train Large Language ModelsMath and OCD - My story with the Thue-Morse sequenceThank you for 100K subscribers! I’m planning tons of new content coming soon, so excited!A friendly introduction to Recurrent Neural NetworksThe math behind Attention: Keys, Queries, and Values matricesWill AI help us, or make us dependent? - A Tale of Two CitiesThe covariance matrix
Luis Serrano Academy |

A friendly introduction to deep reinforcement learning, Q-networks and policy gradients

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER