Uploaded May 2021 | Updated September 2026, 3 hours ago
Have you ever wondered how to do Machine Learning research? Or have you wanted to see how ML research works? Or how to get a research paper published? Then this is the series for you, welcome to Zero to Paper! This series is going to be about taking my Reinforcement Learning and NLP idea from the idea stage and turning it into a full fledged paper. This will entail doing research, reading papers, writing code, running experiments, and submitting to a conference.
I am going to use this channel to document the RL & NLP research experience. I am also going to be talking and collaborating with my subscribers, so if you are interested, join in!
If you are just getting into Reinforcement Learning and want to learn more, check out my playlist on RL theory vvv
youtube.com/watch?v=1OI0uuz9jkI&list=PL_49VD9KwQ_OML1Knh-Yb7FUFkhTLS0jL
If you want to take it a step further I have another whole series on Q-Learning vvv
youtube.com/watch?v=RiSCh4H-GY4&list=PL_49VD9KwQ_OpleGtJhWD24JDrrmXYXEQ
Timestamps
0:00 Overview
1:02 Why This Series
3:34 Research Topic
6:32 Plans
9:48 Wrap Up
Have you ever wondered how to do Machine Learning research? Or have you wanted to see how ML research works? Or how to get a research paper published? Then this is the series for you, welcome to Zero to Paper! This series is going to be about taking my Reinforcement Learning and NLP idea from the idea stage and turning it into a full fledged paper. This will entail doing research, reading papers, writing code, running experiments, and submitting to a conference.
I am going to use this channel to document the RL & NLP research experience. I am also going to be talking and collaborating with my subscribers, so if you are interested, join in!
If you are just getting into Reinforcement Learning and want to learn more, check out my playlist on RL theory vvv
youtube.com/watch?v=1OI0uuz9jkI&list=PL_49VD9KwQ_OML1Knh-Yb7FUFkhTLS0jL
If you want to take it a step further I have another whole series on Q-Learning vvv
youtube.com/watch?v=RiSCh4H-GY4&list=PL_49VD9KwQ_OpleGtJhWD24JDrrmXYXEQ
Timestamps
0:00 Overview
1:02 Why This Series
3:34 Research Topic
6:32 Plans
9:48 Wrap Up








![Self-Supervised RL - Learning Without Data [Zero to Paper]
Inverse Reinforcement Learning with Natural Language Goals (LangGoal IRL) offers a way to do sample-efficient IRL and a way to generalize using self-supervised learning. The paper is novel and is a step forward for general AI and ML algorithms. Though it has its cons, I think it is one of the better papers out there that cover RL, IRL, NLP, and generalization.
Zero to Paper playlist: https://www.youtube.com/playlist?list=PL_49VD9KwQ_ONxENRk11jFEI3_pqAwaug
Inverse Reinforcement Learning video: https://www.youtube.com/watch?v=qo355ALvLRI
Paper covered: https://arxiv.org/pdf/2008.06924.pdf Self-Supervised RL - Learning Without Data [Zero to Paper]](https://i.ytimg.com/vi/CDKsa06xU0o/mqdefault.jpg)

