Post-Training Techniques: How LLMs Learn to Follow Instructions @stanfordonline
Post-Training Techniques: How LLMs Learn to Follow Instructions  @stanfordonline
Uploaded September 2026 | Updated September 2026, 2 weeks ago
CME295 Transformers and Large Language Models is open for enrollment until September 6. Learn more: https://online.stanford.edu/courses/cme295-transformers-and-large-language-models

A pretrained LLM is good at predicting text — but that's not the same as following instructions well. Stanford Adjunct Professors Shervine and Afshine Amidi walk through the techniques that close that gap: fine-tuning, reinforcement learning, preference optimization, verifier-guided training, and distillation.

CME296 Diffusion and Large Vision Models will be available in Spring 2027. Learn more about that course here: https://online.stanford.edu/courses/cme296-diffusion-and-large-vision-models

#PostTraining #LLM #Transformers #AI #MachineLearning #ReinforcementLearning #OnlineCourses
Post-Training Techniques: How LLMs Learn to Follow InstructionsStanford CS547 HCI Seminar | Spring 2026 | Promoting Agency in Human-AI InteractionStanford Online AI Programs Top Questions: When and How to Enroll in Online AI CoursesOverview: Stanford CME295 Transformers and Large Language ModelsStanford CS336 Language Modeling from Scratch | Spring 2026 | Lecture 2: PyTorch (einops)Stanford CS229 Machine Learning | Spring 2026 | Lecture 13: LLMs, Next-Word Prediction LossStanford CS221 | Autumn 2025 | Lecture 9: Policy GradientStanford CS221 | Autumn 2025 | Lecture 19: AI Supply ChainsStanford CS336 Language Modeling from Scratch | Spring 2026 | Lecture 3: ArchitecturesStanford CS229 Machine Learning | Spring 2026 | Lecture 6: Dataset Split, ML AdviceStop second-guessing high-stakes decisions.Stanford MS&E435 Economics of the AI Supercycle | Spring 2026 | Applications, AI in Life Sciences
Stanford Online |

Post-Training Techniques: How LLMs Learn to Follow Instructions

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER