Lecture 71: [ScaleML Series] FlexOlmo: Open Language Models for Flexible Data Use @GPUMODE
Lecture 71: [ScaleML Series] FlexOlmo: Open Language Models for Flexible Data Use  @GPUMODE
Uploaded August 2025 | Updated September 2026, 2 hours ago
Day 1: Overview of the series and a long talk on FlexOlmo by Professor Sewon Min.

Full Schedule: scale-ml.org/bootcamp

The GPU MODE x Scale ML speaker series is a 5-day, online event hosted on the GPU MODE YouTube channel where top researchers in AI will talk about various architectural and system-level advances that are integrated into OpenAI’s frontier open-source model, GPT-OSS.

Each day will consist of ~2 hours of talks and discussions (around noon PST, may start at slightly different times each day so please check frequently), covering a different component of the evolving transformer stack—from quirks in the attention mechanism and positional encodings to quantization, MoEs, and custom GPU kernels.
Lecture 71: [ScaleML Series] FlexOlmo: Open Language Models for Flexible Data Use[Live] ScaleML Series Day 5 — GPU Programming for Foundation ModelsFactorio Learning EnvironmentLecture 27: gpu.cpp - Portable GPU compute using WebGPULecture 1 How to profile CUDA kernels in PyTorchLecture 82 Helion: A high-level DSL for ML kernelsLecture 21: Scan Algorithm Part 2Lecture 103: Fundamentals of CuTe Layout Algebra and Category-theoretic InterpretationConsumer GPU performanceLecture 46: Distributed GEMMLecture 2 Ch1-3 PMPP bookLivestream: FlashInfer
GPU MODE |

Lecture 71: [ScaleML Series] FlexOlmo: Open Language Models for Flexible Data Use

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER