Combining next-token prediction and video diffusion in computer vision and robotics @MITCSAIL
Combining next-token prediction and video diffusion in computer vision and robotics  @MITCSAIL
Uploaded October 2024 | Updated September 2026, 2 weeks ago
Paper: https://boyuan.space/diffusion-forcing/
Authors: Russ Tedrake (MIT CSAIL), Vincent Sitzmann (MIT CSAIL), Boyuan Chen (MIT CSAIL), Diego Marti Monso (MIT CSAIL), Yilun Du (MIT CSAIL), & Max Simchowitz (MIT CSAIL)

Videographer: Mike Grimmett
Director: Rachel Gordon
PA: Alex Shipps
Combining next-token prediction and video diffusion in computer vision and roboticsMIT CSAIL Office Hours: Health | Episode 2Hybrid Drones: Drones that can hover like helicopters and fly like planesMIT talks with Love is Blind’s AI/data scientist Cameron HamiltonEQ-Radio: Emotion Recognition using Wireless SignalsColor-Changing 3D PrintablesMIT App Inventor: Using AI to democratize mobile techSensor skin gives robots a human touchDesign Your Own DronesInside the lab: MIT CSAILMIT CSAIL chats with Neil deGrasse Tyson54(ish) Questions w/an MIT AI & Health researcher
MIT CSAIL |

Combining next-token prediction and video diffusion in computer vision and robotics

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER