WAMs and VLAs for Robot Learning | Cosmos Labs @NVIDIADeveloper
WAMs and VLAs for Robot Learning | Cosmos Labs  @NVIDIADeveloper
Uploaded August 2026 | Updated September 2026, 2 weeks ago
Building capable robot foundation models requires more than one approach. Vision-language-action (VLA) models and World Action Models (WAMs) - each offer different strengths for robot learning, planning, and control.

In this livestream, NVIDIA researchers Danfei Xu US Moritz Benno Reuss CH Thomas Tian US will explore what World Action Models work, how they compare with modern Vision-Language-Action (VLA) models, and why many next-generation robot foundation models combine both approaches.

You'll also see how NVIDIA Cosmos 3 treats action as a native modality, enabling robots to predict future observations while generating actions. We'll walk through how Cosmos 3 provides an open workflow for post-training, evaluation, and deployment using open models, datasets, recipes, and serving resources.

Whether you're building manipulation policies, training robot foundation models, or exploring embodied AI, this livestream provides a practical understanding of where World Action Models fit into the future of robotics.

What You'll Learn:
- What World Action Models are and how they differ from and complement Vision-Language-Action models
- The strengths, limitations, and tradeoffs of WAMs, VLAs, and emerging hybrid approaches
- How Cosmos3-Nano-Policy-DROID imagines future observations while generating robot actions
- How NVIDIA Cosmos 3 represents action as a native modality alongside video
- How to evaluate, post-train, and adapt Cosmos 3 using open models, datasets, recipes, and serving resources

Have questions about how to post-train and deploy NVIDIA Cosmos 3? Drop them live — the NVIDIA team will answer them in real time.

Access more NVIDIA Cosmos developer resources and join our developer community:

📄 Read the Technical Blog → nvda.ws/4c6kK3R
⬇️ Download Cosmos on Hugging Face → huggingface.co/collections/nvidia/cosmos3
📚 Explore Models & Datasets on GitHub → github.com/nvidia/Cosmos
👥 Join the Cosmos Community → discord.com/invite/nvidiaomniverse
WAMs and VLAs for Robot Learning | Cosmos LabsRun NVIDIA Nemotron 3.5 Lightning on DGX SparkHow LangChain and NVIDIA Help Developers Build AI AgentsDev Community Live: NYC Spark Hack WinnersWhats Next in Generative AI | Brad Lightcap, OpenAI COO and Manuvir Das, NVIDIA VP | NVIDIA GTC24Dev Community Live: Run OpenClaw Agents Safely - Cloud AI, Zero Data ExposureAnnouncing the Winners of the NVIDIA Generative AI on RTX Developer ContestHow to Connect Two DGX Sparks with NVIDIA SyncNVIDIA GTC 2026 Developer Community Livestream: OpenClaw, Physical AI & RoboticsAsk the Experts: Nemotron 3 Nano Omni | Nemotron LabsBuild a RAG Agent with NVIDIA Nemotron: A Developers Guide to Agentic AIBuild a Claw: NVIDIA NemoClaw on DGX Spark | Nemotron Labs
NVIDIA Developer |

WAMs and VLAs for Robot Learning | Cosmos Labs

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER