Person

Ming-Yu Liu

VP, NVIDIA Cosmos Lab

Artwork for How Physical AI Learns Across Language, Video and Action — Ming-Yu Liu
Machine Learning Street Talk (MLST)Ming-Yu Liu

How Physical AI Learns Across Language, Video and Action — Ming-Yu Liu

Ming-Yu Liu explains how NVIDIA’s Cosmos 3 combines language reasoning, video generation, and action generation into a single physical-AI stack. The episode focuses on world models, simulator-based policy verification, embodiment transfer from human video to robots, and the open release of Cosmos model sizes from Super to Edge.