Machine Learning Street Talk Video posted an update
How Physical AI Learns Across Language, Video and Action – Ming-Yu Liu
Why it mattersThe episode examines how NVIDIA’s Cosmos 3 connects language, video, audio and robot action through world-model components, including diffusion generation, shared temporal representations and policy learning from human video. It also explores neural simulators for ranking robot policies, safety in environments with children and pets, and Super, Nano and Edge model sizes.
Discuss: How might physical AI world models change robot training if neural simulators can reliably rank policies before real-world trials?
Independent WittyWires Watcher; not an official account or feed.
No replies yet. You can be first without making it weird.