Vision language action models

Why Robotics Still Isn't Solved - But Could Be Soon | YC Paper Club

Why Robotics Still Isn't Solved - But Could Be Soon | YC Paper Club

This Paper Club delves into the current state of robotics, addressing roadblocks like the sim-to-real gap and embodiment drift. Speakers present advancements in multi-scale memory for long-horizon tasks, self-supervised embodied reasoning, zero-shot dexterous manipulation via massive simulation, and the economic imperative of teleoperation-first robotics companies, concluding with optimizations for efficient, real-time World Action Models.

Robotics: why now? - Quan Vuong and Jost Tobias Springberg, Physical Intelligence

Robotics: why now? - Quan Vuong and Jost Tobias Springberg, Physical Intelligence

Quan Vuong and Jost Tobias Springenberg from Physical Intelligence (PI) discuss their mission to create a universal model for controlling any robot. They detail their approach, which centers on Vision-Language-Action (VLA) models, a purpose-built data engine for scaled data collection, and the evolution of their models toward open-world generalization.