Do World Models Make Better Robots? A Survey of Evaluation Benchmarks for Predictive Embodied Intelligence

arXiv:2609.29669v1 Announce Type: cross Abstract: Robot learning now advances along two tracks that rarely meet. On one side, direct Vision-Language-Action (VLA) policies map observations to actions and are scored by closed-loop task success. On the other, predictive and generative world models…

science

Sources