1 paper · 1 filter
Hao Wu, Yuqi Li, Yuan Gao +10
Existing robot video world models are typically trained with low-level objectives such as reconstruction and perceptual similarity, which are poorly aligned with the capabilities t…