1 paper
Hao Wu, Yuqi Li, Yuan Gao +10
Existing robot video world models are typically trained with low-level objectives such as reconstruction and perceptual similarity, which are poorly aligned with the capabilities t…