1 paper
Linghui Shen, Mingyue Cui, Xingyi Yang
Visual world models (VWMs) synthesize interactive, action-conditioned rollouts from a single context image. However, it remains an open question how robust these models are to adve…