1 paper
Andrew Bond, Ilkin Umut Melanlioglu, Erkut Erdem +1
Modern visual world modeling systems increasingly rely on high-capacity architectures and large-scale data to produce plausible motion, yet they often fail to preserve underlying 3…