1 paper
Qikang Zhang, Yingjie Lei, Wei Liu +1
Video generation models have been used as a robot policy to predict the future states of executing a task conditioned on task description and observation. Previous works ignore the…