1 paper
Songyu Xu, Xin Wang, Qiang Chen +6
Recent video generation models (VGMs) have made substantial progress in visual fidelity, yet their ability to follow long, compositional instructions remains insufficiently evaluat…