1 paper · 1 filter
Mining Tan, Yinuo Wang, Ziqi Zhou +6
Unified models for visual understanding and generation have made rapid progress, yet they still lack the ability to understand and manipulate the spatial states of object instances…