1 paper
Kevin Chuanpu Fu, Yongsen Zheng, Zee Kin Yeong +1
World models take multimodal inputs like text, photos, and diagrams to generate dynamic scenes in accordance with the laws of physics, thus opening a compelling application: fusing…