167 citations · 263 across the 31 of their papers we have counts for
4 papers · 2 filters
Unified Discrete Diffusion for Simultaneous Vision-Language Generation
Minghui Hu, Chuanxia Zheng, Heliang Zheng +5
The recently developed discrete diffusion models perform extraordinarily well in the text-to-image task, showing significant promise for handling the multi-modality signals. In thi…
Entry-Flipped Transformer for Inference and Prediction of Participant Behavior
Bo Hu, Tat-Jen Cham
Some group activities, such as team sports and choreographed dances, involve closely coupled interaction between participants. Here we investigate the tasks of inferring and predic…
High-Quality Pluralistic Image Completion via Code Shared VQGAN
Chuanxia Zheng, Guoxian Song, Tat-Jen Cham +3
PICNet pioneered the generation of multiple and diverse results for image completion task, but it required a careful balance between loss (diversity) and reconstruct…
Sem2NeRF: Converting Single-View Semantic Masks to Neural Radiance Fields
Yuedong Chen, Qianyi Wu, Chuanxia Zheng +2
Image translation and manipulation have gain increasing attention along with the rapid development of deep generative models. Although existing approaches have brought impressive r…