10 citations · 49 across the 12 of their papers we have counts for
14 papers · 1 filter
One More Step: A Versatile Plug-and-Play Module for Rectifying Diffusion Schedule Flaws and Enhancing Low-Frequency Controls
Minghui Hu, Jianbin Zheng, Chuanxia Zheng +3
It is well known that many open-released foundational diffusion models have difficulty in generating images that substantially depart from average brightness, despite such images b…
Unified Discrete Diffusion for Simultaneous Vision-Language Generation
Minghui Hu, Chuanxia Zheng, Heliang Zheng +5
The recently developed discrete diffusion models perform extraordinarily well in the text-to-image task, showing significant promise for handling the multi-modality signals. In thi…
High-Quality Pluralistic Image Completion via Code Shared VQGAN
Chuanxia Zheng, Guoxian Song, Tat-Jen Cham +3
PICNet pioneered the generation of multiple and diverse results for image completion task, but it required a careful balance between loss (diversity) and reconstruct…
Visiting the Invisible: Layer-by-Layer Completed Scene Decomposition
Chuanxia Zheng, Duy-Son Dao, Guoxian Song +2
Existing scene understanding systems mainly focus on recognizing the visible parts of a scene, ignoring the intact appearance of physical objects in the real-world. Concurrently, i…
The Spatially-Correlative Loss for Various Image Translation Tasks
Chuanxia Zheng, Tat-Jen Cham, Jianfei Cai
We propose a novel spatially-correlative loss that is simple, efficient and yet effective for preserving scene structure consistency while supporting large appearance changes durin…
Recovering Facial Reflectance and Geometry from Multi-view Images
Guoxian Song, Jianmin Zheng, Jianfei Cai +1
While the problem of estimating shapes and diffuse reflectances of human faces from images has been extensively studied, there is relatively less work done on recovering the specul…