10 citations · 19 across the 4 of their papers we have counts for
5 papers
Unified Discrete Diffusion for Simultaneous Vision-Language Generation
Minghui Hu, Chuanxia Zheng, Heliang Zheng +5
The recently developed discrete diffusion models perform extraordinarily well in the text-to-image task, showing significant promise for handling the multi-modality signals. In thi…
Cross-Modal Contrastive Learning for Robust Reasoning in VQA
Qi Zheng, Chaoyue Wang, Daqing Liu +2
Multi-modal reasoning in visual question answering (VQA) has witnessed rapid progress recently. However, most reasoning models heavily rely on shortcuts learned from training data,…
Neural Maximum A Posteriori Estimation on Unpaired Data for Motion Deblurring
Youjian Zhang, Chaoyue Wang, Dacheng Tao
Real-world dynamic scene deblurring has long been a challenging task since paired blurry-sharp training data is unavailable. Conventional Maximum A Posteriori estimation and deep l…
BatchFormerV2: Exploring Sample Relationships for Dense Representation Learning
Zhi Hou, Baosheng Yu, Chaoyue Wang +2
Attention mechanisms have been very popular in deep neural networks, where the Transformer architecture has achieved great success in not only natural language processing but also…
Exposure Trajectory Recovery from Motion Blur
Youjian Zhang, Chaoyue Wang, Stephen J. Maybank +1
Motion blur in dynamic scenes is an important yet challenging research topic. Recently, deep learning methods have achieved impressive performance for dynamic scene deblurring. How…