most citedBatchFormerV2: Exploring Sample Relationships for Dense Representation Learning

10 citations · 19 across the 4 of their papers we have counts for

collaborators

5 papers

cs.CV20228 cited

Unified Discrete Diffusion for Simultaneous Vision-Language Generation

Minghui Hu, Chuanxia Zheng, Heliang Zheng +5

The recently developed discrete diffusion models perform extraordinarily well in the text-to-image task, showing significant promise for handling the multi-modality signals. In thi…

cs.CV2022

Cross-Modal Contrastive Learning for Robust Reasoning in VQA

Qi Zheng, Chaoyue Wang, Daqing Liu +2

Multi-modal reasoning in visual question answering (VQA) has witnessed rapid progress recently. However, most reasoning models heavily rely on shortcuts learned from training data,…

cs.CV20221 cited

Neural Maximum A Posteriori Estimation on Unpaired Data for Motion Deblurring

Youjian Zhang, Chaoyue Wang, Dacheng Tao

Real-world dynamic scene deblurring has long been a challenging task since paired blurry-sharp training data is unavailable. Conventional Maximum A Posteriori estimation and deep l…

cs.CV202210 cited

BatchFormerV2: Exploring Sample Relationships for Dense Representation Learning

Zhi Hou, Baosheng Yu, Chaoyue Wang +2

Attention mechanisms have been very popular in deep neural networks, where the Transformer architecture has achieved great success in not only natural language processing but also…

cs.CV2020

Exposure Trajectory Recovery from Motion Blur

Youjian Zhang, Chaoyue Wang, Stephen J. Maybank +1

Motion blur in dynamic scenes is an important yet challenging research topic. Recently, deep learning methods have achieved impressive performance for dynamic scene deblurring. How…