24 citations · 27 across the 3 of their papers we have counts for
3 papers
cs.CV2023
A Hybrid CNN-Transformer Architecture with Frequency Domain Contrastive Learning for Image Deraining
Cheng Wang, Wei Li
Image deraining is a challenging task that involves restoring degraded images affected by rain streaks.
cs.CV2023★ 3 cited
Language-Guided 3D Object Detection in Point Cloud for Autonomous Driving
Wenhao Cheng, Junbo Yin, Wei Li +2
This paper addresses the problem of 3D referring expression comprehension (REC) in autonomous driving scenario, which aims to ground a natural language to the targeted region in Li…
cs.CV2023★ 24 cited
DeCap: Decoding CLIP Latents for Zero-Shot Captioning via Text-Only Training
Wei Li, Linchao Zhu, Longyin Wen +1
Large-scale pre-trained multi-modal models (e.g., CLIP) demonstrate strong zero-shot transfer capability in many discriminative tasks. Their adaptation to zero-shot image-condition…