4 citations · 8 across the 3 of their papers we have counts for
3 papers
cs.CV2024★ 1 cited
4D Contrastive Superflows are Dense 3D Representation Learners
Xiang Xu, Lingdong Kong, Hui Shuai +5
In the realm of autonomous driving, accurate 3D perception is the foundation. However, developing such models relies on extensive human annotations -- a process that is both costly…
cs.CV2023★ 4 cited
RoboBEV: Towards Robust Bird's Eye View Perception under Corruptions
Shaoyuan Xie, Lingdong Kong, Wenwei Zhang +4
The recent advances in camera-based bird's eye view (BEV) representation exhibit great potential for in-vehicle 3D perception. Despite the substantial progress achieved on standard…
cs.CV2023★ 3 cited
Aligning Bag of Regions for Open-Vocabulary Object Detection
Size Wu, Wenwei Zhang, Sheng Jin +2
Pre-trained vision-language models (VLMs) learn to align vision and language representations on large-scale datasets, where each image-text pair usually contains a bag of semantic…