17 citations · 21 across the 3 of their papers we have counts for
3 papers
cs.CV2023★ 3 cited
Object2Scene: Putting Objects in Context for Open-Vocabulary 3D Detection
Chenming Zhu, Wenwei Zhang, Tai Wang +2
Point cloud-based open-vocabulary 3D object detection aims to detect 3D categories that do not have ground-truth annotations in the training set. It is extremely challenging becaus…
cs.CV2023★ 1 cited
MVImgNet: A Large-scale Dataset of Multi-view Images
Xianggang Yu, Mutian Xu, Yidan Zhang +10
Being data-driven is one of the most iconic properties of deep learning algorithms. The birth of ImageNet drives a remarkable trend of "learning from large-scale data" in computer…
cs.CV2022★ 17 cited
MV-FCOS3D++: Multi-View Camera-Only 4D Object Detection with Pretrained Monocular Backbones
Tai Wang, Qing Lian, Chenming Zhu +2
In this technical report, we present our solution, dubbed MV-FCOS3D++, for the Camera-Only 3D Detection track in Waymo Open Dataset Challenge 2022. For multi-view camera-only 3D de…