activity
20222024
most citedMV-FCOS3D++: Multi-View Camera-Only 4D Object Detection with Pretrained Monocular Backbones

17 citations · 26 across the 7 of their papers we have counts for

collaborators

5 papers

cs.CV20235 cited

OV-PARTS: Towards Open-Vocabulary Part Segmentation

Meng Wei, Xiaoyu Yue, Wenwei Zhang +3

Segmenting and recognizing diverse object parts is a crucial ability in applications spanning various computer vision and robotic tasks. While significant progress has been made in…

cs.CV20233 cited

Object2Scene: Putting Objects in Context for Open-Vocabulary 3D Detection

Chenming Zhu, Wenwei Zhang, Tai Wang +2

Point cloud-based open-vocabulary 3D object detection aims to detect 3D categories that do not have ground-truth annotations in the training set. It is extremely challenging becaus…

cs.CV20231 cited

MV-JAR: Masked Voxel Jigsaw and Reconstruction for LiDAR-Based Self-Supervised Pre-Training

Runsen Xu, Tai Wang, Wenwei Zhang +4

This paper introduces the Masked Voxel Jigsaw and Reconstruction (MV-JAR) method for LiDAR-based self-supervised pre-training and a carefully designed data-efficient 3D object dete…

cs.CV2023

Position-Guided Point Cloud Panoptic Segmentation Transformer

Zeqi Xiao, Wenwei Zhang, Tai Wang +3

DEtection TRansformer (DETR) started a trend that uses a group of learnable queries for unified visual perception. This work begins by applying this appealing paradigm to LiDAR-bas…

cs.CV202217 cited

MV-FCOS3D++: Multi-View Camera-Only 4D Object Detection with Pretrained Monocular Backbones

Tai Wang, Qing Lian, Chenming Zhu +2

In this technical report, we present our solution, dubbed MV-FCOS3D++, for the Camera-Only 3D Detection track in Waymo Open Dataset Challenge 2022. For multi-view camera-only 3D de…