5 citations · 9 across the 6 of their papers we have counts for
6 papers · 1 filter
RoPETR: Improving Temporal Camera-Only 3D Detection by Integrating Enhanced Rotary Position Embedding
Hang Ji, Tao Ni, Xufeng Huang +4
This technical report introduces a targeted improvement to the StreamPETR framework, specifically aimed at enhancing velocity estimation, a critical factor influencing the overall…
MV-DETR: Multi-modality indoor object detection by Multi-View DEtecton TRansformers
Zichao Dong, Yilin Zhang, Xufeng Huang +4
We introduce a novel MV-DETR pipeline which is effective while efficient transformer based detection method. Given input RGBD data, we notice that there are super strong pretrainin…
LVIC: Multi-modality segmentation by Lifting Visual Info as Cue
Zichao Dong, Bowen Pang, Xufeng Huang +3
Multi-modality fusion is proven an effective method for 3d perception for autonomous driving. However, most current multi-modality fusion pipelines for LiDAR semantic segmentation…
PeP: a Point enhanced Painting method for unified point cloud tasks
Zichao Dong, Hang Ji, Xufeng Huang +3
Point encoder is of vital importance for point cloud recognition. As the very beginning step of whole model pipeline, adding features from diverse sources and providing stronger fe…
OG: Equip vision occupancy with instance segmentation and visual grounding
Zichao Dong, Hang Ji, Weikun Zhang +2
Occupancy prediction tasks focus on the inference of both geometry and semantic labels for each voxel, which is an important perception mission. However, it is still a semantic seg…
OVO: Open-Vocabulary Occupancy
Zhiyu Tan, Zichao Dong, Cheng Zhang +3
Semantic occupancy prediction aims to infer dense geometry and semantics of surroundings for an autonomous agent to operate safely in the 3D environment. Existing occupancy predict…