activity
20232025
most citedOVO: Open-Vocabulary Occupancy

5 citations · 9 across the 6 of their papers we have counts for

collaborators
Showing cs.CVShow all

6 papers · 1 filter

cs.CV2025

RoPETR: Improving Temporal Camera-Only 3D Detection by Integrating Enhanced Rotary Position Embedding

Hang Ji, Tao Ni, Xufeng Huang +4

This technical report introduces a targeted improvement to the StreamPETR framework, specifically aimed at enhancing velocity estimation, a critical factor influencing the overall…

cs.CV2024

MV-DETR: Multi-modality indoor object detection by Multi-View DEtecton TRansformers

Zichao Dong, Yilin Zhang, Xufeng Huang +4

We introduce a novel MV-DETR pipeline which is effective while efficient transformer based detection method. Given input RGBD data, we notice that there are super strong pretrainin…

cs.CV2024

LVIC: Multi-modality segmentation by Lifting Visual Info as Cue

Zichao Dong, Bowen Pang, Xufeng Huang +3

Multi-modality fusion is proven an effective method for 3d perception for autonomous driving. However, most current multi-modality fusion pipelines for LiDAR semantic segmentation…

cs.CV2023

PeP: a Point enhanced Painting method for unified point cloud tasks

Zichao Dong, Hang Ji, Xufeng Huang +3

Point encoder is of vital importance for point cloud recognition. As the very beginning step of whole model pipeline, adding features from diverse sources and providing stronger fe…

cs.CV20231 cited

OG: Equip vision occupancy with instance segmentation and visual grounding

Zichao Dong, Hang Ji, Weikun Zhang +2

Occupancy prediction tasks focus on the inference of both geometry and semantic labels for each voxel, which is an important perception mission. However, it is still a semantic seg…

cs.CV20235 cited

OVO: Open-Vocabulary Occupancy

Zhiyu Tan, Zichao Dong, Cheng Zhang +3

Semantic occupancy prediction aims to infer dense geometry and semantics of surroundings for an autonomous agent to operate safely in the 3D environment. Existing occupancy predict…