activity
20182025
most citedMBEV: Multi-Camera Joint 3D Detection and Segmentation with Unified Birds-Eye View Representation

87 citations · 188 across the 12 of their papers we have counts for

collaborators

22 papers

cs.CV2022

UNITS: Unsupervised Intermediate Training Stage for Scene Text Detection

Youhui Guo, Yu Zhou, Xugong Qin +2

Recent scene text detection methods are almost based on deep learning and data-driven. Synthetic data is commonly adopted for pre-training due to expensive annotation cost. However…

cs.CV202287 cited

MBEV: Multi-Camera Joint 3D Detection and Segmentation with Unified Birds-Eye View Representation

Enze Xie, Zhiding Yu, Daquan Zhou +5

In this paper, we propose MBEV, a unified framework that jointly performs 3D object detection and map segmentation in the Birds Eye View~(BEV) space with multi-camera image inp…

cs.CV2022

Improving Monocular Visual Odometry Using Learned Depth

Libo Sun, Wei Yin, Enze Xie +3

Monocular visual odometry (VO) is an important task in robotics and computer vision. Thus far, how to build accurate and robust monocular VO systems that can work well in diverse s…

cs.CV20223 cited

WegFormer: Transformers for Weakly Supervised Semantic Segmentation

Chunmeng Liu, Enze Xie, Wenjia Wang +3

Although convolutional neural networks (CNNs) have achieved remarkable progress in weakly supervised semantic segmentation (WSSS), the effective receptive field of CNN is insuffici…

cs.CV2021

PolarMask++: Enhanced Polar Representation for Single-Shot Instance Segmentation and Beyond

Enze Xie, Wenhai Wang, Mingyu Ding +2

Reducing the complexity of the pipeline of instance segmentation is crucial for real-world applications. This work addresses this issue by introducing an anchor-box free and single…

cs.CV2021

SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers

Enze Xie, Wenhai Wang, Zhiding Yu +3

We present SegFormer, a simple, efficient yet powerful semantic segmentation framework which unifies Transformers with lightweight multilayer perception (MLP) decoders. SegFormer h…