87 citations · 188 across the 12 of their papers we have counts for
22 papers
UNITS: Unsupervised Intermediate Training Stage for Scene Text Detection
Youhui Guo, Yu Zhou, Xugong Qin +2
Recent scene text detection methods are almost based on deep learning and data-driven. Synthetic data is commonly adopted for pre-training due to expensive annotation cost. However…
MBEV: Multi-Camera Joint 3D Detection and Segmentation with Unified Birds-Eye View Representation
Enze Xie, Zhiding Yu, Daquan Zhou +5
In this paper, we propose MBEV, a unified framework that jointly performs 3D object detection and map segmentation in the Birds Eye View~(BEV) space with multi-camera image inp…
Improving Monocular Visual Odometry Using Learned Depth
Libo Sun, Wei Yin, Enze Xie +3
Monocular visual odometry (VO) is an important task in robotics and computer vision. Thus far, how to build accurate and robust monocular VO systems that can work well in diverse s…
WegFormer: Transformers for Weakly Supervised Semantic Segmentation
Chunmeng Liu, Enze Xie, Wenjia Wang +3
Although convolutional neural networks (CNNs) have achieved remarkable progress in weakly supervised semantic segmentation (WSSS), the effective receptive field of CNN is insuffici…
PolarMask++: Enhanced Polar Representation for Single-Shot Instance Segmentation and Beyond
Enze Xie, Wenhai Wang, Mingyu Ding +2
Reducing the complexity of the pipeline of instance segmentation is crucial for real-world applications. This work addresses this issue by introducing an anchor-box free and single…
SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers
Enze Xie, Wenhai Wang, Zhiding Yu +3
We present SegFormer, a simple, efficient yet powerful semantic segmentation framework which unifies Transformers with lightweight multilayer perception (MLP) decoders. SegFormer h…