147 citations · 396 across the 14 of their papers we have counts for
10 papers · 1 filter
BigDetection: A Large-scale Benchmark for Improved Object Detector Pre-training
Likun Cai, Zhi Zhang, Yi Zhu +3
Multiple datasets and open challenges for object detection have been introduced in recent years. To build more general and powerful object detection systems, in this paper, we cons…
Blending Anti-Aliasing into Vision Transformer
Shengju Qian, Hao Shao, Yi Zhu +2
The transformer architectures, based on self-attention mechanism and convolution-free design, recently found superior performance and booming applications in computer vision. Howev…
Progressive Coordinate Transforms for Monocular 3D Object Detection
Li Wang, Li Zhang, Yi Zhu +4
Recognizing and localizing objects in the 3D space is a crucial ability for an AI agent to perceive its surrounding environment. While significant progress has been achieved with e…
Video Contrastive Learning with Global Context
Haofei Kuang, Yi Zhu, Zhi Zhang +5
Contrastive learning has revolutionized self-supervised image representation learning field, and recently been adapted to video domain. One of the greatest advantages of contrastiv…
A Unified Efficient Pyramid Transformer for Semantic Segmentation
Fangrui Zhu, Yi Zhu, Li Zhang +3
Semantic segmentation is a challenging problem due to difficulties in modeling context in complex scenes and class confusions along boundaries. Most literature either focuses on co…
CrossNorm and SelfNorm for Generalization under Distribution Shifts
Zhiqiang Tang, Yunhe Gao, Yi Zhu +3
Traditional normalization techniques (e.g., Batch Normalization and Instance Normalization) generally and simplistically assume that training and test data follow the same distribu…