21 citations · 84 across the 14 of their papers we have counts for
7 papers · 1 filter
DR.VIC: Decomposition and Reasoning for Video Individual Counting
Tao Han, Lei Bai, Junyu Gao +2
Pedestrian counting is a fundamental tool for understanding pedestrian patterns and crowd flow analysis. Existing works (e.g., image-level pedestrian counting, crossline crowd coun…
Trajectory Forecasting from Detection with Uncertainty-Aware Motion Encoding
Pu Zhang, Lei Bai, Jianru Xue +3
Trajectory forecasting is critical for autonomous platforms to make safe planning and actions. Currently, most trajectory forecasting methods assume that object trajectories have b…
Graph-Based 3D Multi-Person Pose Estimation Using Multi-View Images
Size Wu, Sheng Jin, Wentao Liu +4
This paper studies the task of estimating the 3D human poses of multiple persons from multiple calibrated camera views. Following the top-down paradigm, we decompose the task into…
GLiT: Neural Architecture Search for Global and Local Image Transformer
Boyu Chen, Peixia Li, Chuming Li +6
We introduce the first Neural Architecture Search (NAS) method to find a better transformer architecture for image recognition. Recently, transformers without CNN-based backbones a…
PSViT: Better Vision Transformer via Token Pooling and Attention Sharing
Boyu Chen, Peixia Li, Baopu Li +6
In this paper, we observe two levels of redundancies when applying vision transformers (ViT) for image recognition. First, fixing the number of tokens through the whole network pro…
Temporal-Channel Transformer for 3D Lidar-Based Video Object Detection in Autonomous Driving
Zhenxun Yuan, Xiao Song, Lei Bai +3
The strong demand of autonomous driving in the industry has lead to strong interest in 3D object detection and resulted in many excellent 3D object detection algorithms. However, t…