4 citations · 11 across the 15 of their papers we have counts for
12 papers · 1 filter
Video Individual Counting With Implicit One-to-Many Matching
Xuhui Zhu, Jing Xu, Bingjie Wang +2
Video Individual Counting (VIC) is a recently introduced task that aims to estimate pedestrian flux from a video. It extends conventional Video Crowd Counting (VCC) beyond the per-…
Scaling Multi-Camera 3D Object Detection through Weak-to-Strong Eliciting
Hao Lu, Jiaqi Tang, Xinli Xu +6
The emergence of Multi-Camera 3D Object Detection (MC3D-Det), facilitated by bird's-eye view (BEV) representation, signifies a notable progression in 3D object detection. Scaling M…
In-Context Matting
He Guo, Zixuan Ye, Zhiguo Cao +1
We introduce in-context matting, a novel task setting of image matting. Given a reference image of a certain foreground and guided priors such as points, scribbles, and masks, in-c…
ESTextSpotter: Towards Better Scene Text Spotting with Explicit Synergy in Transformer
Mingxin Huang, Jiaxin Zhang, Dezhi Peng +5
In recent years, end-to-end scene text spotting approaches are evolving to the Transformer-based framework. While previous studies have shown the crucial importance of the intrinsi…
On Point Affiliation in Feature Upsampling
Wenze Liu, Hao Lu, Yuliang Liu +1
We introduce the notion of point affiliation into feature upsampling. By abstracting a feature map into non-overlapped semantic clusters formed by points of identical semantic mean…
rPPG-MAE: Self-supervised Pre-training with Masked Autoencoders for Remote Physiological Measurement
Xin Liu, Yuting Zhang, Zitong Yu +3
Remote photoplethysmography (rPPG) is an important technique for perceiving human vital signs, which has received extensive attention. For a long time, researchers have focused on…