333 citations · 1.3k across the 55 of their papers we have counts for
113 papers
Coordinate Attention for Efficient Mobile Network Design
Qibin Hou, Daquan Zhou, Jiashi Feng
Recent studies on mobile network design have demonstrated the remarkable effectiveness of channel attention (e.g., the Squeeze-and-Excitation attention) for lifting model performan…
Unleashing the Power of Contrastive Self-Supervised Visual Models via Contrast-Regularized Fine-Tuning
Yifan Zhang, Bryan Hooi, Dapeng Hu +2
Contrastive self-supervised learning (CSL) has attracted increasing attention for model pre-training via unlabeled data. The resulted CSL models provide instance-discriminative vis…
ORDNet: Capturing Omni-Range Dependencies for Scene Parsing
Shaofei Huang, Si Liu, Tianrui Hui +4
Learning to capture dependencies between spatial positions is essential to many visual tasks, especially the dense labeling problems like scene parsing. Existing methods can effect…
Improving Generalization in Reinforcement Learning with Mixture Regularization
Kaixin Wang, Bingyi Kang, Jie Shao +1
Deep reinforcement learning (RL) agents trained in a limited set of environments tend to suffer overfitting and fail to generalize to unseen testing environments. To improve their…
Towards Accurate Human Pose Estimation in Videos of Crowded Scenes
Li Yuan, Shuning Chang, Xuecheng Nie +5
Video-based human pose estimation in crowded scenes is a challenging problem due to occlusion, motion blur, scale variation and viewpoint change, etc. Prior approaches always fail…
A Simple Baseline for Pose Tracking in Videos of Crowded Scenes
Li Yuan, Shuning Chang, Ziyuan Huang +6
This paper presents our solution to ACM MM challenge: Large-scale Human-centric Video Analysis in Complex Events\cite{lin2020human}; specifically, here we focus on Track3: Crowd Po…