113 citations · 348 across the 32 of their papers we have counts for
54 papers · 1 filter
ReSTR: Convolution-free Referring Image Segmentation Using Transformers
Namyup Kim, Dongwon Kim, Cuiling Lan +2
Referring image segmentation is an advanced semantic segmentation task where target is not a predefined class but is described in natural language. Most of existing methods for thi…
Correlation-Aware Deep Tracking
Fei Xie, Chunyu Wang, Guangting Wang +3
Robustness and discrimination power are two fundamental requirements in visual object tracking. In most tracking paradigms, we find that the features extracted by the popular Siame…
When Shift Operation Meets Vision Transformer: An Extremely Simple Alternative to Attention Mechanism
Guangting Wang, Yucheng Zhao, Chuanxin Tang +2
Attention mechanism has been widely believed as the key to success of vision transformers (ViTs), since it provides a flexible and powerful way to model spatial relationships. Howe…
Multi-Scale Semantics-Guided Neural Networks for Efficient Skeleton-Based Human Action Recognition
Pengfei Zhang, Cuiling Lan, Wenjun Zeng +3
Skeleton data is of low dimension. However, there is a trend of using very deep and complicated feedforward neural networks to model the skeleton sequence without considering the c…
Pose-Guided Feature Learning with Knowledge Distillation for Occluded Person Re-Identification
Kecheng Zheng, Cuiling Lan, Wenjun Zeng +3
Occluded person re-identification (ReID) aims to match person images with occlusion. It is fundamentally challenging because of the serious occlusion which aggravates the misalignm…
Self-Supervised Visual Representations Learning by Contrastive Mask Prediction
Yucheng Zhao, Guangting Wang, Chong Luo +2
Advanced self-supervised visual representation learning methods rely on the instance discrimination (ID) pretext task. We point out that the ID task has an implicit semantic consis…