39 citations · 78 across the 16 of their papers we have counts for
10 papers
Dynamic Temporal Filtering in Video Models
Fuchen Long, Zhaofan Qiu, Yingwei Pan +3
Video temporal dynamics is conventionally modeled with 3D spatial-temporal kernel or its factorized version comprised of 2D spatial kernel and 1D temporal kernel. The modeling powe…
SPE-Net: Boosting Point Cloud Analysis via Rotation Robustness Enhancement
Zhaofan Qiu, Yehao Li, Yu Wang +3
In this paper, we propose a novel deep architecture tailored for 3D point cloud applications, named as SPE-Net. The embedded ``Selective Position Encoding (SPE)'' procedure relies…
Improving Self-supervised Learning with Automated Unsupervised Outlier Arbitration
Yu Wang, Jingyang Lin, Jingjing Zou +3
Our work reveals a structured shortcoming of the existing mainstream self-supervised learning methods. Whereas self-supervised learning frameworks usually take the prevailing perfe…
A Style and Semantic Memory Mechanism for Domain Generalization
Yang Chen, Yu Wang, Yingwei Pan +3
Mainstream state-of-the-art domain generalization algorithms tend to prioritize the assumption on semantic invariance across domains. Meanwhile, the inherent intra-domain style inv…
Transferrable Contrastive Learning for Visual Domain Adaptation
Yang Chen, Yingwei Pan, Yu Wang +3
Self-supervised learning (SSL) has recently become the favorite among feature learning methodologies. It is therefore appealing for domain adaptation approaches to consider incorpo…
CoCo-BERT: Improving Video-Language Pre-training with Contrastive Cross-modal Matching and Denoising
Jianjie Luo, Yehao Li, Yingwei Pan +3
BERT-type structure has led to the revolution of vision-language pre-training and the achievement of state-of-the-art results on numerous vision-language downstream tasks. Existing…