activity
20142024
most citedBoosting Image Captioning with Attributes

39 citations · 78 across the 16 of their papers we have counts for

collaborators

10 papers

cs.CV2022

Dynamic Temporal Filtering in Video Models

Fuchen Long, Zhaofan Qiu, Yingwei Pan +3

Video temporal dynamics is conventionally modeled with 3D spatial-temporal kernel or its factorized version comprised of 2D spatial kernel and 1D temporal kernel. The modeling powe…

cs.CV2022

SPE-Net: Boosting Point Cloud Analysis via Rotation Robustness Enhancement

Zhaofan Qiu, Yehao Li, Yu Wang +3

In this paper, we propose a novel deep architecture tailored for 3D point cloud applications, named as SPE-Net. The embedded ``Selective Position Encoding (SPE)'' procedure relies…

cs.LG20213 cited

Improving Self-supervised Learning with Automated Unsupervised Outlier Arbitration

Yu Wang, Jingyang Lin, Jingjing Zou +3

Our work reveals a structured shortcoming of the existing mainstream self-supervised learning methods. Whereas self-supervised learning frameworks usually take the prevailing perfe…

cs.CV20212 cited

A Style and Semantic Memory Mechanism for Domain Generalization

Yang Chen, Yu Wang, Yingwei Pan +3

Mainstream state-of-the-art domain generalization algorithms tend to prioritize the assumption on semantic invariance across domains. Meanwhile, the inherent intra-domain style inv…

cs.CV2021

Transferrable Contrastive Learning for Visual Domain Adaptation

Yang Chen, Yingwei Pan, Yu Wang +3

Self-supervised learning (SSL) has recently become the favorite among feature learning methodologies. It is therefore appealing for domain adaptation approaches to consider incorpo…

cs.CV20212 cited

CoCo-BERT: Improving Video-Language Pre-training with Contrastive Cross-modal Matching and Denoising

Jianjie Luo, Yehao Li, Yingwei Pan +3

BERT-type structure has led to the revolution of vision-language pre-training and the achievement of state-of-the-art results on numerous vision-language downstream tasks. Existing…