activity
20202023
most citedDual Contrastive Learning for Spatio-temporal Representation

20 citations · 45 across the 7 of their papers we have counts for

collaborators

7 papers

cs.CV2023

Prune Spatio-temporal Tokens by Semantic-aware Temporal Accumulation

Shuangrui Ding, Peisen Zhao, Xiaopeng Zhang +3

Transformers have become the primary backbone of the computer vision community due to their impressive performance. However, the unfriendly computation cost impedes their potential…

cs.CV202220 cited

Dual Contrastive Learning for Spatio-temporal Representation

Shuangrui Ding, Rui Qian, Hongkai Xiong

Contrastive learning has shown promising potential in self-supervised spatio-temporal representation learning. Most works naively sample different clips to construct positive and n…

cs.CV20212 cited

Contextualized Spatio-Temporal Contrastive Learning with Self-Supervision

Liangzhe Yuan, Rui Qian, Yin Cui +5

Modern self-supervised learning algorithms typically enforce persistency of instance representations across views. While being very effective on learning holistic image and video r…

cs.CV202115 cited

Revisiting 3D ResNets for Video Recognition

Xianzhi Du, Yeqing Li, Yin Cui +3

A recent work from Bello shows that training and scaling strategies may be more significant than model architectures for visual recognition. This short note studies effective train…

cs.CV2021

Enhancing Self-supervised Video Representation Learning via Multi-level Feature Optimization

Rui Qian, Yuxi Li, Huabin Liu +5

The crux of self-supervised video representation learning is to build general features from unlabeled videos. However, most recent works have mainly focused on high-level semantics…

cs.CV2020

Finding Action Tubes with a Sparse-to-Dense Framework

Yuxi Li, Weiyao Lin, Tao Wang +5

The task of spatial-temporal action detection has attracted increasing attention among researchers. Existing dominant methods solve this problem by relying on short-term informatio…