activity
20152024
most citedLearning Spatio-Temporal Representation with Pseudo-3D Residual Networks

251 citations · 585 across the 60 of their papers we have counts for

collaborators
Showing 2020Show all

11 papers · 1 filter

cs.CV2020

Joint Contrastive Learning with Infinite Possibilities

Qi Cai, Yu Wang, Yingwei Pan +2

This paper explores useful modifications of the recent development in contrastive learning via novel probabilistic modeling. We derive a particular form of contrastive loss named J…

cs.CV2020★ 4 cited

Learning to Localize Actions from Moments

Fuchen Long, Ting Yao, Zhaofan Qiu +3

With the knowledge of action moments (i.e., trimmed video clips that each contains an action instance), humans could routinely localize an action temporally in an untrimmed video.…

cs.CV2020

SeCo: Exploring Sequence Supervision for Unsupervised Representation Learning

Ting Yao, Yiheng Zhang, Zhaofan Qiu +2

A steady momentum of innovations and breakthroughs has convincingly pushed the limits of unsupervised image representation learning. Compared to static 2D images, video has one mor…

cs.CV2020★ 1 cited

Pre-training for Video Captioning Challenge 2020 Summary

Yingwei Pan, Jun Xu, Yehao Li +2

The Pre-training for Video Captioning Challenge 2020 Summary: results and challenge participants' technical reports.

cs.CV2020

Single Shot Video Object Detector

Jiajun Deng, Yingwei Pan, Ting Yao +3

Single shot detectors that are potentially faster and simpler than two-stage detectors tend to be more applicable to object detection in videos. Nevertheless, the extension of such…

cs.CV2020★ 27 cited

Auto-captions on GIF: A Large-scale Video-sentence Dataset for Vision-language Pre-training

Yingwei Pan, Yehao Li, Jianjie Luo +3

In this work, we present Auto-captions on GIF, which is a new large-scale pre-training dataset for generic video understanding. All video-sentence pairs are created by automaticall…