activity
20152022
most citedLearning Spatio-Temporal Representation with Pseudo-3D Residual Networks

251 citations · 1k across the 60 of their papers we have counts for

collaborators

97 papers

cs.CV20212 cited

ViDA-MAN: Visual Dialog with Digital Humans

Tong Shen, Jiawei Zuo, Fan Shi +7

We demonstrate ViDA-MAN, a digital-human agent for multi-modal interaction, which offers realtime audio-visual responses to instant speech inquiries. Compared to traditional text o…

cs.CV2021

CoSeg: Cognitively Inspired Unsupervised Generic Event Segmentation

Xiao Wang, Jingen Liu, Tao Mei +1

Some cognitive research has discovered that humans accomplish event segmentation as a side effect of event anticipation. Inspired by this discovery, we propose a simple yet effecti…

cs.CV20219 cited

Semi-Supervised Domain Generalizable Person Re-Identification

Lingxiao He, Wu Liu, Jian Liang +4

Existing person re-identification (re-id) methods are stuck when deployed to a new unseen scenario despite the success in cross-camera person matching. Recent efforts have been sub…

cs.CV20212 cited

Memory-Augmented Non-Local Attention for Video Super-Resolution

Jiyang Yu, Jingen Liu, Liefeng Bo +1

In this paper, we propose a novel video super-resolution method that aims at generating high-fidelity high-resolution (HR) videos from low-resolution (LR) ones. Previous methods pr…

cs.CV20211 cited

X-modaler: A Versatile and High-performance Codebase for Cross-modal Analytics

Yehao Li, Yingwei Pan, Jingwen Chen +2

With the rise and development of deep learning over the past decade, there has been a steady momentum of innovation and breakthroughs that convincingly push the state-of-the-art of…

cs.CV20211 cited

A Low Rank Promoting Prior for Unsupervised Contrastive Learning

Yu Wang, Jingyang Lin, Qi Cai +4

Unsupervised learning is just at a tipping point where it could really take off. Among these approaches, contrastive learning has seen tremendous progress and led to state-of-the-a…