251 citations · 1k across the 60 of their papers we have counts for
97 papers
ViDA-MAN: Visual Dialog with Digital Humans
Tong Shen, Jiawei Zuo, Fan Shi +7
We demonstrate ViDA-MAN, a digital-human agent for multi-modal interaction, which offers realtime audio-visual responses to instant speech inquiries. Compared to traditional text o…
CoSeg: Cognitively Inspired Unsupervised Generic Event Segmentation
Xiao Wang, Jingen Liu, Tao Mei +1
Some cognitive research has discovered that humans accomplish event segmentation as a side effect of event anticipation. Inspired by this discovery, we propose a simple yet effecti…
Semi-Supervised Domain Generalizable Person Re-Identification
Lingxiao He, Wu Liu, Jian Liang +4
Existing person re-identification (re-id) methods are stuck when deployed to a new unseen scenario despite the success in cross-camera person matching. Recent efforts have been sub…
Memory-Augmented Non-Local Attention for Video Super-Resolution
Jiyang Yu, Jingen Liu, Liefeng Bo +1
In this paper, we propose a novel video super-resolution method that aims at generating high-fidelity high-resolution (HR) videos from low-resolution (LR) ones. Previous methods pr…
X-modaler: A Versatile and High-performance Codebase for Cross-modal Analytics
Yehao Li, Yingwei Pan, Jingwen Chen +2
With the rise and development of deep learning over the past decade, there has been a steady momentum of innovation and breakthroughs that convincingly push the state-of-the-art of…
A Low Rank Promoting Prior for Unsupervised Contrastive Learning
Yu Wang, Jingyang Lin, Qi Cai +4
Unsupervised learning is just at a tipping point where it could really take off. Among these approaches, contrastive learning has seen tremendous progress and led to state-of-the-a…