4 citations · 7 across the 6 of their papers we have counts for
5 papers · 1 filter
Learning Generalizable Human Motion Generator with Reinforcement Learning
Yunyao Mao, Xiaoyang Liu, Wengang Zhou +2
Text-driven human motion generation, as one of the vital tasks in computer-aided content creation, has recently attracted increasing attention. While pioneering research has largel…
IMD: 3D Action Representation Learning with Inter- and Intra-modal Mutual Distillation
Yunyao Mao, Jiajun Deng, Wengang Zhou +3
Recent progresses on self-supervised 3D human action representation learning are largely attributed to contrastive learning. However, in conventional contrastive frameworks, the ri…
Text-Only Training for Visual Storytelling
Yuechen Wang, Wengang Zhou, Zhenbo Lu +1
Visual storytelling aims to generate a narrative based on a sequence of images, necessitating both vision-language alignment and coherent story generation. Most existing solutions…
UDoc-GAN: Unpaired Document Illumination Correction with Background Light Prior
Yonghui Wang, Wengang Zhou, Zhenbo Lu +1
Document images captured by mobile devices are usually degraded by uncontrollable illumination, which hampers the clarity of document content. Recently, a series of research effort…
Coordinate-Aligned Multi-Camera Collaboration for Active Multi-Object Tracking
Zeyu Fang, Jian Zhao, Mingyu Yang +3
Active Multi-Object Tracking (AMOT) is a task where cameras are controlled by a centralized system to adjust their poses automatically and collaboratively so as to maximize the cov…