activity
20222024
most citedCoordinate-Aligned Multi-Camera Collaboration for Active Multi-Object Tracking

4 citations · 7 across the 6 of their papers we have counts for

collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV2024

Learning Generalizable Human Motion Generator with Reinforcement Learning

Yunyao Mao, Xiaoyang Liu, Wengang Zhou +2

Text-driven human motion generation, as one of the vital tasks in computer-aided content creation, has recently attracted increasing attention. While pioneering research has largel…

cs.CV20231 cited

IMD: 3D Action Representation Learning with Inter- and Intra-modal Mutual Distillation

Yunyao Mao, Jiajun Deng, Wengang Zhou +3

Recent progresses on self-supervised 3D human action representation learning are largely attributed to contrastive learning. However, in conventional contrastive frameworks, the ri…

cs.CV2023

Text-Only Training for Visual Storytelling

Yuechen Wang, Wengang Zhou, Zhenbo Lu +1

Visual storytelling aims to generate a narrative based on a sequence of images, necessitating both vision-language alignment and coherent story generation. Most existing solutions…

cs.CV2022

UDoc-GAN: Unpaired Document Illumination Correction with Background Light Prior

Yonghui Wang, Wengang Zhou, Zhenbo Lu +1

Document images captured by mobile devices are usually degraded by uncontrollable illumination, which hampers the clarity of document content. Recently, a series of research effort…

cs.CV20224 cited

Coordinate-Aligned Multi-Camera Collaboration for Active Multi-Object Tracking

Zeyu Fang, Jian Zhao, Mingyu Yang +3

Active Multi-Object Tracking (AMOT) is a task where cameras are controlled by a centralized system to adjust their poses automatically and collaboratively so as to maximize the cov…