activity
20172022
most citedHydraPlus-Net: Attentive Deep Features for Pedestrian Analysis

90 citations · 218 across the 14 of their papers we have counts for

collaborators
Showing cs.CVShow all

21 papers · 1 filter

cs.CV20223 cited

Towards Explainable 3D Grounded Visual Question Answering: A New Benchmark and Strong Baseline

Lichen Zhao, Daigang Cai, Jing Zhang +6

Recently, 3D vision-and-language tasks have attracted increasing research interest. Compared to other vision-and-language tasks, the 3D visual question answering (VQA) task is less…

cs.CV2022

X-Learner: Learning Cross Sources and Tasks for Universal Visual Representation

Yinan He, Gengshi Huang, Siyu Chen +7

In computer vision, pre-training models based on largescale supervised learning have been proven effective over the past few years. However, existing works mostly focus on learning…

cs.CV2021

VoteHMR: Occlusion-Aware Voting Network for Robust 3D Human Mesh Recovery from Partial Point Clouds

Guanze Liu, Yu Rong, Lu Sheng

3D human mesh recovery from point clouds is essential for various tasks, including AR/VR and human behavior understanding. Previous works in this field either require high-quality…

cs.CV20215 cited

Back-tracing Representative Points for Voting-based 3D Object Detection in Point Clouds

Bowen Cheng, Lu Sheng, Shaoshuai Shi +2

3D object detection in point clouds is a challenging vision task that benefits various applications for understanding the 3D visual world. Lots of recent research focuses on how to…

cs.CV2021

ForgeryNet: A Versatile Benchmark for Comprehensive Forgery Analysis

Yinan He, Bei Gan, Siyu Chen +6

The rapid progress of photorealistic synthesis techniques has reached at a critical point where the boundary between real and manipulated images starts to blur. Thus, benchmarking…

cs.CV20207 cited

PV-NAS: Practical Neural Architecture Search for Video Recognition

Zihao Wang, Chen Lin, Lu Sheng +2

Recently, deep learning has been utilized to solve video recognition problem due to its prominent representation ability. Deep neural networks for video tasks is highly customized…