activity
20162023
most citedPKU-MMD: A Large Scale Benchmark for Continuous Multi-Modal Human Action Understanding

148 citations · 246 across the 17 of their papers we have counts for

collaborators
Showing cs.CVShow all

34 papers · 1 filter

cs.CV202114 cited

Video Coding for Machine: Compact Visual Representation Compression for Intelligent Collaborative Analytics

Wenhan Yang, Haofeng Huang, Yueyu Hu +2

Video Coding for Machines (VCM) is committed to bridging to an extent separate research tracks of video/image compression and feature compression, and attempts to optimize compactn…

cs.CV2021

Revisit Visual Representation in Analytics Taxonomy: A Compression Perspective

Yueyu Hu, Wenhan Yang, Haofeng Huang +1

Visual analytics have played an increasingly critical role in the Internet of Things, where massive visual signals have to be compressed and fed into machines. But facing such big…

cs.CV20215 cited

HLA-Face: Joint High-Low Adaptation for Low Light Face Detection

Wenjing Wang, Wenhan Yang, Jiaying Liu

Face detection in low light scenarios is challenging but vital to many practical applications, e.g., surveillance video, autonomous driving at night. Most existing face detectors h…

cs.CV20211 cited

Co-Grounding Networks with Semantic Attention for Referring Expression Comprehension in Videos

Sijie Song, Xudong Lin, Jiaying Liu +2

In this paper, we address the problem of referring expression comprehension in videos, which is challenging due to complex expression and scene dynamics. Unlike previous methods wh…

cs.CV20213 cited

Progressive Depth Learning for Single Image Dehazing

Yudong Liang, Bin Wang, Jiaying Liu +3

The formulation of the hazy image is mainly dominated by the reflected lights and ambient airlight. Existing dehazing methods often ignore the depth cues and fail in distant areas…

cs.CV2021

Template-Free Try-on Image Synthesis via Semantic-guided Optimization

Chien-Lung Chou, Chieh-Yun Chen, Chia-Wei Hsieh +3

The virtual try-on task is so attractive that it has drawn considerable attention in the field of computer vision. However, presenting the three-dimensional (3D) physical character…