activity
20152021
most citedJointly Attentive Spatial-Temporal Pooling Networks for Video-based Person Re-Identification

40 citations · 99 across the 8 of their papers we have counts for

collaborators
Showing cs.CVShow all

7 papers · 1 filter

cs.CV2021

Towards Adversarial Patch Analysis and Certified Defense against Crowd Counting

Qiming Wu, Zhikang Zou, Pan Zhou +3

Crowd counting has drawn much attention due to its importance in safety-critical surveillance systems. Especially, deep neural network (DNN) methods have significantly reduced esti…

cs.CV202011 cited

Jointly Cross- and Self-Modal Graph Attention Network for Query-Based Moment Localization

Daizong Liu, Xiaoye Qu, Xiao-Yang Liu +3

Query-based moment localization is a new task that localizes the best matched segment in an untrimmed video according to a given sentence query. In this localization task, one shou…

cs.CV2020

Identity-Aware Attribute Recognition via Real-Time Distributed Inference in Mobile Edge Clouds

Zichuan Xu, Jiangkai Wu, Qiufen Xia +3

With the development of deep learning technologies, attribute recognition and person re-identification (re-ID) have attracted extensive attention and achieved continuous improvemen…

cs.CV202013 cited

Fine-grained Iterative Attention Network for TemporalLanguage Localization in Videos

Xiaoye Qu, Pengwei Tang, Zhikang Zhou +3

Temporal language localization in videos aims to ground one video segment in an untrimmed video based on a given sentence query. To tackle this task, designing an effective model t…

cs.CV2019

EnlightenGAN: Deep Light Enhancement without Paired Supervision

Yifan Jiang, Xinyu Gong, Ding Liu +6

Deep learning-based methods have achieved remarkable success in image restoration and enhancement, but are they still competitive when there is a lack of paired training data? As o…

cs.CV201910 cited

MHP-VOS: Multiple Hypotheses Propagation for Video Object Segmentation

Shuangjie Xu, Daizong Liu, Linchao Bao +2

We address the problem of semi-supervised video object segmentation (VOS), where the masks of objects of interests are given in the first frame of an input video. To deal with chal…