activity
20162021
most citedMulti-modal Factorized Bilinear Pooling with Co-Attention Learning for Visual Question Answering

102 citations · 130 across the 9 of their papers we have counts for

collaborators

11 papers

eess.IV20218 cited

ACN: Adversarial Co-training Network for Brain Tumor Segmentation with Missing Modalities

Yixin Wang, Yang Zhang, Yang Liu +6

Accurate segmentation of brain tumors from magnetic resonance imaging (MRI) is clinically relevant in diagnoses, prognoses and surgery treatment, which requires multiple modalities…

cs.SI20211 cited

"Do You Know You Are Tracked by Photos That You Didn't Take": Location-Aware Multi-Party Image Privacy Protection

Joshua Morris, Sara Newman, Kannappan Palaniappan +2

Most existing image privacy protection works focus mainly on the privacy of photo owners and their friends, but lack the consideration of other people who are in the background of…

cs.CV20202 cited

Cluster-level Feature Alignment for Person Re-identification

Qiuyu Chen, Wei Zhang, Jianping Fan

Instance-level alignment is widely exploited for person re-identification, e.g. spatial alignment, latent semantic alignment and triplet alignment. This paper probes another featur…

cs.CV20201 cited

Automatic Image Labelling at Pixel Level

Xiang Zhang, Wei Zhang, Jinye Peng +1

The performance of deep networks for semantic image segmentation largely depends on the availability of large-scale training images which are labelled at the pixel level. Typically…

cs.CV202012 cited

Adaptive Fractional Dilated Convolution Network for Image Aesthetics Assessment

Qiuyu Chen, Wei Zhang, Ning Zhou +4

To leverage deep learning for image aesthetics assessment, one critical but unsolved issue is how to seamlessly incorporate the information of image aspect ratios to learn more rob…

cs.CV20192 cited

MOD: A Deep Mixture Model with Online Knowledge Distillation for Large Scale Video Temporal Concept Localization

Rongcheng Lin, Jing Xiao, Jianping Fan

In this paper, we present and discuss a deep mixture model with online knowledge distillation (MOD) for large-scale video temporal concept localization, which is ranked 3rd in the…