activity
20192021
most citedContextual Non-Local Alignment over Full-Scale Representation for Text-Based Person Search

61 citations · 62 across the 2 of their papers we have counts for

collaborators

6 papers

cs.CV20211 cited

Learning Canonical View Representation for 3D Shape Recognition with Arbitrary Views

Xin Wei, Yifei Gong, Fudong Wang +2

In this paper, we focus on recognizing 3D shapes from arbitrary views, i.e., arbitrary numbers and positions of viewpoints. It is a challenging and realistic setting for view-based…

cs.CV2021

Ask&Confirm: Active Detail Enriching for Cross-Modal Retrieval with Partial Query

Guanyu Cai, Jun Zhang, Xinyang Jiang +7

Text-based image retrieval has seen considerable progress in recent years. However, the performance of existing methods suffers in real life since the user is likely to provide an…

cs.CV202161 cited

Contextual Non-Local Alignment over Full-Scale Representation for Text-Based Person Search

Chenyang Gao, Guanyu Cai, Xinyang Jiang +6

Text-based person search aims at retrieving target person in an image gallery using a descriptive sentence of that person. It is very challenging since modal gap makes effectively…

cs.LG2020

Learning with Instance-Dependent Label Noise: A Sample Sieve Approach

Hao Cheng, Zhaowei Zhu, Xingyu Li +3

Human-annotated labels are often prone to noise, and the presence of such noise will degrade the performance of the resulting deep neural network (DNN) models. Much of the literatu…

cs.CV2020

Devil's in the Details: Aligning Visual Clues for Conditional Embedding in Person Re-Identification

Fufu Yu, Xinyang Jiang, Yifei Gong +5

Although Person Re-Identification has made impressive progress, difficult cases like occlusion, change of view-pointand similar clothing still bring great challenges. Besides overa…

cs.CV2019

Rethinking Temporal Fusion for Video-based Person Re-identification on Semantic and Time Aspect

Xinyang Jiang, Yifei Gong, Xiaowei Guo +5

Recently, the research interest of person re-identification (ReID) has gradually turned to video-based methods, which acquire a person representation by aggregating frame features…