5 papers
Mixture of Enhanced-View Experts for Multi-Query Vehicle ReID and A Large-Scale Benchmark
Aihua Zheng, Jie Zhen, Chenglong Li +2
Multi-query vehicle ReID aims to leverage complementary information from diverse views for robust feature learning. However, current methods suffer from simplistic feature fusion a…
T2I-VeRW: Part-level Fine-grained Perception for Text-to-Image Vehicle Retrieval
Xiao Wang, Ziwen Wang, Weizhe Kong +5
Vehicle Re-identification (Re-ID) aims to retrieve the most similar image to a given query from images captured by non-overlapping cameras. Extending vehicle Re-ID from image-only…
RefAerial: A Benchmark and Approach for Referring Detection in Aerial Images
Guyue Hu, Hao Song, Yuxing Tong +5
Referring detection refers to locate the target referred by natural languages, which has recently attracted growing research interests. However, existing datasets are limited to gr…
SequencePAR: Understanding Pedestrian Attributes via A Sequence Generation Paradigm
Jiandong Jin, Xiao Wang, Yin Lin +4
Current pedestrian attribute recognition (PAR) algorithms use multi-label or multi-task learning frameworks with specific classification heads. These models often struggle with imb…
Arbitrary Talking Face Generation via Attentional Audio-Visual Coherence Learning
Hao Zhu, Huaibo Huang, Yi Li +2
Talking face generation aims to synthesize a face video with precise lip synchronization as well as a smooth transition of facial motion over the entire video via the given speech…