11 papers
Resolution as a Direction: Vector-Panning Feature Alignment for Cross-Resolution Re-Identification
Zanwu Liu, Chao Yuan, Bo Li +2
Cross-resolution person re-identification (CR-ReID) remains challenging in practical surveillance, where camera quality and capture distance lead to substantial resolution gaps bet…
Towards Robust Text-to-Image Person Retrieval: Multi-View Reformulation for Semantic Compensation
Chao Yuan, Yujian Zhao, Haoxuan Xu +1
In text-to-image person retrieval tasks, the diversity of natural language expressions and the implicitness of visual semantics often lead to the problem of Expression Drift, where…
UniComp: Rethinking Video Compression Through Informational Uniqueness
Chao Yuan, Shimin Chen, Minliang Lin +3
Distinct from attention-based compression methods, this paper presents an information uniqueness driven video compression framework, termed UniComp, which aims to maximize the info…
ERNIE 5.0 Technical Report
Haifeng Wang, Hua Wu, Tian Wu +432
In this report, we introduce ERNIE 5.0, a natively autoregressive foundation model desinged for unified multimodal understanding and generation across text, image, video, and audio…
OmniPerson: Unified Identity-Preserving Pedestrian Generation
Changxiao Ma, Chao Yuan, Xincheng Shi +6
Person re-identification (ReID) suffers from a lack of large-scale high-quality training data due to challenges in data privacy and annotation costs. While previous approaches have…
Modality-Transition Representation Learning for Visible-Infrared Person Re-Identification
Chao Yuan, Zanwu Liu, Guiwei Zhang +4
Visible-infrared person re-identification (VI-ReID) technique could associate the pedestrian images across visible and infrared modalities in the practical scenarios of background…