Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
CPCL: Cross-Modal Prototypical Contrastive Learning for Weakly Supervised Text-based Person Retrieval
Xinpeng Zhao, Yanwei Zheng, Chuanlin Lan +4
Weakly supervised text-based person retrieval seeks to retrieve images of a target person using textual descriptions, without relying on identity annotations and is more challengin…
cs.CV2024
Leveraging Unknown Objects to Construct Labeled-Unlabeled Meta-Relationships for Zero-Shot Object Navigation
Yanwei Zheng, Changrui Li, Chuanlin Lan +5
Zero-shot object navigation (ZSON) addresses situation where an agent navigates to an unseen object that does not present in the training set. Previous works mainly train agent usi…
cs.CV2024
Temporal-Spatial Object Relations Modeling for Vision-and-Language Navigation
Bowen Huang, Yanwei Zheng, Chuanlin Lan +3
Vision-and-Language Navigation (VLN) is a challenging task where an agent is required to navigate to a natural language described location via vision observations. The navigation a…