3 papers
cs.CV2025
Toward Visual Grounding: A Survey
Linhui Xiao, Xiaoshan Yang, Xiangyuan Lan +2
Visual Grounding, also known as Referring Expression Comprehension and Phrase Grounding, aims to ground the specific region(s) within the image(s) based on the given expression tex…
cs.CV2025
SelaVPR++: Towards Seamless Adaptation of Foundation Models for Efficient Place Recognition
Feng Lu, Tong Jin, Xiangyuan Lan +4
Recent studies show that the visual place recognition (VPR) method using pre-trained visual foundation models can achieve promising performance. In our previous work, we propose a…
cs.CV2025
Limb-Aware Virtual Try-On Network with Progressive Clothing Warping
Shengping Zhang, Xiaoyu Han, Weigang Zhang +3
Image-based virtual try-on aims to transfer an in-shop clothing image to a person image. Most existing methods adopt a single global deformation to perform clothing warping directl…