1 paper
Yixuan Du, Chenxiao Yu, Haoyan Xu +3
Vision-Language Models (VLMs) integrate visual and textual knowledge into unified representations that increasingly underpin modern retrieval and recommendation systems. However, i…