2 papers
cs.CV2026
Variational Adapter for Cross-modal Similarity Representation
WenZhang Wei, Zhipeng Gui, Dehua Peng +2
The core of vision-language models lies in measuring cross-modal similarity within a unified representation space. However, most image-text matching or multi-class image classifica…
cs.LG2025
Sampling-enabled scalable manifold learning unveils the discriminative cluster structure of high-dimensional data
Dehua Peng, Zhipeng Gui, Wenzhang Wei +4
As a pivotal branch of machine learning, manifold learning uncovers the intrinsic low-dimensional structure within complex nonlinear manifolds in high-dimensional space for visuali…