1 paper
Giorgio Giannone, Ruoteng Li, Qianli Feng +3
Vision-language models (VLMs) have demonstrated remarkable potential in integrating visual and linguistic information, but their performance is often constrained by the need for ex…