From the 1 of 15 linked papers with an AI index.
1 paper · 1 filter
Wanpeng Zhang, Zilong Xie, Yicheng Feng +4
Multimodal Large Language Models have made significant strides in integrating visual and textual information, yet they often struggle with effectively aligning these modalities. We…