1 paper
Xi Leng, Xinhong Ma, Ziqiang Dong +4
Large multimodal models (LMMs) inherit the self-attention mechanism of pretrained language backbones, yet standard attention can exhibit suboptimal allocation, including cross-moda…