1 paper
Fuwen Luo, Chi Chen, Zihao Wan +11
Multimodal large language models (MLLMs) have demonstrated promising results in a variety of tasks that combine vision and language. As these models become more integral to researc…