1 citations · 1 across the 1 of their papers we have counts for
1 paper
Xiangyu Zhao, Xiangtai Li, Haodong Duan +4
Multi-modal large language models (MLLMs) have made significant strides in various visual understanding tasks. However, the majority of these models are constrained to process low-…