2 citations · 2 across the 1 of their papers we have counts for
1 paper
Ya-Qi Yu, Minghui Liao, Jihao Wu +3
Multimodal Large Language Models (MLLMs) have shown impressive results on various multimodal tasks. However, most existing MLLMs are not well suited for document-oriented tasks, wh…