1 paper
Bozhi Luan, Hao Feng, Hong Chen +3
The advent of Large Multimodal Models (LMMs) has sparked a surge in research aimed at harnessing their remarkable reasoning abilities. However, for understanding text-rich images,…