10 citations · 10 across the 1 of their papers we have counts for
1 paper
Chi Chen, Ruoyu Qin, Fuwen Luo +4
Recently, Multimodal Large Language Models (MLLMs) that enable Large Language Models (LLMs) to interpret images through visual instruction tuning have achieved significant success.…