1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Mirco Bonomo, Simone Bianco
Multimodal Large Language Models (MLLMs) have achieved notable performance in computer vision tasks that require reasoning across visual and textual modalities, yet their capabilit…