1 paper
Tom Hodemon, Mohamed Chaouch, Aboubacar Tuo +1
Multimodal Large Language Models (MLLMs) have achieved remarkable success in Visual Question Answering (VQA), yet their "black-box" nature hinders deployment in critical domains. G…