1 paper
Lele Cao, Valentin Buchner, Zineb Senane +1
Multimodal Large Language Models (MLLMs) are typically assessed using expensive annotated multimodal benchmarks, which often lag behind the rapidly evolving demands of MLLM evaluat…