1 paper · 1 filter
Chunyi Li, Xiaozhe Li, Zicheng Zhang +8
With the emergence of Multimodal Large Language Models (MLLMs), hundreds of benchmarks have been developed to ensure the reliability of MLLMs in downstream tasks. However, the eval…