1 paper · 1 filter
Min-Han Shih, Yu-Hsin Wu, Yu-Wei Chen
We propose a dedicated multimodal Judge Model designed to provide reliable, explainable evaluation across a diverse suite of tasks. Our benchmark spans text, audio, image, and vide…