1 paper
Sherzod Hakimov, Yerkezhan Abdullayeva, Kushal Koshti +4
While the situation has improved for text-only models, it again seems to be the case currently that multimodal (text and image) models develop faster than ways to evaluate them. In…