1 paper
Hyojin Bahng, Caroline Chan, Fredo Durand +1
Measuring alignment between language and vision is a fundamental challenge, especially as multimodal data becomes increasingly detailed and complex. Existing methods often rely on…