1 paper · 1 filter
Hyojin Bahng, Caroline Chan, Fredo Durand +1
Measuring alignment between language and vision is a fundamental challenge, especially as multimodal data becomes increasingly detailed and complex. Existing methods often rely on…