1 paper
Keng Ji Chow, Samson Tan, Min-Yen Kan
Numerous visio-linguistic (V+L) representation learning methods have been developed, yet existing datasets do not adequately evaluate the extent to which they represent visual and…