53 citations · 97 across the 4 of their papers we have counts for
1 paper · 1 filter
Lisa Anne Hendricks, John Mellor, Rosalia Schneider +2
Recently multimodal transformer models have gained popularity because their performance on language and vision tasks suggest they learn rich visual-linguistic representations. Focu…