47 citations · 289 across the 44 of their papers we have counts for
1 paper · 2 filters
Weijie Su, Xizhou Zhu, Yue Cao +4
We introduce a new pre-trainable generic representation for visual-linguistic tasks, called Visual-Linguistic BERT (VL-BERT for short). VL-BERT adopts the simple yet powerful Trans…