32 citations · 73 across the 12 of their papers we have counts for
Showing 2021Show all
2 papers · 1 filter
cs.CV2021
Spatially Consistent Representation Learning
Byungseok Roh, Wuhyun Shin, Ildoo Kim +1
Self-supervised learning has been widely used to obtain transferrable representations from unlabeled images. Especially, recent contrastive learning methods have shown impressive p…
stat.ML2021
ViLT: Vision-and-Language Transformer Without Convolution or Region Supervision
Wonjae Kim, Bokyung Son, Ildoo Kim
Vision-and-Language Pre-training (VLP) has improved performance on various joint vision-and-language downstream tasks. Current approaches to VLP heavily rely on image feature extra…