10 citations · 10 across the 1 of their papers we have counts for
1 paper
Roman Bachmann, David Mizrahi, Andrei Atanov +1
We propose a pre-training strategy called Multi-modal Multi-task Masked Autoencoders (MultiMAE). It differs from standard Masked Autoencoding in two key aspects: I) it can optional…