20 citations · 23 across the 6 of their papers we have counts for
1 paper · 1 filter
Wele Gedara Chaminda Bandara, Naman Patel, Ali Gholami +3
Masked Autoencoders (MAEs) learn generalizable representations for image, text, audio, video, etc., by reconstructing masked input data from tokens of the visible data. Current MAE…