1 citations · 1 across the 1 of their papers we have counts for
1 paper
Shihab Aaqil Ahamed, Malitha Gunawardhana, Liel David +3
Current video-based Masked Autoencoders (MAEs) primarily focus on learning effective spatiotemporal representations from a visual perspective, which may lead the model to prioritiz…