2 citations · 2 across the 1 of their papers we have counts for
1 paper · 1 filter
Xingjian Diao, Ming Cheng, Shitong Cheng
Learning high-quality video representation has shown significant applications in computer vision and remains challenging. Previous work based on mask autoencoders such as ImageMAE…