30 citations · 39 across the 4 of their papers we have counts for
1 paper · 1 filter
Han Fang, Zhifei Yang, Xianghao Zang +2
Recently, masked video modeling has been widely explored and significantly improved the model's understanding ability of visual regions at a local level. However, existing methods…