4 citations · 4 across the 1 of their papers we have counts for
1 paper
Neelu Madan, Andreas Moegelmose, Rajat Modi +2
Video Foundation Models (ViFMs) aim to learn a general-purpose representation for various video understanding tasks. Leveraging large-scale datasets and powerful models, ViFMs achi…