1 citations · 1 across the 1 of their papers we have counts for
1 paper
Sungdong Kim, Jin-Hwa Kim, Jiyoung Lee +1
Efficient video-language modeling should consider the computational cost because of a large, sometimes intractable, number of video frames. Parametric approaches such as the attent…