25 citations · 29 across the 3 of their papers we have counts for
1 paper · 1 filter
Cheng Li, Jiexiong Liu, Yixuan Chen +1
In the field of video-language pretraining, existing models face numerous challenges in terms of inference efficiency and multimodal data processing. This paper proposes a KunLunBa…