3 citations · 3 across the 1 of their papers we have counts for
1 paper
Shuyuan Tu, Qi Dai, Zuxuan Wu +3
Contrastive language-image pretraining (CLIP) has demonstrated remarkable success in various image tasks. However, how to extend CLIP with effective temporal modeling is still an o…