1 citations · 2 across the 4 of their papers we have counts for
1 paper · 2 filters
Boshen Xu, Yuting Mei, Xinbi Liu +2
Egocentric video-language pretraining has significantly advanced video representation learning. Humans perceive and interact with a fully 3D world, developing spatial awareness tha…