most citedVideo-LaVIT: Unified Video-Language Pre-training with Decoupled Visual-Motional Tokenization

5 citations · 6 across the 6 of their papers we have counts for

collaborators
Showing cs.CLShow all

Nothing from them under that filter.

Their other years and fields are still on the left.