7 citations · 13 across the 6 of their papers we have counts for
1 paper · 1 filter
Julio Silva-Rodríguez, Jose Dolz, Ismail Ben Ayed
Vision-language pre-training has recently gained popularity as it allows learning rich feature representations using large-scale data sources. This paradigm has quickly made its wa…