2 citations · 2 across the 4 of their papers we have counts for
1 paper · 1 filter
Leonardo Brusini, Cristian Sbrolli, Eugenio Lomurno +2
Synthetic data offers a scalable solution for vision-language pre-training, yet current state-of-the-art methods typically rely on scaling up a single generative backbone, which in…