1 citations · 1 across the 1 of their papers we have counts for
1 paper
Zekai Chen, Arda Pekis, Kevin Brown
Multi-modal learning has significantly advanced generative AI, especially in vision-language modeling. Innovations like GPT-4V and open-source projects such as LLaVA have enabled r…