6 citations · 11 across the 3 of their papers we have counts for
1 paper · 2 filters
Xiwen Chen, Yen-Chieh Lien, Susan Liu +4
The rapid growth of e-commerce requires robust multimodal representations that capture diverse signals from user-generated listings. Existing vision-language models (VLMs) typicall…