1 citations · 2 across the 3 of their papers we have counts for
1 paper · 1 filter
Bang Yang, Yong Dai, Xuxin Cheng +3
While vision-language pre-trained models (VL-PTMs) have advanced multimodal research in recent years, their mastery in a few languages like English restricts their applicability in…