1 citations · 1 across the 1 of their papers we have counts for
1 paper
Bang Yang, Yong Dai, Xuxin Cheng +3
While vision-language pre-trained models (VL-PTMs) have advanced multimodal research in recent years, their mastery in a few languages like English restricts their applicability in…