1 paper
Boran Zhao, Hetian Liu, Zhenxian Hu +3
The training of large multimodal models fundamentally relies on massive image-text datasets, which inevitably incur prohibitive computational overhead. Dataset selection offers a p…