1 paper
Soyeon Caren Han, Feiqi Cao, Josiah Poon +1
This tutorial explores recent advancements in multimodal pretrained and large models, capable of integrating and processing diverse data forms such as text, images, audio, and vide…