1 paper
Sayna Ebrahimi, Sercan O. Arik, Yihe Dong +1
Multimodal large-scale pretraining has shown impressive performance for unstructured data such as language and image. However, a prevalent real-world scenario involves structured d…