1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2024
Explore the Limits of Omni-modal Pretraining at Scale
Yiyuan Zhang, Handong Li, Jing Liu +1
We propose to build omni-modal intelligence, which is capable of understanding any modality and learning universal representations. In specific, we propose a scalable pretraining p…
cs.CV2024
Multimodal Pathway: Improve Transformers with Irrelevant Data from Other Modalities
Yiyuan Zhang, Xiaohan Ding, Kaixiong Gong +3
We propose to improve transformers of a specific modality with irrelevant data from other modalities, e.g., improve an ImageNet model with audio or point cloud datasets. We would l…
cs.CV2023★ 1 cited
Towards Unified and Effective Domain Generalization
Yiyuan Zhang, Kaixiong Gong, Xiaohan Ding +4
We propose , a novel and fied framework for omain eneralization that is capable of significantly enhancing the out-of-distribu…