From the 1 of 7 linked papers with an AI index.
1 paper · 1 filter
Xiangkai Ma, Han Zhang, Wenzhong Li +1
Large Multimodal Models (LMMs) have achieved remarkable progress in aligning and generating content across text and image modalities. However, the potential of using non-visual, co…