1 paper
Tianyu Guo, Tianming Xu, Xianjie Chen +3
Large multimodal models (LMMs) typically employ an encoding module to transform multimodal data inputs into embeddings, which are then fed to language models for further processing…