2 papers
cs.CL2025
LFAR: Accounting for Layerwise Dynamics to Improve Multimodal Adaptation of Language Models
Santiago Cuervo, Adel Moumen, Yanis Labrak +5
Text-pretrained language models (LMs) encode rich world knowledge, but adapting them to process and generate perceptual modalities such as audio and images while effectively levera…
cs.LG2025
Aligning Multimodal Representations through an Information Bottleneck
Antonio Almudévar, José Miguel Hernández-Lobato, Sameer Khurana +2
Contrastive losses have been extensively used as a tool for multimodal representation learning. However, it has been empirically observed that their use is not effective to learn a…