cross-modal alignment 1layer-wise merging 1model merging 1multimodal pretraining 1parameter interference 1vision-language models 1
From the 1 of 5 linked papers with an AI index.
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
A Step Toward Federated Pretraining of Multimodal Large Language Models
Baochen Xiong, Yifan Xu, Xiaoshan Yang +3
The rapid evolution of Multimodal Large Language Models (MLLMs) is bottlenecked by the saturation of high-quality public data, while vast amounts of diverse multimodal data remain…
cs.LG2025
Pilot: Building the Federated Multimodal Instruction Tuning Framework
Baochen Xiong, Xiaoshan Yang, Yaguang Song +2
In this paper, we explore a novel federated multimodal instruction tuning task(FedMIT), which is significant for collaboratively fine-tuning MLLMs on different types of multimodal…