2 papers
cs.SD2026
CMI-RewardBench: Evaluating Music Reward Models with Compositional Multimodal Instruction
Yinghao Ma, Haiwen Xia, Hewei Gao +9
While music generation models have evolved to handle complex multimodal inputs mixing text, lyrics, and reference audio, evaluation mechanisms have lagged behind. In this paper, we…
cs.LG2025
FedNano: Toward Lightweight Federated Tuning for Pretrained Multimodal Large Language Models
Yao Zhang, Hewei Gao, Haokun Chen +3
Multimodal Large Language Models (MLLMs) excel in tasks like multimodal reasoning and cross-modal retrieval but face deployment challenges in real-world scenarios due to distribute…