2 papers
cs.LG2026
RetentiveKV: State-Space Memory for Uncertainty-Aware Multimodal KV Cache Eviction
Sihao Liu, YuFan Xiong, Zhonghua Jiang +2
Multimodal Large Language Models face severe challenges in computational efficiency and memory consumption due to the substantial expansion of the visual KV cache when processing l…
cs.CV2024
Octavius: Mitigating Task Interference in MLLMs via LoRA-MoE
Zeren Chen, Ziqin Wang, Zhen Wang +7
Recent studies have demonstrated Large Language Models (LLMs) can extend their zero-shot generalization capabilities to multimodal learning through instruction tuning. As more moda…