1 paper · 1 filter
Chin Ting Hsu, Yu-Syuan Xu, Ling Zou +2
Although multimodal Large Language Models (MLLMs) excel in diverse tasks, their scalability remains limited by the memory and computational overhead of KV cache storage. Recent KV…