Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
SAME: Stabilized Mixture-of-Experts for Multimodal Continual Instruction Tuning
Zhen-Hao Xie, Jun-Tao Tang, Yu-Cheng Shi +3
Multimodal Large Language Models (MLLMs) achieve strong performance through instruction tuning, but real-world deployment requires them to continually expand their capabilities, ma…
cs.LG2025
FreqCa: Accelerating Diffusion Models via Frequency-Aware Caching
Jiacheng Liu, Peiliang Cai, Qinming Zhou +9
The application of diffusion transformers is suffering from their significant inference costs. Recently, feature caching has been proposed to solve this problem by reusing features…