5 papers
VTCBench: Can Vision-Language Models Understand Long Context with Vision-Text Compression?
Hongbo Zhao, Meng Wang, Fei Zhu +5
The computational and memory overheads associated with expanding the context window of LLMs severely limit their scalability. A noteworthy solution is vision-text compression (VTC)…
MLLM-CL: Continual Learning for Multimodal Large Language Models
Hongbo Zhao, Fei Zhu, Haiyang Guo +4
Recent Multimodal Large Language Models (MLLMs) excel in vision-language understanding but face challenges in adapting to dynamic real-world scenarios that require continuous integ…
Semi-parametric Memory Consolidation: Towards Brain-like Deep Continual Learning
Geng Liu, Fei Zhu, Rong Feng +4
Humans and most animals inherently possess a distinctive capacity to continually acquire novel experiences and accumulate worldly knowledge over time. This ability, termed continua…
Pareto Continual Learning: Preference-Conditioned Learning and Adaption for Dynamic Stability-Plasticity Trade-off
Song Lai, Zhe Zhao, Fei Zhu +3
Continual learning aims to learn multiple tasks sequentially. A key challenge in continual learning is balancing between two objectives: retaining knowledge from old tasks (stabili…
Practical Continual Forgetting for Pre-trained Vision Models
Hongbo Zhao, Fei Zhu, Bolin Ni +3
For privacy and security concerns, the need to erase unwanted information from pre-trained vision models is becoming evident nowadays. In real-world scenarios, erasure requests ori…