3 papers
cs.CV2026
Mitigating Cross-Image Information Leakage in Multi-Image Understanding with Large Vision-Language Models
Yeji Park, Minyoung Lee, Sanghyuk Chun +1
Large Vision-Language Models (LVLMs) exhibit strong performance on single-image tasks. However, their performance degrades significantly when handling multi-image inputs. While thi…
cs.CV2026
Enhancing Multi-Image Understanding through Delimiter Token Scaling
Minyoung Lee, Yeji Park, Dongjun Hwang +3
Large Vision-Language Models (LVLMs) achieve strong performance on single-image tasks, but their performance declines when multiple images are provided as input. One major reason i…
cs.CV2025
OVS Meets Continual Learning: Towards Sustainable Open-Vocabulary Segmentation
Dongjun Hwang, Yejin Kim, Minyoung Lee +2
Open-Vocabulary Segmentation (OVS) aims to segment classes that are not present in the training dataset. However, most existing studies assume that the training data is fixed in ad…