3 papers
cs.CV2025
Large Language Models Facilitate Vision Reflection in Image Classification
Guoyuan An, JaeYoon Kim, SungEui Yoon
This paper presents several novel findings on the explainability of vision reflection in large multimodal models (LMMs). First, we show that prompting an LMM to verify the predicti…
cs.RO2025
LangPert: Detecting and Handling Task-level Perturbations for Robust Object Rearrangement
Xu Yin, Min-Sung Yoon, Yuchi Huo +2
Task execution for object rearrangement could be challenged by Task-Level Perturbations (TLP), i.e., unexpected object additions, removals, and displacements that can disrupt under…
cs.CV2025
OpenSlot: Mixed Open-Set Recognition with Object-Centric Learning
Xu Yin, Fei Pan, Guoyuan An +3
Existing open-set recognition (OSR) studies typically assume that each image contains only one class label, with the unknown test set (negative) having a disjoint label space from…