3 papers
cs.AI2026
BUZZY: Contrastive Scoring to Mitigate Text-Induced Bias in Multimodal Multiple-Choice QA
Taeyun Roh, Suhyeong Park, Eun-yeong Jo +5
Multimodal multiple-choice question answering (MCQA) provides a standardized and objectively measurable setting for evaluating vision-language models (VLMs). However, because the M…
cs.CV2025
Few-Shot Pattern Detection via Template Matching and Regression
Eunchan Jo, Dahyun Kang, Sanghyun Kim +2
We address the problem of few-shot pattern detection, which aims to detect all instances of a given pattern, typically represented by a few exemplars, from an input image. Although…
cs.CV2025
Memory-Modular Classification: Learning to Generalize with Memory Replacement
Dahyun Kang, Ahmet Iscen, Eunchan Jo +3
We propose a novel memory-modular learner for image classification that separates knowledge memorization from reasoning. Our model enables effective generalization to new classes b…