3 papers
cs.CV2026
SAM3Dual: A 3rd Place Solution to the MOSEv2 Track, 8th LSVOS Challenge
JeongRae Kim, Chaehyun Kim, Changwon Lim
We present SAM3Dual, our third-place solution to the MOSEv2 track of the 8th Large-scale Video Object Segmentation (LSVOS) Challenge at ECCV 2026. SAM3Dual is a training-free infer…
cs.CV2026
Continuity-Driven Representation Learning for Industrial Defect Detection
Minjong Kim, Hyun Jun Kim, Jeongrae Kim +2
Industrial defect detection differs from natural-image object detection because inspection images are captured under controlled conditions and contain large normal-dominant regions…
cs.SD2025
AISTAT lab system for DCASE2025 Task6: Language-based audio retrieval
Hyun Jun Kim, Hyeong Yong Choi, Changwon Lim
This report presents the AISTAT team's submission to the language-based audio retrieval task in DCASE 2025 Task 6. Our proposed system employs dual encoder architecture, where audi…