3 papers
cs.CV2026
Global Geometry Is Not Enough for Vision Representations
Jiwan Chung, Seon Joo Kim
A common assumption in representation learning is that globally well-distributed embeddings support robust and generalizable representations. This focus has shaped both training ob…
cs.CV2023
VISAGE: Video Instance Segmentation with Appearance-Guided Enhancement
Hanjung Kim, Jaehyun Kang, Miran Heo +3
In recent years, online Video Instance Segmentation (VIS) methods have shown remarkable advancement with their powerful query-based detectors. Utilizing the output queries of the d…
cs.CV2023
Leveraging Image Augmentation for Object Manipulation: Towards Interpretable Controllability in Object-Centric Learning
Jinwoo Kim, Janghyuk Choi, Jaehyun Kang +3
The binding problem in artificial neural networks is actively explored with the goal of achieving human-level recognition skills through the comprehension of the world in terms of…