1 citations · 1 across the 4 of their papers we have counts for
4 papers
Different Changes Require Different Reasoning: Change-Type-Specialized Experts for Robust Change Captioning
Jiyoung Park, InJae Oh, Jung Uk Kim
Change captioning is the task of generating natural language descriptions that explain the changes between a pair of images. Although different change types (e.g., color shifts, ob…
Learning to Orchestrate Vision Foundation Models for Multi-Task Dense Prediction
Donghyun Han, Yuseok Bae, Jung Uk Kim +1
Vision foundation models (VFMs) exhibit complementary strengths shaped by their pretraining objectives. Yet prevailing methods for multi-task dense prediction still train an entire…
Leveraging Textual Compositional Reasoning for Robust Change Captioning
Kyu Ri Park, Jiyoung Park, Seong Tae Kim +2
Change captioning aims to describe changes between a pair of images. However, existing works rely on visual features alone, which often fail to capture subtle but meaningful change…
Learning Trimodal Relation for Audio-Visual Question Answering with Missing Modality
Kyu Ri Park, Hong Joo Lee, Jung Uk Kim
Recent Audio-Visual Question Answering (AVQA) methods rely on complete visual and audio input to answer questions accurately. However, in real-world scenarios, issues such as devic…