18 citations · 19 across the 3 of their papers we have counts for
3 papers
cs.CV2022★ 1 cited
Panoramic Vision Transformer for Saliency Detection in 360° Videos
Heeseung Yun, Sehun Lee, Gunhee Kim
360 video saliency detection is one of the challenging benchmarks for 360 video understanding since non-negligible distortion and discontinuity occur in the project…
cs.CL2022★ 18 cited
Multimodal Knowledge Alignment with Reinforcement Learning
Youngjae Yu, Jiwan Chung, Heeseung Yun +8
Large language models readily adapt to novel settings, even without task-specific training data. Can their zero-shot capacity be extended to multimodal inputs? In this work, we pro…
cs.CV2021
Pano-AVQA: Grounded Audio-Visual Question Answering on 360 Videos
Heeseung Yun, Youngjae Yu, Wonsuk Yang +2
360 videos convey holistic views for the surroundings of a scene. It provides audio-visual cues beyond pre-determined normal field of views and displays distinctive spatial…