5 citations · 8 across the 3 of their papers we have counts for
5 papers · 1 filter
Zero-Shot Vision-and-Language Navigation with Collision Mitigation in Continuous Environment
Seongjun Jeong, Gi-Cheon Kang, Joochan Kim +1
We propose the zero-shot Vision-and-Language Navigation with Collision Mitigation (VLN-CM), which takes these considerations. VLN-CM is composed of four modules and predicts the di…
Continual Vision-and-Language Navigation
Seongjun Jeong, Gi-Cheon Kang, Seongho Choi +2
Developing Vision-and-Language Navigation (VLN) agents typically assumes a \textit{train-once-deploy-once} strategy, which is unrealistic as deployed agents continually encounter n…
Attend What You Need: Motion-Appearance Synergistic Networks for Video Question Answering
Ahjeong Seo, Gi-Cheon Kang, Joonhan Park +1
Video Question Answering is a task which requires an AI agent to answer questions grounded in video. This task entails three key challenges: (1) understand the intention of various…
Label Propagation Adaptive Resonance Theory for Semi-supervised Continuous Learning
Taehyeong Kim, Injune Hwang, Gi-Cheon Kang +3
Semi-supervised learning and continuous learning are fundamental paradigms for human-level intelligence. To deal with real-world problems where labels are rarely given and the oppo…
Dual Attention Networks for Visual Reference Resolution in Visual Dialog
Gi-Cheon Kang, Jaeseo Lim, Byoung-Tak Zhang
Visual dialog (VisDial) is a task which requires an AI agent to answer a series of questions grounded in an image. Unlike in visual question answering (VQA), the series of question…