17 citations · 73 across the 14 of their papers we have counts for
24 papers
Distilling Robust and Non-Robust Features in Adversarial Examples by Information Bottleneck
Junho Kim, Byung-Kwan Lee, Yong Man Ro
Adversarial examples, generated by carefully crafted perturbation, have attracted considerable attention in research fields. Recent works have argued that the existence of the robu…
Distinguishing Homophenes Using Multi-Head Visual-Audio Memory for Lip Reading
Minsu Kim, Jeong Hun Yeo, Yong Man Ro
Recognizing speech from silent lip movement, which is called lip reading, is a challenging task due to 1) the inherent information insufficiency of lip movement to fully represent…
Multi-modality Associative Bridging through Memory: Speech Sound Recollected from Face Video
Minsu Kim, Joanna Hong, Se Jin Park +1
In this paper, we introduce a novel audio-visual multi-modal bridging framework that can utilize both audio and visual information, even with uni-modal inputs. We exploit a memory…
Video Prediction Recalling Long-term Motion Context via Memory Alignment Learning
Sangmin Lee, Hak Gu Kim, Dae Hwi Choi +2
Our work addresses long-term motion context issues for predicting future frames. To predict the future precisely, it is required to capture which long-term motion context (e.g., wa…
Comprehensive Facial Expression Synthesis using Human-Interpretable Language
Joanna Hong, Jung Uk Kim, Sangmin Lee +1
Recent advances in facial expression synthesis have shown promising results using diverse expression representations including facial action units. Facial action units for an elabo…
Investigating Vulnerability to Adversarial Examples on Multimodal Data Fusion in Deep Learning
Youngjoon Yu, Hong Joo Lee, Byeong Cheon Kim +2
The success of multimodal data fusion in deep learning appears to be attributed to the use of complementary in-formation between multiple input data. Compared to their predictive p…