49 citations · 133 across the 19 of their papers we have counts for
8 papers · 1 filter
HyperCon: Image-To-Video Model Transfer for Video-To-Video Translation Tasks
Ryan Szeto, Mostafa El-Khamy, Jungwon Lee +1
Video-to-video translation is more difficult than image-to-image translation due to the temporal consistency problem that, if unaddressed, leads to distracting flickering effects.…
End-to-End Multi-Task Denoising for the Joint Optimization of Perceptual Speech Metrics
Jaeyoung Kim, Mostafa El-Khamy, Jungwon Lee
Although supervised learning based on a deep neural network has recently achieved substantial improvement on speech enhancement, the existing schemes have either of two critical is…
T-GSA: Transformer with Gaussian-weighted self-attention for speech enhancement
Jaeyoung Kim, Mostafa El-Khamy, Jungwon Lee
Transformer neural networks (TNN) demonstrated state-of-art performance on many natural language processing (NLP) tasks, replacing recurrent neural networks (RNNs), such as LSTMs o…
Variable Rate Deep Image Compression With a Conditional Autoencoder
Yoojin Choi, Mostafa El-Khamy, Jungwon Lee
In this paper, we propose a novel variable-rate learned image compression framework with a conditional autoencoder. Previous learning-based image compression methods mostly require…
TW-SMNet: Deep Multitask Learning of Tele-Wide Stereo Matching
Mostafa El-Khamy, Haoyu Ren, Xianzhi Du +1
In this paper, we introduce the problem of estimating the real world depth of elements in a scene captured by two cameras with different field of views, where the first field of vi…
Deep Robust Single Image Depth Estimation Neural Network Using Scene Understanding
Haoyu Ren, Mostafa El-khamy, Jungwon Lee
Single image depth estimation (SIDE) plays a crucial role in 3D computer vision. In this paper, we propose a two-stage robust SIDE framework that can perform blind SIDE for both in…