25 citations · 43 across the 2 of their papers we have counts for
3 papers
eess.AS2020★ 25 cited
Universal MelGAN: A Robust Neural Vocoder for High-Fidelity Waveform Generation in Multiple Domains
Won Jang, Dan Lim, Jaesam Yoon
We propose Universal MelGAN, a vocoder that synthesizes high-fidelity speech in multiple domains. To preserve sound quality when the MelGAN-based structure is trained with a datase…
cs.CV2020
Real-time Mask Detection on Google Edge TPU
Keondo Park, Wonyoung Jang, Woochul Lee +4
After the COVID-19 outbreak, it has become important to automatically detect whether people are wearing masks in order to reduce risk of front-line workers. In addition, processing…
eess.AS2020★ 18 cited
JDI-T: Jointly trained Duration Informed Transformer for Text-To-Speech without Explicit Alignment
Dan Lim, Won Jang, Gyeonghwan O +3
We propose Jointly trained Duration Informed Transformer (JDI-T), a feed-forward Transformer with a duration predictor jointly trained without explicit alignments in order to gener…