activity
20182020
most citedA Unified Deep Learning Framework for Short-Duration Speaker Verification in Adverse Environments

20 citations · 28 across the 4 of their papers we have counts for

collaborators

9 papers

eess.AS202020 cited

A Unified Deep Learning Framework for Short-Duration Speaker Verification in Adverse Environments

Youngmoon Jung, Yeunju Choi, Hyungjun Lim +1

Speaker verification (SV) has recently attracted considerable research interest due to the growing popularity of virtual assistants. At the same time, there is an increasing requir…

eess.AS2020

Multi-Task Network for Noise-Robust Keyword Spotting and Speaker Verification using CTC-based Soft VAD and Global Query Attention

Myunghun Jung, Youngmoon Jung, Jahyun Goo +1

Keyword spotting (KWS) and speaker verification (SV) have been studied independently although it is known that acoustic and speaker domains are complementary. In this paper, we pro…

eess.AS2020

Meta-Learning for Short Utterance Speaker Recognition with Imbalance Length Pairs

Seong Min Kye, Youngmoon Jung, Hae Beom Lee +2

In practical settings, a speaker recognition system needs to identify a speaker given a short utterance, while the enrollment utterance may be relatively long. However, existing sp…

eess.AS2020

Dual Attention in Time and Frequency Domain for Voice Activity Detection

Joohyung Lee, Youngmoon Jung, Hoirin Kim

Voice activity detection (VAD) is a challenging task in low signal-to-noise ratio (SNR) environment, especially in non-stationary noise. To deal with this issue, we propose a novel…

cs.LG2020

Meta-Learned Confidence for Few-shot Learning

Seong Min Kye, Hae Beom Lee, Hoirin Kim +1

Transductive inference is an effective means of tackling the data deficiency problem in few-shot learning settings. A popular transductive inference technique for few-shot metric-b…

eess.AS20191 cited

Additional Shared Decoder on Siamese Multi-view Encoders for Learning Acoustic Word Embeddings

Myunghun Jung, Hyungjun Lim, Jahyun Goo +2

Acoustic word embeddings --- fixed-dimensional vector representations of arbitrary-length words --- have attracted increasing interest in query-by-example spoken term detection. Re…