Showing eess.ASShow all
2 papers · 1 filter
eess.AS2025
Hybrid Decoding: Rapid Pass and Selective Detailed Correction for Sequence Models
Yunkyu Lim, Jihwan Park, Hyung Yong Kim +2
Recently, Transformer-based encoder-decoder models have demonstrated strong performance in multilingual speech recognition. However, the decoder's autoregressive nature and large s…
eess.AS2024
Bridging the Gap between Audio and Text using Parallel-attention for User-defined Keyword Spotting
Youkyum Kim, Jaemin Jung, Jihwan Park +2
This paper proposes a novel user-defined keyword spotting framework that accurately detects audio keywords based on text enrollment. Since audio data possesses additional acoustic…