22 citations · 22 across the 2 of their papers we have counts for
2 papers
eess.AS2023★ 22 cited
NeXt-TDNN: Modernizing Multi-Scale Temporal Convolution Backbone for Speaker Verification
Hyun-Jun Heo, Ui-Hyeop Shin, Ran Lee +2
In speaker verification, ECAPA-TDNN has shown remarkable improvement by utilizing one-dimensional(1D) Res2Net block and squeeze-and-excitation(SE) module, along with multi-layer fe…
cs.LG2023
Unsupervised Speech Representation Pooling Using Vector Quantization
Jeongkyun Park, Kwanghee Choi, Hyunjun Heo +1
With the advent of general-purpose speech representations from large-scale self-supervised models, applying a single model to multiple downstream tasks is becoming a de-facto appro…