2 papers
cs.CL2022
End-to-end model for named entity recognition from speech without paired training data
Salima Mdhaffar, Jarod Duret, Titouan Parcollet +1
Recent works showed that end-to-end neural approaches tend to become very popular for spoken language understanding (SLU). Through the term end-to-end, one considers the use of a s…
eess.AS2021
Study on the temporal pooling used in deep neural networks for speaker verification
Mickael Rouvier, Pierre-Michel Bousquet, Jarod Duret
The x-vector architecture has recently achieved state-of-the-art results on the speaker verification task. This architecture incorporates a central layer, referred to as temporal p…