2 citations · 2 across the 5 of their papers we have counts for
Showing eess.ASShow all
2 papers · 1 filter
eess.AS2023
Diagonal State Space Augmented Transformers for Speech Recognition
George Saon, Ankit Gupta, Xiaodong Cui
We improve on the popular conformer architecture by replacing the depthwise temporal convolutions with diagonal state space (DSS) models. DSS is a recently introduced variant of li…
eess.AS2022★ 2 cited
Extending RNN-T-based speech recognition systems with emotion and language classification
Zvi Kons, Hagai Aronowitz, Edmilson Morais +4
Speech transcription, emotion recognition, and language identification are usually considered to be three different tasks. Each one requires a different model with a different arch…