5 citations · 6 across the 6 of their papers we have counts for
Showing eess.ASShow all
2 papers · 1 filter
eess.AS2026
Do we really need Self-Attention for Streaming Automatic Speech Recognition?
Youness Dkhissi, Valentin Vielzeuf, Elys Allesiardo +1
Transformer-based architectures are the most used architectures in many deep learning fields like Natural Language Processing, Computer Vision or Speech processing. It may encourag…
eess.AS2021★ 5 cited
Efficient conformer: Progressive downsampling and grouped attention for automatic speech recognition
Maxime Burchi, Valentin Vielzeuf
The recently proposed Conformer architecture has shown state-of-the-art performances in Automatic Speech Recognition by combining convolution with attention to model both local and…