3 citations · 4 across the 4 of their papers we have counts for
4 papers · 1 filter
Do You Listen with One or Two Microphones? A Unified ASR Model for Single and Multi-Channel Audio
Gokce Keskin, Minhua Wu, Brian King +5
Automatic speech recognition (ASR) models are typically designed to operate on a single input data type, e.g. a single or multi-channel audio streamed from a device. This design de…
Attention-based Neural Beamforming Layers for Multi-channel Speech Recognition
Bhargav Pulugundla, Yang Gao, Brian King +5
Attention-based beamformers have recently been shown to be effective for multi-channel speech recognition. However, they are less capable at capturing local information. In this wo…
End-to-End Multi-Channel Transformer for Speech Recognition
Feng-Ju Chang, Martin Radfar, Athanasios Mouchtaris +2
Transformers are powerful neural architectures that allow integrating different modalities using attention mechanisms. In this paper, we leverage the neural transformer architectur…
Multi-view Frequency LSTM: An Efficient Frontend for Automatic Speech Recognition
Maarten Van Segbroeck, Harish Mallidih, Brian King +3
Acoustic models in real-time speech recognition systems typically stack multiple unidirectional LSTM layers to process the acoustic frames over time. Performance improvements over…