49 citations · 70 across the 11 of their papers we have counts for
14 papers
Speaker Reinforcement Using Target Source Extraction for Robust Automatic Speech Recognition
Catalin Zorila, Rama Doddipatla
Improving the accuracy of single-channel automatic speech recognition (ASR) in noisy conditions is challenging. Strong speech enhancement front-ends are available, however, they ty…
Dialogue Strategy Adaptation to New Action Sets Using Multi-dimensional Modelling
Simon Keizer, Norbert Braunschweiler, Svetlana Stoyanchev +1
A major bottleneck for building statistical spoken dialogue systems for new domains and applications is the need for large amounts of training data. To address this problem, we ado…
Transformer-based Streaming ASR with Cumulative Attention
Mohan Li, Shucong Zhang, Catalin Zorila +1
In this paper, we propose an online attention mechanism, known as cumulative attention (CA), for streaming Transformer-based automatic speech recognition (ASR). Inspired by monoton…
A study on cross-corpus speech emotion recognition and data augmentation
Norbert Braunschweiler, Rama Doddipatla, Simon Keizer +1
Models that can handle a wide range of speakers and acoustic conditions are essential in speech emotion recognition (SER). Often, these models tend to show mixed results when prese…
Towards Handling Unconstrained User Preferences in Dialogue
Suraj Pandey, Svetlana Stoyanchev, Rama Doddipatla
A user input to a schema-driven dialogue information navigation system, such as venue search, is typically constrained by the underlying database which restricts the user to specif…
Teacher-Student MixIT for Unsupervised and Semi-supervised Speech Separation
Jisi Zhang, Catalin Zorila, Rama Doddipatla +1
In this paper, we introduce a novel semi-supervised learning framework for end-to-end speech separation. The proposed method first uses mixtures of unseparated sources and the mixt…