1 citations · 1 across the 3 of their papers we have counts for
4 papers
Speaker conditioning of acoustic models using affine transformation for multi-speaker speech recognition
Midia Yousefi, John H. L. Hanse
This study addresses the problem of single-channel Automatic Speech Recognition of a target speaker within an overlap speech scenario. In the proposed method, the hidden representa…
Real-time Speaker counting in a cocktail party scenario using Attention-guided Convolutional Neural Network
Midia Yousefi, John H. L. Hansen
Most current speech technology systems are designed to operate well even in the presence of multiple active speakers. However, most solutions assume that the number of co-current s…
Frame-based overlapping speech detection using Convolutional Neural Networks
Midia Yousefi, John H. L. Hansen
Naturalistic speech recordings usually contain speech signals from multiple speakers. This phenomenon can degrade the performance of speech technologies due to the complexity of tr…
Probabilistic Permutation Invariant Training for Speech Separation
Midia Yousefi, Soheil Khorram, John H. L. Hansen
Single-microphone, speaker-independent speech separation is normally performed through two steps: (i) separating the specific speech sources, and (ii) determining the best output-l…