2 citations · 3 across the 8 of their papers we have counts for
4 papers · 1 filter
Bayesian Learning for Domain-Invariant Speaker Verification and Anti-Spoofing
Jin Li, Man-Wai Mak, Johan Rohdin +2
The performance of automatic speaker verification (ASV) and anti-spoofing drops seriously under real-world domain mismatch conditions. The relaxed instance frequency-wise normaliza…
Blind Signal Dereverberation for Machine Speech Recognition
Samik Sadhu, Hynek Hermansky
We present a method to remove unknown convolutive noise introduced to speech by reverberations of recording environments, utilizing some amount of training speech data from the rev…
Radically Old Way of Computing Spectra: Applications in End-to-End ASR
Samik Sadhu, Hynek Hermansky
We propose a technique to compute spectrograms using Frequency Domain Linear Prediction (FDLP) that uses all-pole models to fit the squared Hilbert envelope of speech in different…
Multi-Stream End-to-End Speech Recognition
Ruizhi Li, Xiaofei Wang, Sri Harish Mallidi +3
Attention-based methods and Connectionist Temporal Classification (CTC) network have been promising research directions for end-to-end (E2E) Automatic Speech Recognition (ASR). The…