6 citations · 21 across the 24 of their papers we have counts for
8 papers · 1 filter
Analysis of DNN Speech Signal Enhancement for Robust Speaker Recognition
Ondrej Novotny, Oldrich Plchot, Ondrej Glembek +2
In this work, we present an analysis of a DNN-based autoencoder for speech enhancement, dereverberation and denoising. The target application is a robust speaker verification (SV)…
Analysis of Multilingual Sequence-to-Sequence speech recognition systems
Martin Karafiát, Murali Karthick Baskar, Shinji Watanabe +3
This paper investigates the applications of various multilingual approaches developed in conventional hidden Markov model (HMM) systems to sequence-to-sequence (seq2seq) automatic…
Promising Accurate Prefix Boosting for sequence-to-sequence ASR
Murali Karthick Baskar, Lukáš Burget, Shinji Watanabe +3
In this paper, we present promising accurate prefix boosting (PAPB), a discriminative training technique for attention based sequence-to-sequence (seq2seq) ASR. PAPB is devised to…
How to Improve Your Speaker Embeddings Extractor in Generic Toolkits
Hossein Zeinali, Lukas Burget, Johan Rohdin +2
Recently, speaker embeddings extracted with deep neural networks became the state-of-the-art method for speaker verification. In this paper we aim to facilitate its implementation…
Building and Evaluation of a Real Room Impulse Response Dataset
Igor Szoke, Miroslav Skacel, Ladislav Mosner +2
This paper presents BUT ReverbDB - a dataset of real room impulse responses (RIR), background noises and re-transmitted speech data. The retransmitted data includes LibriSpeech tes…
Convolutional Neural Networks and x-vector Embedding for DCASE2018 Acoustic Scene Classification Challenge
Hossein Zeinali, Lukas Burget, Jan Cernocky
In this paper, the Brno University of Technology (BUT) team submissions for Task 1 (Acoustic Scene Classification, ASC) of the DCASE-2018 challenge are described. Also, the analysi…