19 citations · 27 across the 7 of their papers we have counts for
8 papers
Audio-Visual Speaker Diarization: Current Databases, Approaches and Challenges
Victoria Mingote, Alfonso Ortega, Antonio Miguel +1
Nowadays, the large amount of audio-visual content available has fostered the need to develop new robust automatic speaker diarization systems to analyse and characterise it. This…
Generalizing AUC Optimization to Multiclass Classification for Audio Segmentation With Limited Training Data
Pablo Gimeno, Victoria Mingote, Alfonso Ortega +2
Area under the ROC curve (AUC) optimisation techniques developed for neural networks have recently demonstrated their capabilities in different audio and speech related tasks. Howe…
Shouted Speech Compensation for Speaker Verification Robust to Vocal Effort Conditions
Santi Prieto, Alfonso Ortega, Iván López-Espejo +1
The performance of speaker verification systems degrades when vocal effort conditions between enrollment and test (e.g., shouted vs. normal speech) are different. This is a potenti…
Speech Enhancement with Wide Residual Networks in Reverberant Environments
Jorge Llombart, Dayana Ribas, Antonio Miguel +3
This paper proposes a speech enhancement method which exploits the high potential of residual connections in a Wide Residual Network architecture. This is supported on single dimen…
Progressive Speech Enhancement with Residual Connections
Jorge Llombart, Dayana Ribas, Antonio Miguel +3
This paper studies the Speech Enhancement based on Deep Neural Networks. The proposed architecture gradually follows the signal transformation during enhancement by means of a visu…
Optimization of the Area Under the ROC Curve using Neural Network Supervectors for Text-Dependent Speaker Verification
Victoria Mingote, Antonio Miguel, Alfonso Ortega +1
This paper explores two techniques to improve the performance of text-dependent speaker verification systems based on deep neural networks. Firstly, we propose a general alignment…