activity
20182024
most citedGeneralizing AUC Optimization to Multiclass Classification for Audio Segmentation With Limited Training Data

19 citations · 27 across the 7 of their papers we have counts for

collaborators

8 papers

cs.SD2024

Audio-Visual Speaker Diarization: Current Databases, Approaches and Challenges

Victoria Mingote, Alfonso Ortega, Antonio Miguel +1

Nowadays, the large amount of audio-visual content available has fostered the need to develop new robust automatic speaker diarization systems to analyse and characterise it. This…

cs.SD202119 cited

Generalizing AUC Optimization to Multiclass Classification for Audio Segmentation With Limited Training Data

Pablo Gimeno, Victoria Mingote, Alfonso Ortega +2

Area under the ROC curve (AUC) optimisation techniques developed for neural networks have recently demonstrated their capabilities in different audio and speech related tasks. Howe…

eess.AS2020

Shouted Speech Compensation for Speaker Verification Robust to Vocal Effort Conditions

Santi Prieto, Alfonso Ortega, Iván López-Espejo +1

The performance of speaker verification systems degrades when vocal effort conditions between enrollment and test (e.g., shouted vs. normal speech) are different. This is a potenti…

eess.AS2019

Speech Enhancement with Wide Residual Networks in Reverberant Environments

Jorge Llombart, Dayana Ribas, Antonio Miguel +3

This paper proposes a speech enhancement method which exploits the high potential of residual connections in a Wide Residual Network architecture. This is supported on single dimen…

eess.AS2019

Progressive Speech Enhancement with Residual Connections

Jorge Llombart, Dayana Ribas, Antonio Miguel +3

This paper studies the Speech Enhancement based on Deep Neural Networks. The proposed architecture gradually follows the signal transformation during enhancement by means of a visu…

cs.SD2019

Optimization of the Area Under the ROC Curve using Neural Network Supervectors for Text-Dependent Speaker Verification

Victoria Mingote, Antonio Miguel, Alfonso Ortega +1

This paper explores two techniques to improve the performance of text-dependent speaker verification systems based on deep neural networks. Firstly, we propose a general alignment…