activity
20182021
most citedGeneralizing AUC Optimization to Multiclass Classification for Audio Segmentation With Limited Training Data

19 citations · 42 across the 6 of their papers we have counts for

collaborators

10 papers

cs.SD202119 cited

Generalizing AUC Optimization to Multiclass Classification for Audio Segmentation With Limited Training Data

Pablo Gimeno, Victoria Mingote, Alfonso Ortega +2

Area under the ROC curve (AUC) optimisation techniques developed for neural networks have recently demonstrated their capabilities in different audio and speech related tasks. Howe…

eess.AS2020

Robust Sound Source Tracking Using SRP-PHAT and 3D Convolutional Neural Networks

David Diaz-Guerra, Antonio Miguel, Jose R. Beltran

In this paper, we present a new single sound source DOA estimation and tracking system based on the well-known SRP-PHAT algorithm and a three-dimensional Convolutional Neural Netwo…

eess.AS2019

Speech Enhancement with Wide Residual Networks in Reverberant Environments

Jorge Llombart, Dayana Ribas, Antonio Miguel +3

This paper proposes a speech enhancement method which exploits the high potential of residual connections in a Wide Residual Network architecture. This is supported on single dimen…

eess.AS2019

Progressive Speech Enhancement with Residual Connections

Jorge Llombart, Dayana Ribas, Antonio Miguel +3

This paper studies the Speech Enhancement based on Deep Neural Networks. The proposed architecture gradually follows the signal transformation during enhancement by means of a visu…

cs.SD201915 cited

Deep Speech Enhancement for Reverberated and Noisy Signals using Wide Residual Networks

Dayana Ribas, Jorge Llombart, Antonio Miguel +1

This paper proposes a deep speech enhancement method which exploits the high potential of residual connections in a wide neural network architecture, a topology known as Wide Resid…

cs.SD2019

Optimization of the Area Under the ROC Curve using Neural Network Supervectors for Text-Dependent Speaker Verification

Victoria Mingote, Antonio Miguel, Alfonso Ortega +1

This paper explores two techniques to improve the performance of text-dependent speaker verification systems based on deep neural networks. Firstly, we propose a general alignment…