activity
20202023
most citedMultichannel CRNN for Speaker Counting: an Analysis of Performance

2 citations · 2 across the 3 of their papers we have counts for

collaborators

6 papers

cs.SD2023

Efficient bandwidth extension of musical signals using a differentiable harmonic plus noise model

Pierre-Amaury Grumiaux, Mathieu Lagrange

The task of bandwidth extension addresses the generation of missing high frequencies of audio signals based on knowledge of the low-frequency part of the sound. This task applies t…

cs.SD2021

A Survey of Sound Source Localization with Deep Learning Methods

Pierre-Amaury Grumiaux, Srđan Kitić, Laurent Girin +1

This article is a survey on deep learning methods for single and multiple sound source localization. We are particularly interested in sound source localization in indoor/domestic…

cs.SD2021

SALADnet: Self-Attentive multisource Localization in the Ambisonics Domain

Pierre-Amaury Grumiaux, Srdan Kitic, Prerak Srivastava +2

In this work, we propose a novel self-attention based neural network for robust multi-speaker localization from Ambisonics recordings. Starting from a state-of-the-art convolutiona…

cs.SD2021

Improved feature extraction for CRNN-based multiple sound source localization

Pierre-Amaury Grumiaux, Srdan Kitic, Laurent Girin +1

In this work, we propose to extend a state-of-the-art multi-source localization system based on a convolutional recurrent neural network and Ambisonics signals. We significantly im…

cs.SD2021★ 2 cited

Multichannel CRNN for Speaker Counting: an Analysis of Performance

Pierre-Amaury Grumiaux, Srdan Kitic, Laurent Girin +1

Speaker counting is the task of estimating the number of people that are simultaneously speaking in an audio recording. For several audio processing tasks such as speaker diarizati…

cs.SD2020

High-Resolution Speaker Counting In Reverberant Rooms Using CRNN With Ambisonics Features

Pierre-Amaury Grumiaux, Srdjan Kitic, Laurent Girin +1

Speaker counting is the task of estimating the number of people that are simultaneously speaking in an audio recording. For several audio processing tasks such as speaker diarizati…