4 papers · 1 filter
Physics-Guided Variational Model for Unsupervised Sound Source Tracking
Luan VinÃcius Fiorio, Ivana Nikoloska, Bruno Defraene +3
Sound source tracking is commonly performed using classical array-processing algorithms, while machine-learning approaches typically rely on precise source position labels that are…
Unsupervised Variational Acoustic Clustering
Luan VinÃcius Fiorio, Bruno Defraene, Johan David +3
We propose an unsupervised variational acoustic clustering model for clustering audio data in the time-frequency domain. The model leverages variational inference, extended to an a…
Target Speaker Selection for Neural Network Beamforming in Multi-Speaker Scenarios
Luan VinÃcius Fiorio, Bruno Defraene, Johan David +4
We propose a speaker selection mechanism (SSM) for the training of an end-to-end beamforming neural network, based on recent findings that a listener usually looks to the target sp…
Spectral Masking with Explicit Time-Context Windowing for Neural Network-Based Monaural Speech Enhancement
Luan VinÃcius Fiorio, Boris Karanov, Bruno Defraene +4
We propose and analyze the use of an explicit time-context window for neural network-based spectral masking speech enhancement to leverage signal context dependencies between neigh…