Showing 2024Show all
2 papers · 1 filter
cs.SD2024
Audio Spotforming Using Nonnegative Tensor Factorization with Attractor-Based Regularization
Shoma Ayano, Li Li, Shogo Seki +1
Spotforming is a target-speaker extraction technique that uses multiple microphone arrays. This method applies beamforming (BF) to each microphone array, and the common components…
cs.SD2024
Improved Remixing Process for Domain Adaptation-Based Speech Enhancement by Mitigating Data Imbalance in Signal-to-Noise Ratio
Li Li, Shogo Seki
RemixIT and Remixed2Remixed are domain adaptation-based speech enhancement (DASE) methods that use a teacher model trained in full supervision to generate pseudo-paired data by rem…