Showing cs.SDShow all
3 papers · 1 filter
cs.SD2026
From A to B to A: Palindromic Zero-Shot Voice Conversion with Non-Parallel Data
Moshe Mandel, Shlomo E. Chazan
We present a voice conversion (VC) framework that utilizes K-Nearest Neighbors (KNN) retrieval over WavLM representations to align non-parallel source and target speech, constructi…
cs.SD2025
Spectral or spatial? Leveraging both for speaker extraction in challenging data conditions
Aviad Eisenberg, Sharon Gannot, Shlomo E. Chazan
This paper presents a robust multi-channel speaker extraction algorithm designed to handle inaccuracies in reference information. While existing approaches often rely solely on eit…
cs.SD2025
End-to-End Multi-Microphone Speaker Extraction Using Relative Transfer Functions
Aviad Eisenberg, Sharon Gannot, Shlomo E. Chazan
This paper introduces a multi-microphone method for extracting a desired speaker from a mixture involving multiple speakers and directional noise in a reverberant environment. In t…