2 papers
cs.CL2023
Optimized Tokenization for Transcribed Error Correction
Tomer Wullach, Shlomo E. Chazan
The challenges facing speech recognition systems, such as variations in pronunciations, adverse audio conditions, and the scarcity of labeled data, emphasize the necessity for a po…
cs.SD2023
A two-stage speaker extraction algorithm under adverse acoustic conditions using a single-microphone
Aviad Eisenberg, Sharon Gannot, Shlomo E. Chazan
In this work, we present a two-stage method for speaker extraction under reverberant and noisy conditions. Given a reference signal of the desired speaker, the clean, but the still…