3 papers
eess.AS2024
Audio-Visual Approach For Multimodal Concurrent Speaker Detection
Amit Eliav, Sharon Gannot
Concurrent Speaker Detection (CSD), the task of identifying active speakers and their overlaps in an audio signal, is essential for various audio applications, including meeting tr…
eess.AS2024
SingIt! Singer Voice Transformation
Amit Eliav, Aaron Taub, Renana Opochinsky +1
In this paper, we propose a model which can generate a singing voice from normal speech utterance by harnessing zero-shot, many-to-many style transfer learning. Our goal is to give…
eess.AS2024
Concurrent Speaker Detection: A multi-microphone Transformer-Based Approach
Amit Eliav, Sharon Gannot
We present a deep-learning approach for the task of Concurrent Speaker Detection (CSD) using a modified transformer model. Our model is designed to handle multi-microphone data but…