2 papers
eess.AS2025
Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios
Aswin Shanmugam Subramanian, Amit Das, Naoyuki Kanda +3
We extend the frameworks of Serialized Output Training (SOT) to address practical needs of both streaming and offline automatic speech recognition (ASR) applications. Our approach…
cs.SD2025
Summary of the NOTSOFAR-1 Challenge: Highlights and Learnings
Igor Abramovski, Alon Vinnikov, Shalev Shaer +4
The first Natural Office Talkers in Settings of Far-field Audio Recordings (NOTSOFAR-1) Challenge is a pivotal initiative that sets new benchmarks by offering datasets more represe…