2 papers
eess.AS2025
Combining TF-GridNet and Mixture Encoder for Continuous Speech Separation for Meeting Transcription
Peter Vieting, Simon Berger, Thilo von Neumann +3
Many real-life applications of automatic speech recognition (ASR) require processing of overlapped speech. A common method involves first separating the speech into overlap-free st…
cs.CL2025
Efficient Supernet Training with Orthogonal Softmax for Scalable ASR Model Compression
Jingjing Xu, Eugen Beck, Zijian Yang +1
ASR systems are deployed across diverse environments, each with specific hardware constraints. We use supernet training to jointly train multiple encoders of varying sizes, enablin…