activity
20172023
most citedEfficient Retrieval Augmented Generation from Unstructured Knowledge for Task-Oriented Dialog

29 citations · 150 across the 31 of their papers we have counts for

collaborators
Showing cs.SDShow all

5 papers · 1 filter

cs.SD2023

On the Relation between Internal Language Model and Sequence Discriminative Training for Neural Transducers

Zijian Yang, Wei Zhou, Ralf Schlüter +1

Internal language model (ILM) subtraction has been widely applied to improve the performance of the RNN-Transducer with external language model (LM) fusion for speech recognition.…

cs.SD20221 cited

HMM vs. CTC for Automatic Speech Recognition: Comparison Based on Full-Sum Training from Scratch

Tina Raissi, Wei Zhou, Simon Berger +2

In this work, we compare from-scratch sequence-level cross-entropy (full-sum) training of Hidden Markov Model (HMM) and Connectionist Temporal Classification (CTC) topologies for a…

cs.SD2022

Improving Factored Hybrid HMM Acoustic Modeling without State Tying

Tina Raissi, Eugen Beck, Ralf Schlüter +1

In this work, we show that a factored hybrid hidden Markov model (FH-HMM) which is defined without any phonetic state-tying outperforms a state-of-the-art hybrid HMM. The factored…

cs.SD2021

Towards Consistent Hybrid HMM Acoustic Modeling

Tina Raissi, Eugen Beck, Ralf Schlüter +1

High-performance hybrid automatic speech recognition (ASR) systems are often trained with clustered triphone outputs, and thus require a complex training pipeline to generate the c…

cs.SD2019

Analysis of Deep Clustering as Preprocessing for Automatic Speech Recognition of Sparsely Overlapping Speech

Tobias Menne, Ilya Sklyar, Ralf Schlüter +1

Significant performance degradation of automatic speech recognition (ASR) systems is observed when the audio signal contains cross-talk. One of the recently proposed approaches to…