1 paper
Hoseong Ahn, Jeongyun Chae, Yoonji Park +1
Long-form speech recognition with large encoder-decoder models such as Whisper often exhibit hallucinations, repetition loops, and content omissions. These errors can accumulate an…