40 citations · 149 across the 63 of their papers we have counts for
2 papers
eess.AS2021★ 1 cited
ESPnet-ST IWSLT 2021 Offline Speech Translation System
Hirofumi Inaguma, Brian Yan, Siddharth Dalmia +4
This paper describes the ESPnet-ST group's IWSLT 2021 submission in the offline speech translation track. This year we made various efforts on training data, architecture, and audi…
cs.CL2016★ 18 cited
Joint CTC-Attention based End-to-End Speech Recognition using Multi-task Learning
Suyoun Kim, Takaaki Hori, Shinji Watanabe
Recently, there has been an increasing interest in end-to-end speech recognition that directly transcribes speech to text without any predefined alignments. One approach is the att…