57 citations · 57 across the 2 of their papers we have counts for
2 papers
eess.AS2023
Towards Word-Level End-to-End Neural Speaker Diarization with Auxiliary Network
Yiling Huang, Weiran Wang, Guanlong Zhao +3
While standard speaker diarization attempts to answer the question "who spoken when", most of relevant applications in reality are more interested in determining "who spoken what".…
cs.CL2016★ 57 cited
Neural Speech Recognizer: Acoustic-to-Word LSTM Model for Large Vocabulary Speech Recognition
Hagen Soltau, Hank Liao, Hasim Sak
We present results that show it is possible to build a competitive, greatly simplified, large vocabulary continuous speech recognition system with whole words as acoustic units. We…