10 citations · 37 across the 23 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2021★ 10 cited
Lhotse: a speech data representation library for the modern deep learning ecosystem
Piotr Żelasko, Daniel Povey, Jan "Yenda" Trmal +1
Speech data is notoriously difficult to work with due to a variety of codecs, lengths of recordings, and meta-data formats. We present Lhotse, a speech data representation library…
cs.SD2021
An Asynchronous WFST-Based Decoder For Automatic Speech Recognition
Hang Lv, Zhehuai Chen, Hainan Xu +3
We introduce asynchronous dynamic decoder, which adopts an efficient A* algorithm to incorporate big language models in the one-pass decoding for large vocabulary continuous speech…