387 citations · 1.2k across the 47 of their papers we have counts for
1 paper · 2 filters
Bo Li, Yu Zhang, Tara Sainath +2
We present two end-to-end models: Audio-to-Byte (A2B) and Byte-to-Audio (B2A), for multilingual speech recognition and synthesis. Prior work has predominantly used characters, sub-…