3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.CL2021★ 3 cited
Text Normalization for Low-Resource Languages of Africa
Andrew Zupon, Evan Crew, Sandy Ritchie
Training data for machine learning models can come from many different sources, which can be of dubious quality. For resource-rich languages like English, there is a lot of data av…
cs.CL2021
Mining Large-Scale Low-Resource Pronunciation Data From Wikipedia
Tania Chakraborty, Manasa Prasad, Theresa Breiner +2
Pronunciation modeling is a key task for building speech technology in new languages, and while solid grapheme-to-phoneme (G2P) mapping systems exist, language coverage can stand t…