1 paper · 1 filter
Junjie Li, Xuelong Geng, Kun Xie +13
A unified audio model must recognize and understand linguistic, paralinguistic, and environmental information while supporting speech synthesis and editing. A key challenge is repr…