2 citations · 2 across the 5 of their papers we have counts for
1 paper · 1 filter
Yakun Song, Zhuo Chen, Xiaofei Wang +2
The language model (LM) approach based on acoustic and linguistic prompts, such as VALL-E, has achieved remarkable progress in the field of zero-shot audio generation. However, exi…