1 paper
Frédéric Berdoz, Luca A. Lanzendörfer, Antonis Asonitis +1
Speech enhancement language models achieve strong results when trained on discrete audio tokens, but their optimization relies on token-level cross-entropy rather than the perceptu…