2 papers
cs.SD2025
SAO-Instruct: Free-form Audio Editing using Natural Language Instructions
Michael Ungersböck, Florian Grötschla, Luca A. Lanzendörfer +3
Generative models have made significant progress in synthesizing high-fidelity audio from short textual descriptions. However, editing existing audio using natural language has rem…
cs.SD2025
High-Fidelity Music Vocoder using Neural Audio Codecs
Luca A. Lanzendörfer, Florian Grötschla, Michael Ungersböck +1
While neural vocoders have made significant progress in high-fidelity speech synthesis, their application on polyphonic music has remained underexplored. In this work, we propose D…