2 papers
cs.SD2025
Recomposer: Event-roll-guided generative audio editing
Daniel P. W. Ellis, Eduardo Fonseca, Ron J. Weiss +7
Editing complex real-world sound scenes is difficult because individual sound sources overlap in time. Generative models can fill-in missing or corrupted details based on their str…
cs.SD2025
Towards Sub-millisecond Latency Real-Time Speech Enhancement Models on Hearables
Artem Dementyev, Chandan K. A. Reddy, Scott Wisdom +3
Low latency models are critical for real-time speech enhancement applications, such as hearing aids and hearables. However, the sub-millisecond latency space for resource-constrain…