5 citations · 15 across the 11 of their papers we have counts for
4 papers · 1 filter
Automatic multitrack mixing with a differentiable mixing console of neural audio effects
Christian J. Steinmetz, Jordi Pons, Santiago Pascual +1
Applications of deep learning to automatic multitrack mixing are largely unexplored. This is partly due to the limited available data, coupled with the fact that such data is relat…
SESQA: semi-supervised learning for speech quality assessment
Joan Serrà, Jordi Pons, Santiago Pascual
Automatic speech quality assessment is an important, transversal task whose progress is hampered by the scarcity of human annotations, poor generalization to unseen recording condi…
TensorFlow Audio Models in Essentia
Pablo Alonso-Jiménez, Dmitry Bogdanov, Jordi Pons +1
Essentia is a reference open-source C++/Python library for audio and music analysis. In this work, we present a set of algorithms that employ TensorFlow in Essentia, allow predicti…
An empirical study of Conv-TasNet
Berkan Kadioglu, Michael Horgan, Xiaoyu Liu +3
Conv-TasNet is a recently proposed waveform-based deep neural network that achieves state-of-the-art performance in speech source separation. Its architecture consists of a learnab…