4 papers
Audio Transformers
Prateek Verma, Jonathan Berger
Over the past two decades, CNN architectures have produced compelling models of sound perception and cognition, learning hierarchical organizations of features. Analogous to succes…
Diverse Audio Embeddings -- Bringing Features Back Outperforms CLAP!
Prateek Verma
With the advent of modern AI architectures, a shift has happened towards end-to-end architectures. This pivot has led to neural architectures being trained without domain-specific…
Content Adaptive Front End For Audio Classification
Prateek Verma, Chris Chafe
We propose a learnable content adaptive front end for audio signal processing. Before the modern advent of deep learning, we used fixed representation non-learnable front-ends like…
A Language Model With Million Context Length For Raw Audio
Prateek Verma
Modeling long-term dependencies for audio signals is a particularly challenging problem, as even small-time scales yield on the order of a hundred thousand samples. With the recent…