1 paper · 1 filter
Héctor Martel, Joe Hennessy-Priest, Taemin Cho
Audio foundation models are widely adopted as general-purpose feature extractors, yet the internal structure of their learned representations remains insufficiently understood. In…