859 citations · 1.5k across the 35 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2020★ 9 cited
COALA: Co-Aligned Autoencoders for Learning Semantically Enriched Audio Representations
Xavier Favory, Konstantinos Drossos, Tuomas Virtanen +1
Audio representation learning based on deep neural networks (DNNs) emerged as an alternative approach to hand-crafted features. For achieving high performance, DNNs often need a la…
cs.LG2019
Zero-Shot Audio Classification Based on Class Label Embeddings
Huang Xie, Tuomas Virtanen
This paper proposes a zero-shot learning approach for audio classification based on the textual information about class labels without any audio samples from target classes. We pro…