22 citations · 24 across the 6 of their papers we have counts for
6 papers
Attention-Based Methods For Audio Question Answering
Parthasaarathy Sudarsanam, Tuomas Virtanen
Audio question answering (AQA) is the task of producing natural language answers when a system is provided with audio and natural language questions. In this paper, we propose neur…
Few-shot Class-incremental Audio Classification Using Adaptively-refined Prototypes
Wei Xie, Yanxiong Li, Qianhua He +2
New classes of sounds constantly emerge with a few samples, making it challenging for models to adapt to dynamic acoustic environments. This challenge motivates us to address the n…
Multi-Channel Masking with Learnable Filterbank for Sound Source Separation
Wang Dai, Archontis Politis, Tuomas Virtanen
This work proposes a learnable filterbank based on a multi-channel masking framework for multi-channel source separation. The learnable filterbank is a 1D Conv layer, which transfo…
Subjective Evaluation of Deep Neural Network Based Speech Enhancement Systems in Real-World Conditions
Gaurav Naithani, Kirsi Pietilä, Riitta Niemistö +3
Subjective evaluation results for two low-latency deep neural networks (DNN) are compared to a matured version of a traditional Wiener-filter based noise suppressor. The target use…
Domestic Activity Clustering from Audio via Depthwise Separable Convolutional Autoencoder Network
Yanxiong Li, Wenchang Cao, Konstantinos Drossos +1
Automatic estimation of domestic activities from audio can be used to solve many problems, such as reducing the labor cost for nursing the elderly people. This study focuses on sol…
Low-complexity acoustic scene classification in DCASE 2022 Challenge
Irene Martín-Morató, Francesco Paissan, Alberto Ancilotto +5
This paper presents an analysis of the Low-Complexity Acoustic Scene Classification task in DCASE 2022 Challenge. The task was a continuation from the previous years, but the low-c…