6 citations · 15 across the 5 of their papers we have counts for
7 papers
Task 1A DCASE 2021: Acoustic Scene Classification with mismatch-devices using squeeze-excitation technique and low-complexity constraint
Javier Naranjo-Alcazar, Sergi Perez-Castanos, Maximo Cobos +2
Acoustic scene classification (ASC) is one of the most popular problems in the field of machine listening. The objective of this problem is to classify an audio clip into one of th…
TASK3 DCASE2021 Challenge: Sound event localization and detection using squeeze-excitation residual CNNs
Javier Naranjo-Alcazar, Sergi Perez-Castanos, Pedro Zuccarello +2
Sound event localisation and detection (SELD) is a problem in the field of automatic listening that aims at the temporal detection and localisation (direction of arrival estimation…
Squeeze-Excitation Convolutional Recurrent Neural Networks for Audio-Visual Scene Classification
Javier Naranjo-Alcazar, Sergi Perez-Castanos, Aaron Lopez-Garcia +3
The use of multiple and semantically correlated sources can provide complementary information to each other that may not be evident when working with individual modalities on their…
Listen carefully and tell: an audio captioning system based on residual learning and gammatone audio representation
Sergi Perez-Castanos, Javier Naranjo-Alcazar, Pedro Zuccarello +1
Automated audio captioning is machine listening task whose goal is to describe an audio using free text. An automated audio captioning system has to be implemented as it accepts an…
Anomalous Sound Detection using unsupervised and semi-supervised autoencoders and gammatone audio representation
Sergi Perez-Castanos, Javier Naranjo-Alcazar, Pedro Zuccarello +1
Anomalous sound detection (ASD) is, nowadays, one of the topical subjects in machine listening discipline. Unsupervised detection is attracting a lot of interest due to its immedia…
Acoustic Scene Classification with Squeeze-Excitation Residual Networks
Javier Naranjo-Alcazar, Sergi Perez-Castanos, Pedro Zuccarello +1
Acoustic scene classification (ASC) is a problem related to the field of machine listening whose objective is to classify/tag an audio clip in a predefined label describing a scene…