activity
20172025
most citedStacked Convolutional and Recurrent Neural Networks for Music Emotion Recognition

45 citations · 117 across the 18 of their papers we have counts for

collaborators
Showing cs.SDShow all

21 papers · 1 filter

cs.SD20213 cited

Continual Learning for Automated Audio Captioning Using The Learning Without Forgetting Approach

Jan Berg, Konstantinos Drossos

Automated audio captioning (AAC) is the task of automatically creating textual descriptions (i.e. captions) for the contents of a general audio signal. Most AAC methods are using e…

cs.SD20215 cited

Assessment of Self-Attention on Learned Features For Sound Event Localization and Detection

Parthasaarathy Sudarsanam, Archontis Politis, Konstantinos Drossos

Joint sound event localization and detection (SELD) is an emerging audio signal processing task adding spatial dimensions to acoustic scene analysis and sound event detection. A po…

cs.SD202123 cited

Enriched Music Representations with Multiple Cross-modal Contrastive Learning

Andres Ferraro, Xavier Favory, Konstantinos Drossos +2

Modeling various aspects that make a music piece unique is a challenging task, requiring the combination of multiple sources of information. Deep learning is commonly used to obtai…

cs.SD2021

Towards Citizen Science for Smart Cities: A Framework for a Collaborative Game of Bird Call Recognition Based on Internet of Sound Practices

Emmanuel Rovithis, Nikolaos Moustakas, Konstantinos Vogklis +2

Citizen Science aims to engage people in research activities on important issues related to their well-being. Smart Cities aim to provide them with services that improve the qualit…

cs.SD2020

Learning Contextual Tag Embeddings for Cross-Modal Alignment of Audio and Tags

Xavier Favory, Konstantinos Drossos, Tuomas Virtanen +1

Self-supervised audio representation learning offers an attractive alternative for obtaining generic audio embeddings, capable to be employed into various downstream tasks. Publish…

cs.SD2020

WaveTransformer: A Novel Architecture for Audio Captioning Based on Learning Temporal and Time-Frequency Information

An Tran, Konstantinos Drossos, Tuomas Virtanen

Automated audio captioning (AAC) is a novel task, where a method takes as an input an audio sample and outputs a textual description (i.e. a caption) of its contents. Most AAC meth…