5 papers
Uncertainty Estimation in the Real World: A Study on Music Emotion Recognition
Karn N. Watcharasupat, Yiwei Ding, T. Aleksandra Ma +2
Any data annotation for subjective tasks shows potential variations between individuals. This is particularly true for annotations of emotional responses to musical stimuli. While…
FRCRN: Boosting Feature Representation using Frequency Recurrence for Monaural Speech Enhancement
Shengkui Zhao, Bin Ma, Karn N. Watcharasupat +1
Convolutional recurrent networks (CRN) integrating a convolutional encoder-decoder (CED) structure and a recurrent structure have achieved promising performance for monaural speech…
A Stem-Agnostic Single-Decoder System for Music Source Separation Beyond Four Stems
Karn N. Watcharasupat, Alexander Lerch
Despite significant recent progress across multiple subtasks of audio source separation, few music source separation systems support separation beyond the four-stem vocals, drums,…
Autonomous Soundscape Augmentation with Multimodal Fusion of Visual and Participant-linked Inputs
Kenneth Ooi, Karn N. Watcharasupat, Bhan Lam +2
Autonomous soundscape augmentation systems typically use trained models to pick optimal maskers to effect a desired perceptual change. While acoustic information is paramount to su…
ARAUS: A Large-Scale Dataset and Baseline Models of Affective Responses to Augmented Urban Soundscapes
Kenneth Ooi, Zhen-Ting Ong, Karn N. Watcharasupat +3
Choosing optimal maskers for existing soundscapes to effect a desired perceptual change via soundscape augmentation is non-trivial due to extensive varieties of maskers and a deart…