1 citations · 1 across the 2 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2024
Improving Speaker Representations Using Contrastive Losses on Multi-scale Features
Satvik Dixit, Massa Baali, Rita Singh +1
Speaker verification systems have seen significant advancements with the introduction of Multi-scale Feature Aggregation (MFA) architectures, such as MFA-Conformer and ECAPA-TDNN.…
cs.SD2024★ 1 cited
Explaining Deep Learning Embeddings for Speech Emotion Recognition by Predicting Interpretable Acoustic Features
Satvik Dixit, Daniel M. Low, Gasser Elbanna +2
Pre-trained deep learning embeddings have consistently shown superior performance over handcrafted acoustic features in speech emotion recognition (SER). However, unlike acoustic f…