2 papers
cs.CL2025
EM2LDL: A Multilingual Speech Corpus for Mixed Emotion Recognition through Label Distribution Learning
Xingfeng Li, Xiaohan Shi, Junjie Li +4
This study introduces EM2LDL, a novel multilingual speech corpus designed to advance mixed emotion recognition through label distribution learning. Addressing the limitations of pr…
cs.SD2024
Modeling and Estimation of Vocal Tract and Glottal Source Parameters Using ARMAX-LF Model
Kai Lia, Masato Akagia, Yongwei Lib +1
Modeling and estimation of the vocal tract and glottal source parameters of vowels from raw speech can be typically done by using the Auto-Regressive with eXogenous input (ARX) mod…