2 papers
cs.SD2025
Foundation Model Hidden Representations for Heart Rate Estimation from Auscultation
Jingping Nie, Dung T. Tran, Karan Thakkar +5
Auscultation, particularly heart sound, is a non-invasive technique that provides essential vital sign information. Recently, self-supervised acoustic representation foundation mod…
cs.SD2025
Modeling speech emotion with label variance and analyzing performance across speakers and unseen acoustic conditions
Vikramjit Mitra, Amrit Romana, Dung T. Tran +1
Spontaneous speech emotion data usually contain perceptual grades where graders assign emotion score after listening to the speech files. Such perceptual grades introduce uncertain…