2 papers
cs.SD2024
Improving speaker verification robustness with synthetic emotional utterances
Nikhil Kumar Koditala, Chelsea Jui-Ting Ju, Ruirui Li +3
A speaker verification (SV) system offers an authentication service designed to confirm whether a given speech sample originates from a specific speaker. This technology has paved…
cs.CL2024
Large Language Model Based Generative Error Correction: A Challenge and Baselines for Speech Recognition, Speaker Tagging, and Emotion Recognition
Chao-Han Huck Yang, Taejin Park, Yuan Gong +18
Given recent advances in generative AI technology, a key question is how large language models (LLMs) can enhance acoustic modeling tasks using text decoding results from a frozen,…