Showing cs.SDShow all
2 papers · 1 filter
cs.SD2024
Comparative Analysis of Audio Feature Extraction for Real-Time Talking Portrait Synthesis
Pegah Salehi, Sajad Amouei Sheshkal, Vajira Thambawita +5
This paper examines the integration of real-time talking-head generation for interviewer training, focusing on overcoming challenges in Audio Feature Extraction (AFE), which often…
cs.SD2024
SoccerNet-Echoes: A Soccer Game Audio Commentary Dataset
Sushant Gautam, Mehdi Houshmand Sarkhoosh, Jan Held +7
The application of Automatic Speech Recognition (ASR) technology in soccer offers numerous opportunities for sports analytics. Specifically, extracting audio commentaries with ASR…