204 citations · 218 across the 17 of their papers we have counts for
Showing 2024 · cs.SDShow all
2 papers · 2 filters
cs.SD2024
Comparative Analysis of Audio Feature Extraction for Real-Time Talking Portrait Synthesis
Pegah Salehi, Sajad Amouei Sheshkal, Vajira Thambawita +5
This paper examines the integration of real-time talking-head generation for interviewer training, focusing on overcoming challenges in Audio Feature Extraction (AFE), which often…
cs.SD2024★ 7 cited
SoccerNet-Echoes: A Soccer Game Audio Commentary Dataset
Sushant Gautam, Mehdi Houshmand Sarkhoosh, Jan Held +7
The application of Automatic Speech Recognition (ASR) technology in soccer offers numerous opportunities for sports analytics. Specifically, extracting audio commentaries with ASR…