17 citations · 24 across the 9 of their papers we have counts for
3 papers · 1 filter
Easy, Interpretable, Effective: openSMILE for voice deepfake detection
Octavian Pascu, Dan Oneata, Horia Cucu +1
In this paper, we demonstrate that attacks in the latest ASVspoof5 dataset -- a de facto standard in the field of voice authenticity and deepfake detection -- can be identified wit…
Speaker disentanglement in video-to-speech conversion
Dan Oneata, Adriana Stan, Horia Cucu
The task of video-to-speech aims to translate silent video of lip movement to its corresponding audio signal. Previous approaches to this task are generally limited to the case of…
An evaluation of word-level confidence estimation for end-to-end automatic speech recognition
Dan Oneata, Alexandru Caranica, Adriana Stan +1
Quantifying the confidence (or conversely the uncertainty) of a prediction is a highly desirable trait of an automatic system, as it improves the robustness and usefulness in downs…