4 citations · 5 across the 3 of their papers we have counts for
3 papers
eess.AS2021★ 1 cited
Speaker disentanglement in video-to-speech conversion
Dan Oneata, Adriana Stan, Horia Cucu
The task of video-to-speech aims to translate silent video of lip movement to its corresponding audio signal. Previous approaches to this task are generally limited to the case of…
eess.AS2021★ 4 cited
An evaluation of word-level confidence estimation for end-to-end automatic speech recognition
Dan Oneata, Alexandru Caranica, Adriana Stan +1
Quantifying the confidence (or conversely the uncertainty) of a prediction is a highly desirable trait of an automatic system, as it improves the robustness and usefulness in downs…
eess.AS2020
RECOApy: Data recording, pre-processing and phonetic transcription for end-to-end speech-based applications
Adriana Stan
Deep learning enables the development of efficient end-to-end speech processing applications while bypassing the need for expert linguistic and signal processing features. Yet, rec…