17 citations · 43 across the 26 of their papers we have counts for
Showing 2024 · eess.ASShow all
3 papers · 2 filters
eess.AS2024
Easy, Interpretable, Effective: openSMILE for voice deepfake detection
Octavian Pascu, Dan Oneata, Horia Cucu +1
In this paper, we demonstrate that attacks in the latest ASVspoof5 dataset -- a de facto standard in the field of voice authenticity and deepfake detection -- can be identified wit…
eess.AS2024
WavLM model ensemble for audio deepfake detection
David Combei, Adriana Stan, Dan Oneata +1
Audio deepfake detection has become a pivotal task over the last couple of years, as many recent speech synthesis and voice cloning systems generate highly realistic speech samples…
eess.AS2024
Translating speech with just images
Dan Oneata, Herman Kamper
Visually grounded speech models link speech to images. We extend this connection by linking images to text via an existing image captioning system, and as a result gain the ability…