1 paper
Yi Zhu, Brahmi Dwivedi, Jayaram Raghuram +1
Existing voice deepfake detection and localization models rely heavily on representations extracted from speech foundation models (SFMs). However, downstream finetuning has now rea…