79 citations · 82 across the 11 of their papers we have counts for
4 papers · 1 filter
WeDefense: A Toolkit to Defend Against Fake Audio
Lin Zhang, Johan Rohdin, Xin Wang +8
The advances in generative AI have enabled the creation of synthetic audio which is perceptually indistinguishable from real, genuine audio. Although this stellar progress enables…
BUT Systems and Analyses for the ASVspoof 5 Challenge
Johan Rohdin, Lin Zhang, Oldřich Plchot +8
This paper describes the BUT submitted systems for the ASVspoof 5 challenge, along with analyses. For the conventional deepfake detection task, we use ResNet18 and self-supervised…
Do End-to-End Neural Diarization Attractors Need to Encode Speaker Characteristic Information?
Lin Zhang, Themos Stafylakis, Federico Landini +3
In this paper, we apply the variational information bottleneck approach to end-to-end neural diarization with encoder-decoder attractors (EEND-EDA). This allows us to investigate w…
Toroidal Probabilistic Spherical Discriminant Analysis
Anna Silnova, Niko Brümmer, Albert Swart +1
In speaker recognition, where speech segments are mapped to embeddings on the unit hypersphere, two scoring back-ends are commonly used, namely cosine scoring and PLDA. We have rec…