activity
20172025
most citedBUT System Description to VoxCeleb Speaker Recognition Challenge 2019

79 citations · 84 across the 13 of their papers we have counts for

collaborators
Showing eess.ASShow all

17 papers · 1 filter

eess.AS2025

BUT Systems for WildSpoof Challenge: SASV in the Wild

Junyi Peng, Jin Li, Johan Rohdin +3

This paper presents the BUT submission to the WildSpoof Challenge, focusing on the Spoofing-robust Automatic Speaker Verification (SASV) track. We propose a SASV framework designed…

eess.AS2025

BUT Systems for Environmental Sound Deepfake Detection in the ESDD 2026 Challenge

Junyi Peng, Lin Zhang, Jin Li +2

This paper describes the BUT submission to the ESDD 2026 Challenge, specifically focusing on Track 1: Environmental Sound Deepfake Detection with Unseen Generators. To address the…

eess.AS2024

Challenging margin-based speaker embedding extractors by using the variational information bottleneck

Themos Stafylakis, Anna Silnova, Johan Rohdin +2

Speaker embedding extractors are typically trained using a classification loss over the training speakers. During the last few years, the standard softmax/cross-entropy loss has be…

eess.AS2024

Probing Self-supervised Learning Models with Target Speech Extraction

Junyi Peng, Marc Delcroix, Tsubasa Ochiai +4

Large-scale pre-trained self-supervised learning (SSL) models have shown remarkable advancements in speech-related tasks. However, the utilization of these models in complex multi-…

eess.AS2024

Target Speech Extraction with Pre-trained Self-supervised Learning Models

Junyi Peng, Marc Delcroix, Tsubasa Ochiai +3

Pre-trained self-supervised learning (SSL) models have achieved remarkable success in various speech tasks. However, their potential in target speech extraction (TSE) has not been…

eess.AS20221 cited

Extracting speaker and emotion information from self-supervised speech models via channel-wise correlations

Themos Stafylakis, Ladislav Mosner, Sofoklis Kakouros +3

Self-supervised learning of speech representations from large amounts of unlabeled data has enabled state-of-the-art results in several speech processing tasks. Aggregating these s…