activity
20182022
most citedBUT VOiCES 2019 System Description

2 citations · 5 across the 5 of their papers we have counts for

collaborators
Showing eess.ASShow all

8 papers · 1 filter

eess.AS2022

Speech-based emotion recognition with self-supervised models using attentive channel-wise correlations and label smoothing

Sofoklis Kakouros, Themos Stafylakis, Ladislav Mosner +1

When recognizing emotions from speech, we encounter two common problems: how to optimally capture emotion-relevant information from the speech signal and how to best quantify or ca…

eess.AS20221 cited

Extracting speaker and emotion information from self-supervised speech models via channel-wise correlations

Themos Stafylakis, Ladislav Mosner, Sofoklis Kakouros +3

Self-supervised learning of speech representations from large amounts of unlabeled data has enabled state-of-the-art results in several speech processing tasks. Aggregating these s…

eess.AS20221 cited

An attention-based backend allowing efficient fine-tuning of transformer models for speaker verification

Junyi Peng, Oldrich Plchot, Themos Stafylakis +3

In recent years, self-supervised learning paradigm has received extensive attention due to its great success in various down-stream tasks. However, the fine-tuning strategies for a…

eess.AS20221 cited

Analyzing speaker verification embedding extractors and back-ends under language and channel mismatch

Anna Silnova, Themos Stafylakis, Ladislav Mosner +6

In this paper, we analyze the behavior and performance of speaker embeddings and the back-end scoring model under domain and language mismatch. We present our findings regarding Re…

eess.AS2020

BUT System for the Second DIHARD Speech Diarization Challenge

Federico Landini, Shuai Wang, Mireia Diez +9

This paper describes the winning systems developed by the BUT team for the four tracks of the Second DIHARD Speech Diarization Challenge. For tracks 1 and 2 the systems were mainly…

eess.AS20192 cited

BUT VOiCES 2019 System Description

Hossein Zeinali, Pavel Matějka, Ladislav Mošner +6

This is a description of our effort in VOiCES 2019 Speaker Recognition challenge. All systems in the fixed condition are based on the x-vector paradigm with different features and…