activity
20172026
most citedSeeing wake words: Audio-visual Keyword Spotting

22 citations · 29 across the 26 of their papers we have counts for

collaborators
Showing cs.CVShow all

11 papers · 1 filter

cs.CV2026

Sparsity-Inducing Divergence Losses for Biometric Verification

Dimitrios Koutsianos, Ladislav Mošner, Yannis Panagakis +1

Performance in face and speaker verification is largely driven by margin-penalty softmax losses such as CosFace and ArcFace. Recently introduced -divergence loss functions offer…

cs.CV2025

Alpha Divergence Losses for Biometric Verification

Dimitrios Koutsianos, Ladislav Mosner, Yannis Panagakis +1

Performance in face and speaker verification is largely driven by margin-based softmax losses such as CosFace and ArcFace. Recently introduced -divergence loss functions offer a…

cs.CV2023

A Simple Baseline for Knowledge-Based Visual Question Answering

Alexandros Xenos, Themos Stafylakis, Ioannis Patras +1

This paper is on the problem of Knowledge-Based Visual Question Answering (KB-VQA). Recent works have emphasized the significance of incorporating both explicit (through external d…

cs.CV202022 cited

Seeing wake words: Audio-visual Keyword Spotting

Liliane Momeni, Triantafyllos Afouras, Themos Stafylakis +2

The goal of this work is to automatically determine whether and when a word of interest is spoken by a talking face, with or without the audio. We propose a zero-shot method suitab…

cs.CV2019

Detecting Spoofing Attacks Using VGG and SincNet: BUT-Omilia Submission to ASVspoof 2019 Challenge

Hossein Zeinali, Themos Stafylakis, Georgia Athanasopoulou +4

In this paper, we present the system description of the joint efforts of Brno University of Technology (BUT) and Omilia -- Conversational Intelligence for the ASVSpoof2019 Spoofing…

cs.CV2019

Self-supervised speaker embeddings

Themos Stafylakis, Johan Rohdin, Oldrich Plchot +2

Contrary to i-vectors, speaker embeddings such as x-vectors are incapable of leveraging unlabelled utterances, due to the classification loss over training speakers. In this paper,…