activity
20212025
most citedThe VoxCeleb Speaker Recognition Challenge: A Retrospective

22 citations · 35 across the 14 of their papers we have counts for

collaborators
Showing eess.ASShow all

9 papers · 1 filter

eess.AS2025

SEED: Speaker Embedding Enhancement Diffusion Model

KiHyun Nam, Jungwoo Heo, Jee-weon Jung +4

A primary challenge when deploying speaker recognition systems in real-world applications is performance degradation caused by environmental mismatch. We propose a diffusion-based…

eess.AS2024

To what extent can ASV systems naturally defend against spoofing attacks?

Jee-weon Jung, Xin Wang, Nicholas Evans +6

The current automatic speaker verification (ASV) task involves making binary decisions on two types of trials: target and non-target. However, emerging advancements in speech gener…

eess.AS2023

Rethinking Session Variability: Leveraging Session Embeddings for Session Robustness in Speaker Verification

Hee-Soo Heo, KiHyun Nam, Bong-Jin Lee +4

In the field of speaker verification, session or channel variability poses a significant challenge. While many contemporary methods aim to disentangle session information from spea…

eess.AS2022

Metric Learning for User-defined Keyword Spotting

Jaemin Jung, Youkyum Kim, Jihwan Park +4

The goal of this work is to detect new spoken terms defined by users. While most previous works address Keyword Spotting (KWS) as a closed-set classification problem, this limits t…

eess.AS20223 cited

Pushing the limits of raw waveform speaker recognition

Jee-weon Jung, You Jin Kim, Hee-Soo Heo +3

In recent years, speaker recognition systems based on raw waveform inputs have received increasing attention. However, the performance of such systems are typically inferior to the…

eess.AS2021

Multi-scale speaker embedding-based graph attention networks for speaker diarisation

Youngki Kwon, Hee-Soo Heo, Jee-weon Jung +3

The objective of this work is effective speaker diarisation using multi-scale speaker embeddings. Typically, there is a trade-off between the ability to recognise short speaker seg…