2 papers
eess.AS2025
SV-Mixer: Replacing the Transformer Encoder with Lightweight MLPs for Self-Supervised Model Compression in Speaker Verification
Jungwoo Heo, Hyun-seo Shin, Chan-yeong Lim +4
Self-supervised learning (SSL) has pushed speaker verification accuracy close to state-of-the-art levels, but the Transformer backbones used in most SSL encoders hinder on-device a…
eess.AS2025
Token-based Attractors and Cross-attention in Spoof Diarization
Kyo-Won Koo, Chan-yeong Lim, Jee-weon Jung +2
Spoof diarization identifies ``what spoofed when" in a given speech by temporally locating spoofed regions and determining their manipulation techniques. As a first step toward thi…